الباحثون

Alan C. Bovik

المنشورات 2

نسخة أولية وصول مفتوح

MEND: RL For Flow Models via Proximal Velocity Matching

Shreshth Saini, Neil Birkbeck, Yilin Wang وآخرون · 2026

Reward post-training of flow models either reweights the model's own samples under a KL penalty or a frozen reference, often for thousands of updates, or backpropagates the reward and moves every sample without checking that the move is worth its size. We introduce MEND, a reinforcement learning method built on proxima …

نسخة أولية وصول مفتوح

JLD: Perceptual Distance Through A Jacobian Lens

Image compression, restoration, and generation all require a way to measure how different two images look to a person. Pixel error ignores how people see, while the most accurate perceptual distances are typically fitted to human judgments, tying them to a fixed data and resolution. For example, when image resolution i …

المؤلفون المشاركون