arxiv
PublishedSeptember 30, 2026 at 4:00 AM
CoRe: Co-Evolving Reward Models for Mitigating Latent Reward Hacking in Video Diffusion Models
Publisher summary· verbatim
arXiv:2609.36245v1 Announce Type: new Abstract: Latent reward models (LRMs) enable efficient alignment of video diffusion models by scoring intermediate states directly in latent space. However, we find that optimizing against a fixed latent reward rapidly leads to latent reward hacking: the predict
Stay posted· Newsletter
A 5-min weekly brief — top movers, price watch, story of the week.
Discussion
No replies yet. Be first.
Related coverage
More from ARXIV
arxivRight Words, Wrong Moment: A Clinician-Grounded Analysis of Distress in 19,930 Conversations between Young People and ChatGPT59marxivSAGE: A Statistical Acceptance Gate for Self-Evolving Agents59marxivPowerZooJax: A JAX-based Power System Benchmark for Reinforcement Learning59marxivGeoWind2Plan: Mission-Time 3D Urban Wind Prediction for Energy-Efficient UAV Planning59mThe Bubble Brief
WEEKLYRead AI insights every Tuesday — top movers, new releases, story of the week.
Originally published on arxiv ↗