CoRe: Co-Evolving Reward Models for Mitigating Latent Reward Hacking in Video Diffusion Models
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2609.36245v1 Announce Type: new Abstract: Latent reward models (LRMs) enable efficient alignment of video diffusion models by scoring intermediate states directly in latent space. However, we find that optimizing against a fixed latent reward rapidly leads to latent reward hacking: the predicted reward stays high while perceptual and motion quality deteriorate. Our analysis identifies distributional escape as the central cause: within a few hundred updates, the generator moves beyond the reward model's training support, where its scores no longer reflect video quality. Based on this insight, we introduce CoRe, a co-evolving reward framework that treats latent-space alignment as a dynam
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Bild
Related AI news
- China’s DeepSeek open-sources tools to help Huawei chips supplant Nvidia in AISCMP Tech · September 30, 2026
- Column: Memory and logic collide as AI shifts the balanceDIGITIMES · September 30, 2026
- Sources: US-based PaleBlueDot AI is seeking $600M in private credit to buy chips for its South Korea site, to be used by Chinese social media app Xiaohongshu (Megawati Wijaya/Bloomberg)Techmeme · September 30, 2026
- DeepSeek says it has partnered with Huawei to develop programming tools for Huawei's Ascend chips, including TileLang, an open-source CUDA alternative (Reuters)Techmeme · September 30, 2026
- Chinese firms trail global peers on profits, but AI power boom offers bright spot: NatixisSCMP Tech · September 30, 2026
- AI broke the job application. What replaces it?CNBC Technology · September 30, 2026