Why Do Video Diffusion Models Violate Physics? Unveiling the Flaws in Attention Mechanisms Paper • 2609.23658 • Published 5 days ago • 25
Reading Images Like Texts: Sequential Image Understanding in Vision-Language Models Paper • 2509.19191 • Published Feb 9 • 1
GoLongRL: Capability-Oriented Long Context Reinforcement Learning with Multitask Alignment Paper • 2605.19577 • Published May 19 • 59