VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 12 days ago • 179
EnvHarness: Awakening Static Worlds for Agent Learning Paper • 2608.19880 • Published 18 days ago • 274
Demystifying Agent Skills: Why They Work-Until They Don't Paper • 2608.14036 • Published 24 days ago • 169
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published 26 days ago • 291
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 28 days ago • 342
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Paper • 2608.01755 • Published Aug 3 • 142
TARS: Timestep-Aware Data Scaling for 3D-Free Video Re-Shooting Paper • 2607.28261 • Published Jul 30 • 116
Are LLMs Ready for Scientific Discovery? A Capability-Oriented Benchmark for AI Scientists Paper • 2607.11079 • Published Jul 13 • 10
InstanceControl: Controllable Complex Image Generation without Instance Labeling Paper • 2606.31924 • Published Jun 30 • 15
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models Paper • 2606.11324 • Published Jun 9 • 173
Domino: Decoupling Causal Modeling from Autoregressive Drafting in Speculative Decoding Paper • 2605.29707 • Published May 28 • 152
Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information Paper • 2605.11609 • Published May 12 • 196
OpenSearch-VL: An Open Recipe for Frontier Multimodal Search Agents Paper • 2605.05185 • Published May 6 • 108
From Context to Skills: Can Language Models Learn from Context Skillfully? Paper • 2604.27660 • Published May 3 • 172
Phantom: Physics-Infused Video Generation via Joint Modeling of Visual and Latent Physical Dynamics Paper • 2604.08503 • Published Apr 9 • 7