Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 9 days ago • 91
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published 9 days ago • 177
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 17 days ago • 179
Apodex 1.1: Scaling Agentic Intelligence for Complex Work Paper • 2608.23283 • Published 19 days ago • 206
Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discoverative Artificial Intelligence Paper • 2608.11341 • Published Aug 11 • 67
AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses Paper • 2608.12307 • Published about 1 month ago • 115