MIMESIS: Learning User Simulators as Training Environments for Interactive Agents Paper • 2610.09484 • Published 3 days ago • 20
DiVeR: Decision-Critical Verifier Learning for VLA Test-Time Scaling Paper • 2610.04933 • Published 6 days ago • 14
Science or Slop?: Benchmarking and Mitigating Scientific Slop in AI-Generated Papers Paper • 2610.00531 • Published 10 days ago • 60
Mistral Large 3 Collection A state-of-the-art, open-weight, general-purpose multimodal model with a granular Mixture-of-Experts architecture. • 4 items • Updated about 7 hours ago • 104
nvidia/Nemotron-Research-Reasoning-Qwen-1.5B Text Generation • 2B • Updated Nov 21, 2025 • 1.14k • • 244
Hide-and-Seek in Trajectories: Discovering Failure Signals for VLA Runtime Monitoring Paper • 2605.30834 • Published May 29 • 9
Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents Paper • 2606.26080 • Published Jun 24 • 13
WTF GENIUS PAPERS Collection Papers that made me appreciate my major and my life a little more. obs=Observation, innov=Innovation. Most papers are abt improving tiny models. • 453 items • Updated 1 day ago • 99