mim-chess-vlas/train_800_sparse__mask__overlay_a75__sim__all_cameras__live__pi05__seed_0 Robotics • 4B • Updated 5 days ago • 20 • 2
Reference-Based Bias Detection in LLMs via Relative Representations of Hidden States Paper • 2609.10060 • Published 11 days ago • 4
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 12 days ago • 171
Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models Paper • 2608.25518 • Published 25 days ago • 59
VGI-Bench: Probing Visual Intelligence in Video Generation Models Paper • 2608.19583 • Published 25 days ago • 88
WarpSAC: Towards the Pinnacle of Scalable Off-policy RL by Rethinking Exploration and Exploitation Paper • 2608.24479 • Published 26 days ago • 50
Learn What's Left, Not What's Mastered: Saturation Aware Advantage Reweighting for Multi-Reward Policy Optimization Paper • 2608.16072 • Published Aug 17 • 50
HarnessEval-W: Agentifying the Evaluation of Visual Worlds Paper • 2608.16859 • Published Aug 17 • 121
Spark-to-Paper: End-to-End Research Paper Generation as a Composable Skill Paper • 2608.11924 • Published Aug 12 • 108
Running Featured 789 Agent Memory Leaderboard 🧠 789 Unified memory evaluation · Results expected August 12.