ScienceIDE: Turning World's Scientific Codebase into Agent Learnable Environments Paper • 2609.19134 • Published 2 days ago • 70
Beyond Solver Verdicts: Generative Reward Models for Autoformalization Paper • 2609.11085 • Published 8 days ago • 34
NeoHorse-1: Towards Recursive Self-Improvement via Agentic Post-Training with Routing Harness Paper • 2609.08183 • Published 10 days ago • 421