RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 19 days ago • 225
Mid-Harness: Scaling Actions Between Model and Harness for Terminal Agents Paper • 2609.39982 • Published 10 days ago • 119
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published Aug 4 • 56
Beyond Final Scores: A Systematic Evaluation of Agents for Long-Horizon AI Research and Development Paper • 2608.13417 • Published Aug 13 • 59
Securing the AI Agent: A Unified Framework for Multi-Layer Agent Red Teaming Paper • 2606.31227 • Published Jun 30 • 14
Multi-Turn Agentic Scientific Literature Search via Workflow Induction Paper • 2607.00597 • Published Jul 1 • 28
ResearchStudio-Idea: An Evidence-Grounded Research-Ideation Skill Suite from ML Conference Outcomes Paper • 2607.04439 • Published Jul 5 • 64
ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog Paper • 2607.04438 • Published Jul 5 • 62