Woah! Dockerless: Environment-Free Program Verifier for Coding Agents Paper • 2606.28436 • Published 25 days ago • 111 Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published 19 days ago • 120
Dockerless: Environment-Free Program Verifier for Coding Agents Paper • 2606.28436 • Published 25 days ago • 111
Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published 19 days ago • 120
Pre-Training Tokens HRM-Text: Efficient Pretraining Beyond Scaling Paper • 2605.20613 • Published May 20 • 323 LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models Paper • 2605.11011 • Published May 10 • 10 Less is More: Recursive Reasoning with Tiny Networks Paper • 2510.04871 • Published Oct 6, 2025 • 518 Characterizing, Evaluating, and Optimizing Complex Reasoning Paper • 2602.08498 • Published Jun 3 • 1
LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models Paper • 2605.11011 • Published May 10 • 10
Less is More: Recursive Reasoning with Tiny Networks Paper • 2510.04871 • Published Oct 6, 2025 • 518
Characterizing, Evaluating, and Optimizing Complex Reasoning Paper • 2602.08498 • Published Jun 3 • 1
Demixing Models & Datasets Moisesdb: A dataset for source separation beyond 4-stems Paper • 2307.15913 • Published Jul 29, 2023 • 1 Music Source Separation with Band-Split RoPE Transformer Paper • 2309.02612 • Published Sep 5, 2023 • 2 Hybrid Transformers for Music Source Separation Paper • 2211.08553 • Published Nov 15, 2022 • 1 nvidia/RE-USE Audio-to-Audio • 9.61M • Updated May 14 • 13.8k • 86
Moisesdb: A dataset for source separation beyond 4-stems Paper • 2307.15913 • Published Jul 29, 2023 • 1
Music Source Separation with Band-Split RoPE Transformer Paper • 2309.02612 • Published Sep 5, 2023 • 2
Agent Loops, Character, Work Ethics & Behavior Close the Loop: Synthesizing Infinite Tool-Use Data via Multi-Agent Role-Playing Paper • 2512.23611 • Published Dec 29, 2025 • 7 Context as a Tool: Context Management for Long-Horizon SWE-Agents Paper • 2512.22087 • Published Dec 26, 2025 • 4 AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications Paper • 2508.16279 • Published Aug 22, 2025 • 67 Very Large-Scale Multi-Agent Simulation in AgentScope Paper • 2407.17789 • Published Jul 25, 2024 • 44
Close the Loop: Synthesizing Infinite Tool-Use Data via Multi-Agent Role-Playing Paper • 2512.23611 • Published Dec 29, 2025 • 7
Context as a Tool: Context Management for Long-Horizon SWE-Agents Paper • 2512.22087 • Published Dec 26, 2025 • 4
AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications Paper • 2508.16279 • Published Aug 22, 2025 • 67
Very Large-Scale Multi-Agent Simulation in AgentScope Paper • 2407.17789 • Published Jul 25, 2024 • 44
Sample Upscaling & Denoising. nvidia/RE-USE Audio-to-Audio • 9.61M • Updated May 14 • 13.8k • 86 weya-ai/hush Audio-to-Audio • Updated Mar 31 • 35 • 37 drbaph/AudioSR Audio-to-Audio • Updated Jan 27 • 26 oulianov/slp-rl-aero Audio-to-Audio • Updated Mar 20 • 1
Evaluation Methods & Metrics RubricBench: Aligning Model-Generated Rubrics with Human Standards Paper • 2603.01562 • Published Mar 2 • 64 T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasoning Paper • 2603.03790 • Published Mar 4 • 122 SWE-rebench: An Automated Pipeline for Task Collection and Decontaminated Evaluation of Software Engineering Agents Paper • 2505.20411 • Published May 26, 2025 • 97 SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale Paper • 2602.23866 • Published Feb 27 • 92
RubricBench: Aligning Model-Generated Rubrics with Human Standards Paper • 2603.01562 • Published Mar 2 • 64
T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasoning Paper • 2603.03790 • Published Mar 4 • 122
SWE-rebench: An Automated Pipeline for Task Collection and Decontaminated Evaluation of Software Engineering Agents Paper • 2505.20411 • Published May 26, 2025 • 97
SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale Paper • 2602.23866 • Published Feb 27 • 92
Py serafeimdossas/ai-code-detection Viewer • Updated Aug 12, 2025 • 11.8k • 36 argilla-warehouse/python-lib-tools-v0.1 Viewer • Updated Oct 1, 2024 • 50.7k • 90 • 6 Dream-org/Dream-Coder-RL-17k Viewer • Updated Aug 6, 2025 • 17k • 107 • 5
Woah! Dockerless: Environment-Free Program Verifier for Coding Agents Paper • 2606.28436 • Published 25 days ago • 111 Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published 19 days ago • 120
Dockerless: Environment-Free Program Verifier for Coding Agents Paper • 2606.28436 • Published 25 days ago • 111
Program-as-Weights: A Programming Paradigm for Fuzzy Functions Paper • 2607.02512 • Published 19 days ago • 120
Pre-Training Tokens HRM-Text: Efficient Pretraining Beyond Scaling Paper • 2605.20613 • Published May 20 • 323 LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models Paper • 2605.11011 • Published May 10 • 10 Less is More: Recursive Reasoning with Tiny Networks Paper • 2510.04871 • Published Oct 6, 2025 • 518 Characterizing, Evaluating, and Optimizing Complex Reasoning Paper • 2602.08498 • Published Jun 3 • 1
LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models Paper • 2605.11011 • Published May 10 • 10
Less is More: Recursive Reasoning with Tiny Networks Paper • 2510.04871 • Published Oct 6, 2025 • 518
Characterizing, Evaluating, and Optimizing Complex Reasoning Paper • 2602.08498 • Published Jun 3 • 1
Sample Upscaling & Denoising. nvidia/RE-USE Audio-to-Audio • 9.61M • Updated May 14 • 13.8k • 86 weya-ai/hush Audio-to-Audio • Updated Mar 31 • 35 • 37 drbaph/AudioSR Audio-to-Audio • Updated Jan 27 • 26 oulianov/slp-rl-aero Audio-to-Audio • Updated Mar 20 • 1
Demixing Models & Datasets Moisesdb: A dataset for source separation beyond 4-stems Paper • 2307.15913 • Published Jul 29, 2023 • 1 Music Source Separation with Band-Split RoPE Transformer Paper • 2309.02612 • Published Sep 5, 2023 • 2 Hybrid Transformers for Music Source Separation Paper • 2211.08553 • Published Nov 15, 2022 • 1 nvidia/RE-USE Audio-to-Audio • 9.61M • Updated May 14 • 13.8k • 86
Moisesdb: A dataset for source separation beyond 4-stems Paper • 2307.15913 • Published Jul 29, 2023 • 1
Music Source Separation with Band-Split RoPE Transformer Paper • 2309.02612 • Published Sep 5, 2023 • 2
Evaluation Methods & Metrics RubricBench: Aligning Model-Generated Rubrics with Human Standards Paper • 2603.01562 • Published Mar 2 • 64 T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasoning Paper • 2603.03790 • Published Mar 4 • 122 SWE-rebench: An Automated Pipeline for Task Collection and Decontaminated Evaluation of Software Engineering Agents Paper • 2505.20411 • Published May 26, 2025 • 97 SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale Paper • 2602.23866 • Published Feb 27 • 92
RubricBench: Aligning Model-Generated Rubrics with Human Standards Paper • 2603.01562 • Published Mar 2 • 64
T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasoning Paper • 2603.03790 • Published Mar 4 • 122
SWE-rebench: An Automated Pipeline for Task Collection and Decontaminated Evaluation of Software Engineering Agents Paper • 2505.20411 • Published May 26, 2025 • 97
SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale Paper • 2602.23866 • Published Feb 27 • 92
Agent Loops, Character, Work Ethics & Behavior Close the Loop: Synthesizing Infinite Tool-Use Data via Multi-Agent Role-Playing Paper • 2512.23611 • Published Dec 29, 2025 • 7 Context as a Tool: Context Management for Long-Horizon SWE-Agents Paper • 2512.22087 • Published Dec 26, 2025 • 4 AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications Paper • 2508.16279 • Published Aug 22, 2025 • 67 Very Large-Scale Multi-Agent Simulation in AgentScope Paper • 2407.17789 • Published Jul 25, 2024 • 44
Close the Loop: Synthesizing Infinite Tool-Use Data via Multi-Agent Role-Playing Paper • 2512.23611 • Published Dec 29, 2025 • 7
Context as a Tool: Context Management for Long-Horizon SWE-Agents Paper • 2512.22087 • Published Dec 26, 2025 • 4
AgentScope 1.0: A Developer-Centric Framework for Building Agentic Applications Paper • 2508.16279 • Published Aug 22, 2025 • 67
Very Large-Scale Multi-Agent Simulation in AgentScope Paper • 2407.17789 • Published Jul 25, 2024 • 44
Py serafeimdossas/ai-code-detection Viewer • Updated Aug 12, 2025 • 11.8k • 36 argilla-warehouse/python-lib-tools-v0.1 Viewer • Updated Oct 1, 2024 • 50.7k • 90 • 6 Dream-org/Dream-Coder-RL-17k Viewer • Updated Aug 6, 2025 • 17k • 107 • 5