EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents Paper • 2609.17632 • Published 13 days ago • 45
StableVQ: Practical Guidelines for Stable Vector-Quantized Tokenizer Training Paper • 2609.26774 • Published 6 days ago • 55
Emergent Collusion in Long-Horizon LLM Agent Interaction Paper • 2609.24967 • Published 7 days ago • 19
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 6 days ago • 159
LatentPort: Beyond KV Cache - Cross-Model Transfer of Recurrent Memory in Hybrid Language Models: A 4B-to-9B Hybrid-State Handoff Without Target Prefix Replay Paper • 2609.25053 • Published 21 days ago • 17
interstellarninja/hermes_interleaved_reasoning_tool_use Viewer • Updated Jun 30, 2025 • 1.71k • 102 • 9
qualifire-oss/mcp-tool-use-quality-ranger-0.6b-GGUF Text Generation • 0.6B • Updated Sep 15, 2025 • 12
Harness-Zero: Harness Distillation via Agent-as-Harness Paper • 2609.24974 • Published 7 days ago • 37
EDGEGEN: Improving Tool-Calling Agents Beyond Happy Paths with Synthetic Edge Case Generation Paper • 2609.24115 • Published 7 days ago • 5
One to More, More to One: Category-Aware Iterative Expert Training for Software Engineering Agents Paper • 2609.23377 • Published 8 days ago • 50