Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 22 days ago • 375
ActReview: Rebuttal-Guided Training Data and Rubric Rewards for Actionable Peer Review Generation Paper • 2609.09076 • Published 21 days ago • 24
Scaling Automatic Research Agents via World Models Paper • 2608.12564 • Published about 1 month ago • 482
CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements Paper • 2609.07498 • Published 22 days ago • 34
Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space Paper • 2608.29188 • Published about 1 month ago • 11
orcarouter/Qwen3.8-Flash-Next-Uncensored-GGUF Image-Text-to-Text • 177B • Updated 18 days ago • 199k • 424
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published Aug 27 • 155