FocusVTC: Efficient and High-Performance Visual Text Compression with Adaptive Resolution Paper • 2609.36651 • Published 4 days ago • 25
SoL-Refiner: Speed-of-Light One-Step Refinement for High-Resolution Video Paper • 2609.37969 • Published 4 days ago • 31
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 6 days ago • 508
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 6 days ago • 41
Groupwise Agentic Grading and Advantage Redistribution for Code Agent RL Paper • 2609.32577 • Published 7 days ago • 127
Rethinking Training-Inference Mismatch in LLM Reinforcement Learning: Where It Arises and How to Correct It Paper • 2609.32444 • Published 7 days ago • 30
Imprint Reader: From Weight-Update Readout to Behavioral Intervention Paper • 2609.35261 • Published 5 days ago • 16
Do Implicit Personalization and Explicit Styles Conflict? PsPLUG: A Lightweight Plug-in for Balancing Personalization and Style in Customized LLMs Paper • 2601.06362 • Published 13 days ago • 10
AV-GRPO: Modality-Anchored Decoupling Diffusion Reinforcement Learning for Joint Audio-Video Generation Paper • 2609.29816 • Published 9 days ago • 12
Qwen-Planner-Agent: A Closed-Loop AI-for-AI Framework for Real-World Mobile Planner Agents Paper • 2609.29892 • Published 9 days ago • 32
RewardVerse: Rubric-Guided Policy Optimization for Video Reward Modeling Paper • 2609.22947 • Published 14 days ago • 41
Don't Mask the Environment: Observation Supervision Changes How Agents Explore Under RL Paper • 2609.20715 • Published 16 days ago • 44