Morphometric Imitation: From Morphology and Contact Aware Hand Retargeting to Sim-to-Real Visuomotor Policy Paper • 2609.28660 • Published 11 days ago • 15
EVO-WAM: Evolving World Action Models through Video-Action Verification Paper • 2609.38057 • Published 5 days ago • 36
FocusVTC: Efficient and High-Performance Visual Text Compression with Adaptive Resolution Paper • 2609.36651 • Published 5 days ago • 28
SoL-Refiner: Speed-of-Light One-Step Refinement for High-Resolution Video Paper • 2609.37969 • Published 5 days ago • 35
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 7 days ago • 518
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 7 days ago • 41
Groupwise Agentic Grading and Advantage Redistribution for Code Agent RL Paper • 2609.32577 • Published 8 days ago • 130
Rethinking Training-Inference Mismatch in LLM Reinforcement Learning: Where It Arises and How to Correct It Paper • 2609.32444 • Published 8 days ago • 30
Imprint Reader: From Weight-Update Readout to Behavioral Intervention Paper • 2609.35261 • Published 6 days ago • 16
Do Implicit Personalization and Explicit Styles Conflict? PsPLUG: A Lightweight Plug-in for Balancing Personalization and Style in Customized LLMs Paper • 2601.06362 • Published 14 days ago • 10
AV-GRPO: Modality-Anchored Decoupling Diffusion Reinforcement Learning for Joint Audio-Video Generation Paper • 2609.29816 • Published 10 days ago • 12
Qwen-Planner-Agent: A Closed-Loop AI-for-AI Framework for Real-World Mobile Planner Agents Paper • 2609.29892 • Published 10 days ago • 32
RewardVerse: Rubric-Guided Policy Optimization for Video Reward Modeling Paper • 2609.22947 • Published 15 days ago • 43