TRIAGE: Direction-Aware Mismatch Stabilization of Native NVFP4 Reinforcement Learning Paper • 2610.07043 • Published 4 days ago • 29
Routing Drift Alone Does Not Diagnose Failure in Merged MoE LLMs Paper • 2609.32821 • Published 13 days ago • 5
Access Sets Matter: Budgeting Expert Reads for Scalable Weight-Space Model Merging Paper • 2605.29489 • Published May 28 • 4
Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation Paper • 2605.26844 • Published May 26 • 24
Geometry Conflict: Explaining and Controlling Forgetting in LLM Continual Post-Training Paper • 2605.09608 • Published May 10 • 52
InfiAlign: A Scalable and Sample-Efficient Framework for Aligning LLMs to Enhance Reasoning Capabilities Paper • 2508.05496 • Published Aug 7, 2025 • 9