COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization Paper • 2609.11682 • Published 14 days ago • 45
Rethinking On-Policy Distillation of Large Language Models II: One Training Example Paper • 2609.04172 • Published 21 days ago • 101
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation Paper • 2608.04419 • Published Aug 5 • 30
T-POP: Test-Time Personalization with Online Preference Feedback Paper • 2509.24696 • Published Sep 29, 2025 • 1
T-POP: Test-Time Personalization with Online Preference Feedback Paper • 2509.24696 • Published Sep 29, 2025 • 1
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation Paper • 2608.04419 • Published Aug 5 • 30
InCoder-32B: Code Foundation Model for Industrial Scenarios Paper • 2603.16790 • Published Mar 17 • 189
From Code Foundation Models to Agents and Applications: A Practical Guide to Code Intelligence Paper • 2511.18538 • Published Nov 23, 2025 • 307