arxiv:2610.01509
🔄 In a Training Loop
Changdae Oh
changdae
·
AI & ML interests
Generalization; Distribution Shift; Uncertainty Quantification; Reward Modeling; Post-training
Recent Activity
upvoted a paper about 22 hours ago
DiVeR: Decision-Critical Verifier Learning for VLA Test-Time Scaling authored a paper 6 days ago
Sharpening Tax in Post-Training