Gradient-based Data Diversification Boosts Generalization in LLM Reasoning
Jaehun Jung
Jaehun
AI & ML interests
None yet
Recent Activity
authored a paper about 1 month ago
LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL authored a paper about 1 month ago
How to Instruct Your Robot: Dense Language Annotations Power Robot Policy Learning authored a paper about 1 month ago
ProCUA-SFT Technical Report