arxiv:2601.12730
Chen Wang
wc597358816
AI & ML interests
None yet
Recent Activity
upvoted a paper 27 days ago
Distilled Reinforcement Learning for LLM Post-training submitted a paper 27 days ago
Distilled Reinforcement Learning for LLM Post-training updated a model about 1 month ago
wc597358816/Qwen3-8B-GRPOOrganizations
None yet