yicheng qiu
MaXWe1l1
ยท
AI & ML interests
AI for science
RL theory
Recent Activity
liked a dataset 19 days ago
SciDataOcean/ReasonEM upvoted a paper 3 months ago
Smaller Models are Natural Explorers for Policy-Level Diversity in GRPO