Yufan Gong
yufan-gong
ยท
AI & ML interests
offline RL + reinforcement learning
Recent Activity
upvoted a paper about 18 hours ago
StudentSim: Training LLM-based Student Simulators upvoted a paper about 18 hours ago
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders liked a model 5 days ago
fb700/chatglm-fitness-RLHFOrganizations
None yet