Yufan Gong
yufan-gong
ยท
AI & ML interests
offline RL + reinforcement learning
Recent Activity
upvoted a paper about 4 hours ago
StudentSim: Training LLM-based Student Simulators upvoted a paper about 4 hours ago
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders liked a model 4 days ago
fb700/chatglm-fitness-RLHFOrganizations
None yet