Omni-Reward: Towards Generalist Omni-Modal Reward Modeling with Free-Form Preferences
Zhuoran Jin
jinzhuoran
AI & ML interests
NLP
Recent Activity
submitted a paper about 9 hours ago
The Physics of Multi-Turn Long-Horizon Planning: From Pre-training to Post-training via Single- and Multi-Teacher On-Policy Agentic Distillation liked a dataset 13 days ago
ByteDance-Seed/EdgeBenchOrganizations
None yet