Omni-Reward: Towards Generalist Omni-Modal Reward Modeling with Free-Form Preferences
Zhuoran Jin
jinzhuoran
AI & ML interests
NLP
Recent Activity
upvoted a paper about 1 hour ago
SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMsOrganizations
None yet