13 citations · 40 across the 10 of their papers we have counts for
1 paper · 1 filter
Zichao Li, Xueru Wen, Jie Lou +5
Multimodal Reward Models (MM-RMs) are crucial for aligning Large Language Models (LLMs) with human preferences, particularly as LLMs increasingly interact with multimodal data. How…