1 paper
Zichao Li, Xueru Wen, Jie Lou +5
Multimodal Reward Models (MM-RMs) are crucial for aligning Large Language Models (LLMs) with human preferences, particularly as LLMs increasingly interact with multimodal data. How…