2 papers
cs.CL2025
Sotopia-RL: Reward Design for Social Intelligence
Haofei Yu, Zhengyang Qi, Yining Zhao +6
Social intelligence has become a critical capability for large language models (LLMs), enabling them to engage effectively in real-world social tasks such as collaboration and nego…
cs.LG2025
MINT: Multimodal Instruction Tuning with Multimodal Interaction Grouping
Xiaojun Shan, Qi Cao, Xing Han +2
Recent advances in multimodal foundation models have achieved state-of-the-art performance across a range of tasks. These breakthroughs are largely driven by new pre-training parad…