1 paper
Guanghao Zhou, Panjia Qiu, Cen Chen +4
The application of reinforcement learning (RL) to enhance the reasoning capabilities of Multimodal Large Language Models (MLLMs) constitutes a rapidly advancing research area. Whil…