1 paper
Bingrui Sima, Linhua Cong, Wenxuan Wang +1
The emergence of Multimodal Large Language Models (MLRMs) has enabled sophisticated visual reasoning capabilities by integrating reinforcement learning and Chain-of-Thought (CoT) s…