1 paper
Wen Luo, Xiaohan Yi, Xiaotao Huang +1
Multimodal large reasoning models often rely on long Chain-of-Thought (CoT) traces in which a substantial fraction of tokens, such as repeated visual descriptions, self-reflection,…