2 papers
cs.CV2026
Truth in the Few: High-Value Data Selection for Efficient Multi-Modal Reasoning
Shenshen Li, Xing Xu, Kaiyuan Deng +3
While multi-modal large language models (MLLMs) have made significant progress in complex reasoning tasks via reinforcement learning, it is commonly believed that extensive trainin…
cs.AI2025
What Makes Reasoning Invalid: Echo Reflection Mitigation for Large Language Models
Chen He, Xun Jiang, Lei Wang +5
Large Language Models (LLMs) have demonstrated remarkable performance across a wide range of reasoning tasks. Recent methods have further improved LLM performance in complex mathem…