Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
The Ebb and Flow of Multimodal Focus: Scheduling Visual Relay Windows for Grounded VLM Reasoning
Wencheng Ye, Yi Bin, Yujuan Ding +7
Vision-language models increasingly succeed on multimodal reasoning benchmarks, yet their visual evidence often becomes unstable once it enters the language stack, weakening eviden…
cs.AI2025
What Makes Reasoning Invalid: Echo Reflection Mitigation for Large Language Models
Chen He, Xun Jiang, Lei Wang +5
Large Language Models (LLMs) have demonstrated remarkable performance across a wide range of reasoning tasks. Recent methods have further improved LLM performance in complex mathem…