3 papers
cs.CV2026
Structure-Token Evidence-Anchored Reasoning for Scientific Chart Understanding
Alberlucia Rafael Soarez, Camila Ferreira, Daniel Kim +2
Scientific charts encode quantities in axes, legends, and geometric marks, yet large vision-language models still treat them as natural photographs. Visual in-context examples do n…
eess.IV2026
SpaCE: Rethinking Spatial Capacity and Generalization in Multi-Frame Multimodal Large Language Models
Mariana Costa, Camila Ferreira, Alberlucia Rafael Soarez +1
Multi-modal large language models (MLLMs) have achieved remarkable empirical progress in spatial understanding through large-scale training on spatial visual question answering dat…
cs.CL2026
Enhancing Self-Correction in Large Language Models through Multi-Perspective Reflection
Mariana Costa, Alberlucia Rafael Soarez, Daniel Kim +1
While Chain-of-Thought (CoT) prompting advances LLM reasoning, challenges persist in consistency, accuracy, and self-correction, especially for complex or ethically sensitive tasks…