3 papers
cs.CV2026
Enhancing Multimodal In-Context Learning via Inductive-Deductive Reasoning
Haoyu Wang, Haonan Wang, Yuyan Chen +5
In-context learning (ICL) allows large models to adapt to tasks using a few examples, yet its extension to vision-language models (VLMs) remains fragile. Our analysis reveals that…
cs.CL2026
Why Did Apple Fall: Evaluating Curiosity in Large Language Models
Haoyu Wang, Sihang Jiang, Yuyan Chen +4
Curiosity serves as a pivotal conduit for human beings to discover and learn new knowledge. Recent advancements of large language models (LLMs) in natural language processing have…
cs.AI2026
Thinking with Constructions: A Benchmark and Policy Optimization for Visual-Text Interleaved Geometric Reasoning
Haokun Zhao, Wanshi Xu, Haidong Yuan +3
Geometric reasoning inherently requires "thinking with constructions" -- the dynamic manipulation of visual aids to bridge the gap between problem conditions and solutions. However…