Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Self-Evolving Visual Questioner
Yijun Liang, Hengguang Zhou, Ming Li +3
Vision-language models (VLMs) are typically trained as passive answerers, while their ability to actively ask diverse, non-trivial, visual-centric and grounded questions remains un…
cs.CV2026
V-REX: Benchmarking Exploratory Visual Reasoning via Chain-of-Questions
Chenrui Fan, Yijun Liang, Shweta Bhardwaj +3
While many vision-language models (VLMs) are developed to answer well-defined, straightforward questions with highly specified targets, as in most benchmarks, they often struggle i…