2 papers
cs.AI2026
Do VLMs Reason Like Engineers? A Benchmark and a Stage-wise Evaluation
Syed Wasiq, Syed Mohamad Tawseeq, Yashwant Pravinrao Bangde +1
Vision-Language Models (VLMs) demonstrate strong performance on general multimodal reasoning benchmarks, yet their ability to perform engineering reasoning remains largely unexplor…
cs.CV2026
Instruction-Evidence Contrastive Dual-Stream Decoding for Grounded Vision-Language Reasoning
Yashwant Pravinrao Bangde, Debaditya Roy
Vision-Language Models (VLMs) exhibit strong performance in instruction following and open-ended vision-language reasoning, yet they frequently generate fluent outputs that are wea…