Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
MedObvious: Exposing the Medical Moravec's Paradox in VLMs via Clinical Triage
Ufaq Khan, Umair Nawaz, L D M S S Teja +5
Vision Language Models (VLMs) are increasingly used for tasks like medical report generation and visual question answering. However, fluent diagnostic text does not guarantee safe…
cs.CV2026
AURORA: Adaptive Unified Representation for Robust Ultrasound Analysis
Ufaq Khan, L. D. M. S. Sai Teja, Ayuba Shakiru +4
Ultrasound images vary widely across scanners, operators, and anatomical targets, which often causes models trained in one setting to generalize poorly to new hospitals and clinica…
cs.CV2025
AGIC: Attention-Guided Image Captioning to Improve Caption Relevance
L. D. M. S. Sai Teja, Ashok Urlana, Pruthwik Mishra
Despite significant progress in image captioning, generating accurate and descriptive captions remains a long-standing challenge. In this study, we propose Attention-Guided Image C…