1 citations · 1 across the 19 of their papers we have counts for
7 papers · 1 filter
MapRoute++: Surrogate-Guided Semantic Routing for Visual Concept Unlearning
Ashok Urlana, L. D. M. S. Sai Teja, Vivek Hruday Kavuri +1
We present our submission to Task 3 of the Gen 2.0 Challenge on visual concept unlearning. Building on MapRoute, we introduce task-specific training objectives, richer concept r…
GridVQA-X: A Framework for Evaluating Multimodal Explainability Methods
Sujay Belsare, Sudarshan Nikhil, Sushant Kumar +2
With the increasing development of Vision-Language Models, it becomes imperative that their predictions are readily explainable to relevant stakeholders. However, the field of expl…
Communicating about Space: Language-Mediated Spatial Integration Across Partial Views
Ankur Sikarwar, Debangan Mishra, Sudarshan Nikhil +2
Humans build shared spatial understanding by communicating partial, viewpoint-dependent observations. We ask whether Multimodal Large Language Models (MLLMs) can do the same, align…
LABELING COPILOT: A Deep Research Agent for Automated Data Curation in Computer Vision
Debargha Ganguly, Sumit Kumar, Ishwar Balappanawar +6
Curating high-quality, domain-specific datasets is a major bottleneck for deploying robust vision systems, requiring complex trade-offs between data quality, diversity, and cost wh…
Freeze and Reveal: Exposing Modality Bias in Vision-Language Models
Vivek Hruday Kavuri, Vysishtya Karanam, Venkata Jahnavi Venkamsetty +3
Vision Language Models achieve impressive multi-modal performance but often inherit gender biases from their training data. This bias might be coming from both the vision and text…
Random Representations Outperform Online Continually Learned Representations
Ameya Prabhu, Shiven Sinha, Ponnurangam Kumaraguru +3
Continual learning has primarily focused on the issue of catastrophic forgetting and the associated stability-plasticity tradeoffs. However, little attention has been paid to the e…