5 papers
SCLARO: A Dataset for Grounded Scenario-Level Scene Understanding and ScenarioCLIP for Benchmarking
Advik Sinha, Saurabh Atreya, Aashutosh A +2
In the paradigm of computer vision-based precise real-world scene understanding, joint reasoning in terms of contextual understanding about the objects present in a scene, their in…
How Many Counterfactuals Does It Take? Probing VLM Hallucinations Through Circuits and Causal Effects
Abhivansh Gupta, Simardeep Singh, Advika Sinha +2
Visual Language Models (VLMs) are known to produce hallucinated predictions that are not grounded in visual evidence, yet existing approaches lack a principled understanding of how…
DensePercept-NCSSD: Vision Mamba towards Real-time Dense Visual Perception with Non-Causal State Space Duality
Tushar Anand, Advik Sinha, Abhijit Das
In this work, we propose an accurate and real-time optical flow and disparity estimation model by fusing pairwise input images in the proposed non-causal selective state space for…
Towards Obstacle-Avoiding Control of Planar Snake Robots Exploring Neuro-Evolution of Augmenting Topologies
Advik Sinha, Akshay Arjun, Abhijit Das +1
This work aims to develop a resource-efficient solution for obstacle-avoiding tracking control of a planar snake robot in a densely cluttered environment with obstacles. Particular…
Impact of Language Guidance: A Reproducibility Study
Cherish Puniani, Advika Sinha, Shree Singhi +1
Modern deep-learning architectures need large amounts of data to produce state-of-the-art results. Annotating such huge datasets is time-consuming, expensive, and prone to human er…