5 papers
SCLARO: A Dataset for Grounded Scenario-Level Scene Understanding and ScenarioCLIP for Benchmarking
Advik Sinha, Saurabh Atreya, Aashutosh A +2
In the paradigm of computer vision-based precise real-world scene understanding, joint reasoning in terms of contextual understanding about the objects present in a scene, their in…
Continual-learning for Modelling Low-Resource Languages from Large Language Models
Santosh Srinath K, Mudit Somani, Varun Reddy Padala +2
Modelling a language model for a multi-lingual scenario includes several potential challenges, among which catastrophic forgetting is the major challenge. For example, small langua…
DensePercept-NCSSD: Vision Mamba towards Real-time Dense Visual Perception with Non-Causal State Space Duality
Tushar Anand, Advik Sinha, Abhijit Das
In this work, we propose an accurate and real-time optical flow and disparity estimation model by fusing pairwise input images in the proposed non-causal selective state space for…
Towards Obstacle-Avoiding Control of Planar Snake Robots Exploring Neuro-Evolution of Augmenting Topologies
Advik Sinha, Akshay Arjun, Abhijit Das +1
This work aims to develop a resource-efficient solution for obstacle-avoiding tracking control of a planar snake robot in a densely cluttered environment with obstacles. Particular…
A Survey on Improving Human Robot Collaboration through Vision-and-Language Navigation
Nivedan Yakolli, Avinash Gautam, Abhijit Das +2
Vision-and-Language Navigation (VLN) is a multi-modal, cooperative task requiring agents to interpret human instructions, navigate 3D environments, and communicate effectively unde…