4 papers
MotionSync: Non-Causal Refinement of Causal Tracker for Label-Efficient 3D Perception
Rahul Ahuja, Bala Murali Manoghar Sai Sudhakar, Shashwata Gupta +3
Three-dimensional box-and-track annotation is the cost bottleneck in autonomous-driving data engines, and the offline systems built to relieve it replace the online perception stac…
RedLight-VLA: Models for traffic-rule grounding and behavioral emphasis in driving policies
Bala Murali Manoghar Sai Sudhakar, Sourab Bapu Sridhar, Sandipan Das +5
Behavior-cloned Vision-Language-Action (VLA) driving policies struggle with rare rule-governed maneuvers at signalized intersections. Braking and launching examples contribute litt…
FishRoPE: Projective Rotary Position Embeddings for Omnidirectional Visual Perception
Rahul Ahuja, Mudit Jain, Bala Murali Manoghar Sai Sudhakar +4
Vision foundation models (VFMs) and Bird's Eye View (BEV) representation have advanced visual perception substantially, yet their internal spatial representations assume the rectil…
MambaFusion: Adaptive State-Space Fusion for Multimodal 3D Object Detection
Venkatraman Narayanan, Bala Sai, Rahul Ahuja +3
Reliable 3D object detection is fundamental to autonomous driving, and multimodal fusion algorithms using cameras and LiDAR remain a persistent challenge. Cameras provide dense vis…