3 papers
cs.CV2026
ATLAS: A Large-Scale Evaluation Benchmark for Adversarial LiDAR Perception
Mellon M. Zhang, Siddhant Panse, Zimo Fan +3
Autonomous driving perception is typically evaluated on clean benchmark data, yet real-world deployment requires robustness to rare, structured, and potentially adversarial sensor…
cs.CV2025
MAPS: Preserving Vision-Language Representations via Module-Wise Proximity Scheduling for Better Vision-Language-Action Generalization
Chengyue Huang, Mellon M. Zhang, Robert Azarcon +2
Vision-Language-Action (VLA) models inherit strong priors from pretrained Vision-Language Models (VLMs), but naive fine-tuning often disrupts these representations and harms genera…
cs.CV2025
Towards Streaming LiDAR Object Detection with Point Clouds as Egocentric Sequences
Mellon M. Zhang, Glen Chou, Saibal Mukhopadhyay
Accurate and low-latency 3D object detection is essential for autonomous driving, where safety hinges on both rapid response and reliable perception. While rotating LiDAR sensors a…