3 papers
cs.CV2026
Adapting Vision Foundation Models to Acoustics for Pose-Free 3D Sonar Reconstruction
Kevin Zhang, Jingxi Chen, Mohamad Qadri +5
Vision foundation models trained on Internet-scale RGB datasets enable remarkable capabilities across a range of tasks, from text-to-video generation to few-shot 3D scene reconstru…
cs.CV2025
Underwater Monocular Metric Depth Estimation: Real-World Benchmarks and Synthetic Fine-Tuning with Vision Foundation Models
Zijie Cai, Christopher Metzler
Monocular depth estimation has recently progressed beyond ordinal depth to provide metric depth predictions. However, its reliability in underwater environments remains limited due…
eess.SP2025
Acoustic Neural 3D Reconstruction Under Pose Drift
Tianxiang Lin, Mohamad Qadri, Kevin Zhang +3
We consider the problem of optimizing neural implicit surfaces for 3D reconstruction using acoustic images collected with drifting sensor poses. The accuracy of current state-of-th…