2 papers
cs.CV2025
FastViDAR: Real-Time Omnidirectional Depth Estimation via Alternative Hierarchical Attention
Hangtian Zhao, Xiang Chen, Yizhe Li +3
In this paper we propose FastViDAR, a novel framework that takes four fisheye camera inputs and produces a full depth map along with per-camera depth, fusion depth, and…
cs.CV2025
MARS2 2025 Challenge on Multimodal Reasoning: Datasets, Methods, Results, Discussion, and Outlook
Peng Xu, Shengwu Xiong, Jiajun Zhang +125
This paper reviews the MARS2 2025 Challenge on Multimodal Reasoning. We aim to bring together different approaches in multimodal machine learning and LLMs via a large benchmark. We…