activity
20192026
most citedStereo R-CNN based 3D Object Detection for Autonomous Driving

37 citations · 70 across the 11 of their papers we have counts for

collaborators
Showing cs.CVShow all

9 papers · 1 filter

cs.CV2025

Learning better representations for crowded pedestrians in offboard LiDAR-camera 3D tracking-by-detection

Shichao Li, Peiliang Li, Qing Lian +2

Perceiving pedestrians in highly crowded urban environments is a difficult long-tail problem for learning-based autonomous perception. Speeding up 3D ground truth generation for su…

cs.CV2024

Adaptive Fusion of Single-View and Multi-View Depth for Autonomous Driving

JunDa Cheng, Wei Yin, Kaixuan Wang +3

Multi-view depth estimation has achieved impressive performance over various benchmarks. However, almost all current multi-view systems rely on given ideal camera poses, which are…

cs.CV20242 cited

GIM: Learning Generalizable Image Matcher From Internet Videos

Xuelun Shen, Zhipeng Cai, Wei Yin +5

Image matching is a fundamental computer vision problem. While learning-based methods achieve state-of-the-art performance on existing benchmarks, they generalize poorly to in-the-…

cs.CV2023

UC-NeRF: Neural Radiance Field for Under-Calibrated Multi-view Cameras in Autonomous Driving

Kai Cheng, Xiaoxiao Long, Wei Yin +6

Multi-camera setups find widespread use across various applications, such as autonomous driving, as they greatly expand sensing capabilities. Despite the fast development of Neural…

cs.CV20234 cited

Metric3D: Towards Zero-shot Metric 3D Prediction from A Single Image

Wei Yin, Chi Zhang, Hao Chen +5

Reconstructing accurate 3D scenes from images is a long-standing vision task. Due to the ill-posedness of the single-image reconstruction problem, most well-established methods are…

cs.CV202222 cited

Multi-Camera Collaborative Depth Prediction via Consistent Structure Estimation

Jialei Xu, Xianming Liu, Yuanchao Bai +4

Depth map estimation from images is an important task in robotic systems. Existing methods can be categorized into two groups including multi-view stereo and monocular depth estima…