1 citations · 1 across the 1 of their papers we have counts for
4 papers
Structure-Guided Self-Supervised Matching for One-Shot Medical Landmark Detection
Qingsong Yao, Zhen Huang, Ao Wang +4
Medical landmark detection usually requires accurate expert annotations, which are laborious and difficult to scale across anatomical regions. In this work, we study an extreme ann…
Test-time Correction: An Online 3D Detection System via Visual Prompting
Hanxue Zhang, Zetong Yang, Yanan Sun +4
This paper introduces Test-time Correction (TTC), an online 3D detection system designed to rectify test-time errors using various auxiliary feedback, aiming to enhance the safety…
Detect Anything 3D in the Wild
Hanxue Zhang, Haoran Jiang, Qingsong Yao +6
Despite the success of deep learning in close-set 3D object detection, existing approaches struggle with zero-shot generalization to novel objects and camera configurations. We int…
DriveLM: Driving with Graph Visual Question Answering
Chonghao Sima, Katrin Renz, Kashyap Chitta +7
We study how vision-language models (VLMs) trained on web-scale data can be integrated into end-to-end driving systems to boost generalization and enable interactivity with human u…