3 papers
cs.CV2026
Zero-shot 2D Grounding with Novel Affordance Types
Haomeng Zhang, Raymond A. Yeh
2D affordance grounding aims to locate the region of an object that a human can interact with. Existing research focuses on recognizing affordance types seen during training and do…
cs.CV2025
Auto-Vocabulary 3D Object Detection
Haomeng Zhang, Kuan-Chuan Peng, Suhas Lohit +1
Open-vocabulary 3D object detection methods are able to localize 3D boxes of classes unseen during training. Despite the name, existing methods rely on user-specified classes both…
cs.CV2024
Multi-Object 3D Grounding with Dynamic Modules and Language-Informed Spatial Attention
Haomeng Zhang, Chiao-An Yang, Raymond A. Yeh
Multi-object 3D Grounding involves locating 3D boxes based on a given query phrase from a point cloud. It is a challenging and significant task with numerous applications in visual…