19 citations · 21 across the 3 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024★ 1 cited
AffordanceLLM: Grounding Affordance from Vision Language Models
Shengyi Qian, Weifeng Chen, Min Bai +3
Affordance grounding refers to the task of finding the area of an object with which one can interact. It is a fundamental but challenging task, as a successful solution requires th…
cs.CV2023★ 1 cited
LiDAR-Based 3D Object Detection via Hybrid 2D Semantic Scene Generation
Haitao Yang, Zaiwei Zhang, Xiangru Huang +5
Bird's-Eye View (BEV) features are popular intermediate scene representations shared by the 3D backbone and the detector head in LiDAR-based object detectors. However, little resea…
cs.CV2016★ 19 cited
TorontoCity: Seeing the World with a Million Eyes
Shenlong Wang, Min Bai, Gellert Mattyus +7
In this paper we introduce the TorontoCity benchmark, which covers the full greater Toronto area (GTA) with 712.5 of land, 8439 of road and around 400,000 buildings. Ou…