4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2024
PD-APE: A Parallel Decoding Framework with Adaptive Position Encoding for 3D Visual Grounding
Chenshu Hou, Liang Peng, Xiaopei Wu +2
3D visual grounding aims to identify objects in 3D point cloud scenes that match specific natural language descriptions. This requires the model to not only focus on the target obj…
cs.CV2024★ 4 cited
VastTrack: Vast Category Visual Object Tracking
Liang Peng, Junyuan Gao, Xinran Liu +5
In this paper, we introduce a novel benchmark, dubbed VastTrack, towards facilitating the development of more general visual tracking via encompassing abundant classes and videos.…