Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Towards Visual Query Localization in the 3D World
Liang Peng, Bohan Tan, Zhipeng Zhang +4
Visual query localization (VQL) aims to predict the spatio-temporal response of the most recent occurrence in a sequence given a query. Currently, most research focuses on visual q…
cs.CV2025
DEGround: An Effective Baseline for Ego-centric 3D Visual Grounding with a Homogeneous Framework
Yani Zhang, Dongming Wu, Hao Shi +3
A core task in embodied intelligence is ego-centric 3D visual grounding. Existing methods typically adopt two-stage, heterogeneous pipelines that pair a detector with a separate gr…