1 paper
Feng Xiao, Hongbin Xu, Guocan Zhao +1
3D visual grounding aims to localize the unique target described by natural languages in 3D scenes. The significant gap between 3D and language modalities makes it a notable challe…