1 paper
Yunsong Wang, Hanlin Chen, Gim Hee Lee
Recent advancements in vision-language foundation models have significantly enhanced open-vocabulary 3D scene understanding. However, the generalizability of existing methods is co…