1 paper
Junjie Wen, Junlin He, Fei Ma +1
Accurate open-vocabulary 3D scene understanding requires semantic representations that are both language-aligned and spatially precise at the pixel level, while remaining scalable…