2 citations · 2 across the 5 of their papers we have counts for
1 paper · 1 filter
Jiarun Liu, Qifeng Chen, Yiru Zhao +3
While visual-language models have profoundly linked features between texts and images, the incorporation of 3D modality data, such as point clouds and 3D Gaussians, further enables…