1 citations · 1 across the 2 of their papers we have counts for
4 papers
Are VLMs Ready for Autonomous Driving? An Empirical Study from the Reliability, Data, and Metric Perspectives
Shaoyuan Xie, Lingdong Kong, Yuhao Dong +5
Recent advancements in Vision-Language Models (VLMs) have sparked interest in their use for autonomous driving, particularly in generating interpretable driving decisions through n…
OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies
Runnan Chen, Xiangyu Sun, Zhaoqing Wang +8
Open-vocabulary scene understanding using 3D Gaussian (3DGS) representations has garnered considerable attention. However, existing methods mostly lift knowledge from large 2D visi…
GS-VTON: Controllable 3D Virtual Try-on with Gaussian Splatting
Yukang Cao, Masoud Hadi, Liang Pan +1
Diffusion-based 2D virtual try-on (VTON) techniques have recently demonstrated strong performance, while the development of 3D VTON has largely lagged behind. Despite recent advanc…
Multi-Space Alignments Towards Universal LiDAR Segmentation
Youquan Liu, Lingdong Kong, Xiaoyang Wu +5
A unified and versatile LiDAR segmentation model with strong robustness and generalizability is desirable for safe autonomous driving perception. This work presents M3Net, a one-of…