2 papers
cs.CV2025
LLaVA: Representing 3D Scenes like a Cubist Painter to Boost 3D Scene Understanding of VLMs
Doriand Petit, Steve Bourgeois, Vincent Gay-Bellile +2
Developing a multi-modal language model capable of understanding 3D scenes remains challenging due to the limited availability of 3D training data, in contrast to the abundance of…
cs.CV2025
DiSCO-3D : Discovering and segmenting Sub-Concepts from Open-vocabulary queries in NeRF
Doriand Petit, Steve Bourgeois, Vincent Gay-Bellile +2
3D semantic segmentation provides high-level scene understanding for applications in robotics, autonomous systems, \textit{etc}. Traditional methods adapt exclusively to either tas…