4 citations · 4 across the 2 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
3D Primitives are a Spatial Language for VLMs
Junze Liu, Kun Qian, Florian Dubost +8
Vision-language models (VLMs) exhibit a striking paradox: they can generate executable code that reconstructs a 3D scene from geometric primitives with correct object counts, class…
cs.CV2025
Restereo: Diffusion stereo video generation and restoration
Xingchang Huang, Ashish Kumar Singh, Florian Dubost +6
Stereo video generation has been gaining increasing attention with recent advancements in video diffusion models. However, most existing methods focus on generating 3D stereoscopic…