Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models
Junming Huang, Chi Wang, Letian Li +5
Large Language Models(LLMs) have revolutionized text generation and multimodal perception,but their capabilities in 3D content generation remain underexplored. Existing methods com…
cs.CV2025
Enhanced Velocity Field Modeling for Gaussian Video Reconstruction
Zhenyang Li, Xiaoyang Bai, Tongchen Zhang +3
High-fidelity 3D video reconstruction is essential for enabling real-time rendering of dynamic scenes with realistic motion in virtual and augmented reality (VR/AR). The deformatio…