3 papers
cs.CV2025
Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation
Chun-Peng Chang, Chen-Yu Wang, Julian Schmidt +2
Recent advancements in video generation have substantially improved visual quality and temporal coherence, making these models increasingly appealing for applications such as auton…
cs.CV2024
3D Spatial Understanding in MLLMs: Disambiguation and Evaluation
Chun-Peng Chang, Alain Pagani, Didier Stricker
Multimodal Large Language Models (MLLMs) have made significant progress in tasks such as image captioning and question answering. However, while these models can generate realistic…
cs.CV2024
Uni-SLAM: Uncertainty-Aware Neural Implicit SLAM for Real-Time Dense Indoor Scene Reconstruction
Shaoxiang Wang, Yaxu Xie, Chun-Peng Chang +3
Neural implicit fields have recently emerged as a powerful representation method for multi-view surface reconstruction due to their simplicity and state-of-the-art performance. How…