3 papers
cs.CV2025
Learning Temporal 3D Semantic Scene Completion via Optical Flow Guidance
Meng Wang, Fan Wu, Ruihui Li +3
3D Semantic Scene Completion (SSC) provides comprehensive scene geometry and semantics for autonomous driving perception, which is crucial for enabling accurate and reliable decisi…
cs.CV2025
Vision-based 3D Semantic Scene Completion via Capture Dynamic Representations
Meng Wang, Fan Wu, Yunchuan Qin +3
The vision-based semantic scene completion task aims to predict dense geometric and semantic 3D scene representations from 2D images. However, the presence of dynamic objects in th…
cs.CV2024
PAR: Prompt-Aware Token Reduction Method for Efficient Large Multimodal Models
Yingen Liu, Fan Wu, Ruihui Li +2
Multimodal large language models (MLLMs) demonstrate strong performance across visual tasks, but their efficiency is hindered by significant computational and memory demands from p…