3 papers
cs.CV2025
Vision-based 3D Semantic Scene Completion via Capture Dynamic Representations
Meng Wang, Fan Wu, Yunchuan Qin +3
The vision-based semantic scene completion task aims to predict dense geometric and semantic 3D scene representations from 2D images. However, the presence of dynamic objects in th…
cs.CV2025
Learning Temporal 3D Semantic Scene Completion via Optical Flow Guidance
Meng Wang, Fan Wu, Ruihui Li +3
3D Semantic Scene Completion (SSC) provides comprehensive scene geometry and semantics for autonomous driving perception, which is crucial for enabling accurate and reliable decisi…
cs.CV2024
PAR: Prompt-Aware Token Reduction Method for Efficient Large Multimodal Models
Yingen Liu, Fan Wu, Ruihui Li +2
Multimodal large language models (MLLMs) demonstrate strong performance across visual tasks, but their efficiency is hindered by significant computational and memory demands from p…