Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
CoFiE: Coarse-to-Fine Evidence Selection for Efficient Streaming Video Understanding
Jing Jiang, Yiran Ling, Ruonan Li +2
Streaming video understanding requires Vision Language Models (VLLMs) to process growing video streams and answer user questions under tight latency constraints. Existing methods i…
cs.CV2026
Distill, Diffuse, Segment: Unsupervised 3D Semantic Segmentation for Autonomous Driving Based on Multi-Level Distillation and Graph Diffusion
Yijing Wang, Ruonan Li, Qilin Wang +2
LiDAR-based semantic segmentation is essential for autonomous-driving perception, yet dense point-wise annotations are costly, and long-tailed outdoor scenes make small safety-crit…