1 citations · 1 across the 9 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Em-Garde: A Propose-Match Framework for Proactive Streaming Video Understanding
Yikai Zheng, Xin Ding, Yifan Yang +6
Recent advances in Streaming Video Understanding has enabled a new interaction paradigm where models respond proactively to user queries. Current proactive VideoLLMs rely on per-fr…
cs.CV2025
Selective Structured State Space for Multispectral-fused Small Target Detection
Qianqian Zhang, WeiJun Wang, Yunxing Liu +4
Target detection in high-resolution remote sensing imagery faces challenges due to the low recognition accuracy of small targets and high computational costs. The computational com…
cs.CV2024
MobiFuse: A High-Precision On-device Depth Perception System with Multi-Data Fusion
Jinrui Zhang, Deyu Zhang, Tingting Long +6
We present MobiFuse, a high-precision depth perception system on mobile devices that combines dual RGB and Time-of-Flight (ToF) cameras. To achieve this, we leverage physical princ…