2 papers
cs.LG2025
VLM in a flash: I/O-Efficient Sparsification of Vision-Language Model via Neuron Chunking
Kichang Yang, Seonjun Kim, Minjae Kim +3
Edge deployment of large Vision-Language Models (VLMs) increasingly relies on flash-based weight offloading, where activation sparsification is used to reduce I/O overhead. However…
cs.CV2024
MobiFuse: A High-Precision On-device Depth Perception System with Multi-Data Fusion
Jinrui Zhang, Deyu Zhang, Tingting Long +6
We present MobiFuse, a high-precision depth perception system on mobile devices that combines dual RGB and Time-of-Flight (ToF) cameras. To achieve this, we leverage physical princ…