3 papers
cs.CV2026
DinoLink: A Token-Centric Representation Compression Framework for Bandwidth-Constrained Collaborative V2X Perception
Tianle Zhu, Haohua Que, Handong Yao +2
High-precision remote perception is often hindered by the severe bandwidth constraints of Vehicle-to-Everything (V2X) networks. We propose \textit{DinoLink}, a token-centric compre…
cs.CV2026
CABLE: Cloud-Assisted Bandwidth-efficient LMM-based Encoding for V2X Systems
Haohua Que, Zhipeng Bao, Qianyi Wu +1
Cloud-hosted large multimodal models (LMMs) can provide strong open-vocabulary perception for Vehicle-to-Everything systems, but naively transmitting full-resolution frames from ed…
cs.AI2026
BlazeEdit: Generalist Image Editing on Mobile Devices with Image-to-Image Diffusion Models
Fei Deng, Yanwu Xu, Zhipeng Bao +4
The remarkable generation quality of modern diffusion models often comes at the cost of massive parameter counts, which necessitate server-side inference with significant computati…