3 papers
cs.CV2026
AdaDINO: Pair-Aware In-Backbone Adaptation of Frozen DINO for Efficient Remote Sensing Change Detection
Xu Zhang, Xinqing Li, Jianpeng Xie +3
Vision foundation models (VFMs) such as DINO are pretrained for single-image representation, whereas remote sensing change detection requires reasoning over a bi-temporal pair. Exi…
cs.CV2026
LAD-COD: Language-Aligned Dense Perception for Camouflaged Object Detection
Shangye Song, Tianzhi Zhu, Syed Ariff Syed Hesham +2
Camouflaged object detection (COD) aims to segment objects that exhibit high visual similarity to their surroundings, which reduces foreground-background discriminability and weake…
cs.CV2026
Towards Joint Quantization and Token Pruning of Vision-Language Models
Xinqing Li, Xin He, Xindong Zhang +3
Deploying Vision-Language Models (VLMs) under aggressive low-bit inference remains challenging because inference cost is dominated by the long visual-token prefix during prefill an…