2 citations · 2 across the 2 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
ZSPAPrune: Zero-Shot Prompt-Aware Token Pruning for Vision-Language Models
Pu Zhang, Yuwei Li, Xingyuan Xian +1
As the capabilities of Vision-Language Models (VLMs) advance, they can process increasingly large inputs, which, unlike in LLMs, generates significant visual token redundancy and l…
cs.CV2025
Falcon: A Remote Sensing Vision-Language Foundation Model (Technical Report)
Kelu Yao, Nuo Xu, Rong Yang +8
This paper introduces a holistic vision-language foundation model tailored for remote sensing, named Falcon. Falcon offers a unified, prompt-based paradigm that effectively execute…