activity
20242026
most citedIFShip: Interpretable Fine-grained Ship Classification with Domain Knowledge-Enhanced Vision-Language Models

12 citations · 12 across the 10 of their papers we have counts for

collaborators
Showing cs.CVShow all

7 papers · 1 filter

cs.CV2026

Think with Extra-Image: A Farmland Segmentation Agent Driven by Spatio-Temporal Information Gain

Haiyang Wu, Weiliang Mu, Zhuofei Du +4

Existing farmland remote sensing image (FRSI) segmentation follows a "Think with Intra-Image" paradigm, assuming that the current image contains sufficient visual evidence for reli…

cs.CV2026

iGSP:Implicit Gradient Subspace Projection for Efficient Continual Learning of Vision-Language Models

Xuezhi Cui, Dongbo Zhou, Wang Guo +8

Vision-Language Models require efficient adaptation to continually emerging downstream tasks. While Parameter-Efficient Fine-Tuning mitigates catastrophic forgetting, assigning iso…

cs.CV2025

Towards Accurate UAV Image Perception: Guiding Vision-Language Models with Stronger Task Prompts

Mingning Guo, Mengwei Wu, Shaoxian Li +2

Existing image perception methods based on VLMs generally follow a paradigm wherein models extract and analyze image content based on user-provided textual task prompts. However, s…

cs.CV2025

Enhancing Scene Classification in Cloudy Image Scenarios: A Collaborative Transfer Method with Information Regulation Mechanism using Optical Cloud-Covered and SAR Remote Sensing Images

Yuze Wang, Rong Xiao, Haifeng Li +2

In remote sensing scene classification, leveraging the transfer methods with well-trained optical models is an efficient way to overcome label scarcity. However, cloud contaminatio…

cs.CV2024

SeFi-CD: A Semantic First Change Detection Paradigm That Can Detect Any Change You Want

Ling Zhao, Zhenyang Huang, Dongsheng Kuang +3

The existing change detection(CD) methods can be summarized as the visual-first change detection (ViFi-CD) paradigm, which first extracts change features from visual differences an…

cs.CV2024

Homogeneous Tokenizer Matters: Homogeneous Visual Tokenizer for Remote Sensing Image Understanding

Run Shao, Zhaoyang Zhang, Chao Tao +3

The tokenizer, as one of the fundamental components of large models, has long been overlooked or even misunderstood in visual tasks. One key factor of the great comprehension power…