3 papers
cs.CV2026
FIRM: Fine-Grained Intra-Token Representation of Masks for Remote Sensing Reasoning Segmentation
Weidong Tang, Kaiyu Li, Yikai Wang +4
Reasoning segmentation requires multimodal large language models (MLLMs) to translate implicit instructions into precise pixel-level masks. MLLMs encode an image as visual tokens,…
cs.CV2025
Annotation-Free Open-Vocabulary Segmentation for Remote-Sensing Images
Kaiyu Li, Xiangyong Cao, Ruixun Liu +4
Semantic segmentation of remote sensing (RS) images is pivotal for comprehensive Earth observation, but the demand for interpreting new object categories, coupled with the high exp…
cs.CV2024
Class Similarity Transition: Decoupling Class Similarities and Imbalance from Generalized Few-shot Segmentation
Shihong Wang, Ruixun Liu, Kaiyu Li +2
In Generalized Few-shot Segmentation (GFSS), a model is trained with a large corpus of base class samples and then adapted on limited samples of novel classes. This paper focuses o…