2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.CV2026
Semantic-Driven Scale and Spatial Selection for Efficient Cross-Modal Alignment in Referring Remote Sensing Image Segmentation
Kun Li, Shengxi Gui, Francesco Nex +1
Referring Remote Sensing Image Segmentation (RRSIS) seeks to localize and segment the target object or region specified by a natural language expression in a remote sensing image.…
cs.CV2024★ 2 cited
AUG: A New Dataset and An Efficient Model for Aerial Image Urban Scene Graph Generation
Yansheng Li, Kun Li, Yongjun Zhang +2
Scene graph generation (SGG) aims to understand the visual objects and their semantic relationships from one given image. Until now, lots of SGG datasets with the eyelevel view are…
cs.CV2023
Interactive Image Segmentation with Cross-Modality Vision Transformers
Kun Li, George Vosselman, Michael Ying Yang
Interactive image segmentation aims to segment the target from the background with the manual guidance, which takes as input multimodal data such as images, clicks, scribbles, and…