5 papers
Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions
Kai Ye, Bowen Liu, Jianghang Lin +3
Referring Remote Sensing Image Segmentation (RRSIS) aims to segment instances in remote sensing images according to referring expressions. Unlike Referring Image Segmentation on ge…
FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification
Zhen Sun, Lei Tan, Yunhang Shen +5
Multimodal person re-identification (Re-ID) aims to match pedestrian images across different modalities. However, most existing methods focus on limited cross-modal settings and fa…
RIS-LAD: A Benchmark and Model for Referring Low-Altitude Drone Image Segmentation
Kai Ye, YingShi Luan, Zhudi Chen +3
Referring Image Segmentation (RIS), which aims to segment specific objects based on natural language descriptions, plays an essential role in vision-language understanding. Despite…
More Clear, More Flexible, More Precise: A Comprehensive Oriented Object Detection benchmark for UAV
Kai Ye, Haidi Tang, Bowen Liu +3
Applications of unmanned aerial vehicle (UAV) in logistics, agricultural automation, urban management, and emergency response are highly dependent on oriented object detection (OOD…
Knowing Where to Focus: Attention-Guided Alignment for Text-based Person Search
Lei Tan, Weihao Li, Pingyang Dai +3
In the realm of Text-Based Person Search (TBPS), mainstream methods aim to explore more efficient interaction frameworks between text descriptions and visual data. However, recent…