collaborators

5 papers

eess.IV2025

Understanding What Is Not Said:Referring Remote Sensing Image Segmentation with Scarce Expressions

Kai Ye, Bowen Liu, Jianghang Lin +3

Referring Remote Sensing Image Segmentation (RRSIS) aims to segment instances in remote sensing images according to referring expressions. Unlike Referring Image Segmentation on ge…

cs.CV2025

FlexiReID: Adaptive Mixture of Expert for Multi-Modal Person Re-Identification

Zhen Sun, Lei Tan, Yunhang Shen +5

Multimodal person re-identification (Re-ID) aims to match pedestrian images across different modalities. However, most existing methods focus on limited cross-modal settings and fa…

cs.CV2025

RIS-LAD: A Benchmark and Model for Referring Low-Altitude Drone Image Segmentation

Kai Ye, YingShi Luan, Zhudi Chen +3

Referring Image Segmentation (RIS), which aims to segment specific objects based on natural language descriptions, plays an essential role in vision-language understanding. Despite…

cs.CV2025

More Clear, More Flexible, More Precise: A Comprehensive Oriented Object Detection benchmark for UAV

Kai Ye, Haidi Tang, Bowen Liu +3

Applications of unmanned aerial vehicle (UAV) in logistics, agricultural automation, urban management, and emergency response are highly dependent on oriented object detection (OOD…

cs.CV2024

Knowing Where to Focus: Attention-Guided Alignment for Text-based Person Search

Lei Tan, Weihao Li, Pingyang Dai +3

In the realm of Text-Based Person Search (TBPS), mainstream methods aim to explore more efficient interaction frameworks between text descriptions and visual data. However, recent…