3 papers
cs.CV2026
An Open-Source Benchmark and Baseline for Multi-temporal Referring Segmentation
Bingyu Li, Da Zhang, Tao Huo +3
Large Vision-Language Models (LVLMs) have shown strong visual understanding and language-guided grounding abilities, yet their capacity for multi-temporal visual reasoning remains…
cs.CV2026
Towards Realistic Open-Vocabulary Remote Sensing Segmentation: Benchmark and Baseline
Bingyu Li, Tao Huo, Haocheng Dong +4
Open-vocabulary remote sensing image segmentation (OVRSIS) remains underexplored due to fragmented datasets, limited training diversity, and the lack of evaluation benchmarks that…
cs.CV2025
Exploring the Underwater World Segmentation without Extra Training
Bingyu Li, Tao Huo, Da Zhang +3
Accurate segmentation of marine organisms is vital for biodiversity monitoring and ecological assessment, yet existing datasets and models remain largely limited to terrestrial sce…