1 citations · 1 across the 2 of their papers we have counts for
3 papers
cs.CV2026★ 1 cited
Are Multimodal Large Language Models Good Annotators for Image Tagging?
Ming-Kun Xie, Jia-Hao Xiao, Zhiqiang Kou +3
Image tagging, a fundamental vision task, traditionally relies on human-annotated datasets to train multi-label classifiers, which incurs significant labor and costs. While Multimo…
cs.CV2026
LoGoSeg: Integrating Local and Global Features for Open-Vocabulary Semantic Segmentation
Junyang Chen, Xiangbo Lv, Zhiqiang Kou +3
Open-vocabulary semantic segmentation (OVSS) extends traditional closed-set segmentation by enabling pixel-wise annotation for both seen and unseen categories using arbitrary textu…
cs.CV2025
Object-level Correlation for Few-Shot Segmentation
Chunlin Wen, Yu Zhang, Jie Fan +5
Few-shot semantic segmentation (FSS) aims to segment objects of novel categories in the query images given only a few annotated support samples. Existing methods primarily build th…