Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
LoGoSeg: Integrating Local and Global Features for Open-Vocabulary Semantic Segmentation
Junyang Chen, Xiangbo Lv, Zhiqiang Kou +3
Open-vocabulary semantic segmentation (OVSS) extends traditional closed-set segmentation by enabling pixel-wise annotation for both seen and unseen categories using arbitrary textu…
cs.CV2025
HLV-1K: A Large-scale Hour-Long Video Benchmark for Time-Specific Long Video Understanding
Heqing Zou, Tianze Luo, Guiyang Xie +7
Multimodal large language models have become a popular topic in deep visual understanding due to many promising real-world applications. However, hour-long video understanding, spa…