25 citations · 25 across the 2 of their papers we have counts for
3 papers
cs.CV2026
Standalone DINOv3 for Training-Free Open-Vocabulary Semantic Segmentation in Remote Sensing
Changhao Zhao, Haoxiang Li, Yuke Li +2
Remote sensing semantic segmentation is hindered by costly pixel-level annotations, motivating training-free open-vocabulary methods. Recently, the recent release of DINOv3 brings…
cs.CV2026
Natural Language Camera Movement Understanding
Yuwen Tan, Joey Huang, Jin Huang +2
Understanding camera movement in natural language is critical for training and evaluating video generation models, among other applications. However, we demonstrate that existing v…
cs.CV2017★ 25 cited
VQS: Linking Segmentations to Questions and Answers for Supervised Attention in VQA and Question-Focused Semantic Segmentation
Chuang Gan, Yandong Li, Haoxiang Li +2
Rich and dense human labeled datasets are among the main enabling factors for the recent advance on vision-language understanding. Many seemingly distant annotations (e.g., semanti…