2 papers
cs.CV2026
Standalone DINOv3 for Training-Free Open-Vocabulary Semantic Segmentation in Remote Sensing
Changhao Zhao, Haoxiang Li, Yuke Li +2
Remote sensing semantic segmentation is hindered by costly pixel-level annotations, motivating training-free open-vocabulary methods. Recently, the recent release of DINOv3 brings…
cs.CV2026
Natural Language Camera Movement Understanding
Yuwen Tan, Joey Huang, Jin Huang +2
Understanding camera movement in natural language is critical for training and evaluating video generation models, among other applications. However, we demonstrate that existing v…