3 papers
cs.CV2026
Progressive Reasoning with Primitive Correction for Compositional Zero-Shot Learning
Ziyi Chen, Haoyan Shi, Sunhan Xu +1
Compositional Zero-Shot Learning (CZSL) aims to combine known attributes and objects as primitives for recognizing previously unseen attribute-object pairs. Prior works either pred…
cs.CV2026
How Far Are Video Models from True Multimodal Reasoning?
Xiaotian Zhang, Jianhui Wei, Yuan Wang +9
Despite remarkable progress toward general-purpose video models, a critical question remains unanswered: how far are these models from achieving true multimodal reasoning? Existing…
cs.CV2025
Query-aware Hub Prototype Learning for Few-Shot 3D Point Cloud Semantic Segmentation
YiLin Zhou, Lili Wei, Zheming Xu +2
Few-shot 3D point cloud semantic segmentation (FS-3DSeg) aims to segment novel classes with only a few labeled samples. However, existing metric-based prototype learning methods ge…