3 papers
cs.AI2026
Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models
Yang Zhang, Nada Mimouni, Jean-Claude Moissinac +1
The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for their ability to structure vast…
cs.CV2026
WOW-Seg: A Word-free Open World Segmentation Model
Danyang Li, Tianhao Wu, Bin Li +5
Open world image segmentation aims to achieve precise segmentation and semantic understanding of targets within images by addressing the infinitely open set of object categories en…
cs.CV2025
Leveraging MLLM Embeddings and Attribute Smoothing for Compositional Zero-Shot Learning
Xudong Yan, Songhe Feng, Yang Zhang +3
Compositional zero-shot learning (CZSL) aims to recognize novel compositions of attributes and objects learned from seen compositions. Previous works disentangle attributes and obj…