9 citations · 22 across the 10 of their papers we have counts for
6 papers · 1 filter
Exploring the Role of Explicit Temporal Modeling in Multimodal Large Language Models for Video Understanding
Yun Li, Zhe Liu, Yajing Kong +6
Applying Multimodal Large Language Models (MLLMs) to video understanding presents significant challenges due to the need to model temporal relations across frames. Existing approac…
Compositional Zero-Shot Learning with Contextualized Cues and Adaptive Contrastive Training
Yun Li, Zhe Liu, Lina Yao
Compositional Zero-Shot Learning (CZSL) aims to recognize unseen combinations of seen attributes and objects. Current CLIP-based methods in CZSL, despite their advancements, often…
Context-based and Diversity-driven Specificity in Compositional Zero-Shot Learning
Yun Li, Zhe Liu, Hang Chen +1
Compositional Zero-Shot Learning (CZSL) aims to recognize unseen attribute-object pairs based on a limited set of observed examples. Current CZSL methodologies, despite their advan…
Simple Primitives with Feasibility- and Contextuality-Dependence for Open-World Compositional Zero-shot Learning
Zhe Liu, Yun Li, Lina Yao +4
The task of Compositional Zero-Shot Learning (CZSL) is to recognize images of novel state-object compositions that are absent during the training stage. Previous methods of learnin…
An Entropy-guided Reinforced Partial Convolutional Network for Zero-Shot Learning
Yun Li, Zhe Liu, Lina Yao +3
Zero-Shot Learning (ZSL) aims to transfer learned knowledge from observed classes to unseen classes via semantic correlations. A promising strategy is to learn a global-local repre…
Task Aligned Generative Meta-learning for Zero-shot Learning
Zhe Liu, Yun Li, Lina Yao +2
Zero-shot learning (ZSL) refers to the problem of learning to classify instances from the novel classes (unseen) that are absent in the training set (seen). Most ZSL methods infer…