5 citations · 7 across the 4 of their papers we have counts for
4 papers
Exploring the Distinctiveness and Fidelity of the Descriptions Generated by Large Vision-Language Models
Yuhang Huang, Zihan Wu, Chongyang Gao +2
Large Vision-Language Models (LVLMs) are gaining traction for their remarkable ability to process and integrate visual and textual data. Despite their popularity, the capacity of L…
Beyond Static Evaluation: A Dynamic Approach to Assessing AI Assistants' API Invocation Capabilities
Honglin Mu, Yang Xu, Yunlong Feng +4
With the rise of Large Language Models (LLMs), AI assistants' ability to utilize tools, especially through API calls, has advanced notably. This progress has necessitated more accu…
Automatically Discovering Novel Visual Categories with Self-supervised Prototype Learning
Lu Zhang, Lu Qi, Xu Yang +3
This paper tackles the problem of novel category discovery (NCD), which aims to discriminate unknown categories in large-scale image collections. The NCD task is challenging due to…
Siamese Contrastive Embedding Network for Compositional Zero-Shot Learning
Xiangyu Li, Xu Yang, Kun Wei +2
Compositional Zero-Shot Learning (CZSL) aims to recognize unseen compositions formed from seen state and object during training. Since the same state may be various in the visual a…