3 citations · 6 across the 4 of their papers we have counts for
4 papers
Masked AutoDecoder is Effective Multi-Task Vision Generalist
Han Qiu, Jiaxing Huang, Peng Gao +3
Inspired by the success of general-purpose models in NLP, recent studies attempt to unify different vision tasks in the same sequence format and employ autoregressive Transformers…
LLMs Meet VLMs: Boost Open Vocabulary Object Detection with Fine-grained Descriptors
Sheng Jin, Xueying Jiang, Jiaxing Huang +2
Inspired by the outstanding zero-shot capability of vision language models (VLMs) in image classification tasks, open-vocabulary object detection has attracted increasing interest…
Domain Generalization via Balancing Training Difficulty and Model Capability
Xueying Jiang, Jiaxing Huang, Sheng Jin +1
Domain generalization (DG) aims to learn domain-generalizable models from one or multiple source domains that can perform well in unseen target domains. Despite its recent progress…
Black-box Unsupervised Domain Adaptation with Bi-directional Atkinson-Shiffrin Memory
Jingyi Zhang, Jiaxing Huang, Xueying Jiang +1
Black-box unsupervised domain adaptation (UDA) learns with source predictions of target data without accessing either source data or source models during training, and it has clear…