78 citations · 83 across the 5 of their papers we have counts for
4 papers · 1 filter
Context-Aware Token Selection and Packing for Enhanced Vision Transformer
Tianyi Zhang, Baoxin Li, Jae-sun Seo +1
In recent years, the long-range attention mechanism of vision transformers has driven significant performance breakthroughs across various computer vision tasks. However, the tradi…
Domain Adaptation Using Pseudo Labels
Sachin Chhabra, Hemanth Venkateswara, Baoxin Li
In the absence of labeled target data, unsupervised domain adaptation approaches seek to align the marginal distributions of the source and target domains in order to train a class…
Instance Adaptive Prototypical Contrastive Embedding for Generalized Zero Shot Learning
Riti Paul, Sahil Vora, Baoxin Li
Generalized zero-shot learning(GZSL) aims to classify samples from seen and unseen labels, assuming unseen labels are not accessible during training. Recent advancements in GZSL ha…
Hierarchical Attention Network for Action Recognition in Videos
Yilin Wang, Suhang Wang, Jiliang Tang +3
Understanding human actions in wild videos is an important task with a broad range of applications. In this paper we propose a novel approach named Hierarchical Attention Network (…