most citedVision Permutator: A Permutable MLP-Like Architecture for Visual Recognition

25 citations · 54 across the 4 of their papers we have counts for

collaborators

9 papers

cs.CV202123 cited

VOLO: Vision Outlooker for Visual Recognition

Li Yuan, Qibin Hou, Zihang Jiang +2

Visual recognition has been dominated by convolutional neural networks (CNNs) for years. Though recently the prevailing vision transformers (ViTs) have shown great potential of sel…

cs.CV202125 cited

Vision Permutator: A Permutable MLP-Like Architecture for Visual Recognition

Qibin Hou, Zihang Jiang, Li Yuan +3

In this paper, we present Vision Permutator, a conceptually simple and data efficient MLP-like architecture for visual recognition. By realizing the importance of the positional in…

cs.CV2021

All Tokens Matter: Token Labeling for Training Better Vision Transformers

Zihang Jiang, Qibin Hou, Li Yuan +5

In this paper, we present token labeling -- a new training objective for training high-performance vision transformers (ViTs). Different from the standard training objective of ViT…

cs.CV2019

Revisiting Knowledge Distillation via Label Smoothing Regularization

Li Yuan, Francis E. H. Tay, Guilin Li +2

Knowledge Distillation (KD) aims to distill the knowledge of a cumbersome teacher model into a lightweight student model. Its success is generally attributed to the privileged info…

cs.LG2019

Modular Deep Reinforcement Learning with Temporal Logic Specifications

Lim Zun Yuan, Mohammadhosein Hasanbeig, Alessandro Abate +1

We propose an actor-critic, model-free, and online Reinforcement Learning (RL) framework for continuous-state continuous-action Markov Decision Processes (MDPs) when the reward is…

cs.CV2019

Central Similarity Quantization for Efficient Image and Video Retrieval

Li Yuan, Tao Wang, Xiaopeng Zhang +4

Existing data-dependent hashing methods usually learn hash functions from pairwise or triplet data relationships, which only capture the data similarity locally, and often suffer f…