activity
20212024
most citedDesigning BERT for Convolutional Networks: Sparse and Hierarchical Masked Modeling

37 citations · 46 across the 10 of their papers we have counts for

collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV2024

OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation

Junke Wang, Yi Jiang, Zehuan Yuan +3

Tokenizer, serving as a translator to map the intricate visual data into a compact latent space, lies at the core of visual generative models. Based on the finding that existing to…

cs.CV2023

EGC: Image Generation and Classification via a Diffusion Energy-Based Model

Qiushan Guo, Chuofan Ma, Yi Jiang +3

Learning image classification and image generation using the same set of network parameters is a challenging problem. Recent advanced approaches perform well in one task often exhi…

cs.CV20231 cited

Multi-Level Contrastive Learning for Dense Prediction Task

Qiushan Guo, Yizhou Yu, Yi Jiang +3

In this work, we present Multi-Level Contrastive Learning for Dense Prediction Task (MCL), an efficient self-supervised method for learning region-level feature representation for…

cs.CV2023

InstMove: Instance Motion for Object-centric Video Segmentation

Qihao Liu, Junfeng Wu, Yi Jiang +3

Despite significant efforts, cutting-edge video segmentation methods still remain sensitive to occlusion and rapid movement, due to their reliance on the appearance of objects in t…

cs.CV202337 cited

Designing BERT for Convolutional Networks: Sparse and Hierarchical Masked Modeling

Keyu Tian, Yi Jiang, Qishuai Diao +3

We identify and overcome two key obstacles in extending the success of BERT-style pre-training, or the masked image modeling, to convolutional networks (convnets): (i) convolution…

cs.CV2022

Towards Grand Unification of Object Tracking

Bin Yan, Yi Jiang, Peize Sun +4

We present a unified method, termed Unicorn, that can simultaneously solve four tracking problems (SOT, MOT, VOS, MOTS) with a single network using the same model parameters. Due t…