activity
20142025
most citedTaming Encoder for Zero Fine-tuning Image Customization with Text-to-Image Diffusion Models

21 citations · 68 across the 19 of their papers we have counts for

collaborators

6 papers

cs.CV2022

Open Long-Tailed Recognition in a Dynamic World

Ziwei Liu, Zhongqi Miao, Xiaohang Zhan +3

Real world data often exhibits a long-tailed and open-ended (with unseen classes) distribution. A practical recognition system must balance between majority (head) and minority (ta…

cs.CV20219 cited

Towards a Unified Foundation Model: Jointly Pre-Training Transformers on Unpaired Images and Text

Qing Li, Boqing Gong, Yin Cui +4

In this paper, we explore the possibility of building a unified foundation model that can be adapted to both vision-only and text-only tasks. Starting from BERT and ViT, we design…

cs.CV20214 cited

Exploring Temporal Granularity in Self-Supervised Video Representation Learning

Rui Qian, Yeqing Li, Liangzhe Yuan +7

This work presents a self-supervised learning framework named TeG to explore Temporal Granularity in learning video representations. In TeG, we sample a long clip from a video and…

cs.CV20212 cited

Contextualized Spatio-Temporal Contrastive Learning with Self-Supervision

Liangzhe Yuan, Rui Qian, Yin Cui +5

Modern self-supervised learning algorithms typically enforce persistency of instance representations across views. While being very effective on learning holistic image and video r…

cs.CV2016

Query-Focused Extractive Video Summarization

Aidean Sharghi, Boqing Gong, Mubarak Shah

Video data is explosively growing. As a result of the "big video data", intelligent algorithms for automatic video summarization have re-emerged as a pressing need. We develop a pr…

stat.ML201414 cited

Large-Margin Determinantal Point Processes

Boqing Gong, Wei-lun Chao, Kristen Grauman +1

Determinantal point processes (DPPs) offer a powerful approach to modeling diversity in many applications where the goal is to select a diverse subset. We study the problem of lear…