activity
20182021
most citedFBNetV2: Differentiable Neural Architecture Search for Spatial and Channel Dimensions

29 citations · 79 across the 8 of their papers we have counts for

collaborators

17 papers

cs.CV20241 cited

Imagine yourself: Tuning-Free Personalized Image Generation

Zecheng He, Bo Sun, Felix Juefei-Xu +14

Diffusion models have demonstrated remarkable efficacy across various image-to-image tasks. In this research, we introduce Imagine yourself, a state-of-the-art model designed for p…

cs.CV20224 cited

3D-Aware Encoding for Style-based Neural Radiance Fields

Yu-Jhe Li, Tao Xu, Bichen Wu +6

We tackle the task of NeRF inversion for style-based neural radiance fields, (e.g., StyleNeRF). In the task, we aim to learn an inversion function to project an input image to the…

cs.CV202268 cited

UmeTrack: Unified multi-view end-to-end hand tracking for VR

Shangchen Han, Po-chen Wu, Yubo Zhang +13

Real-time tracking of 3D hand pose in world space is a challenging problem and plays an important role in VR interaction. Existing work in this space are limited to either producin…

cs.CV20222 cited

Hydra Attention: Efficient Attention with Many Heads

Daniel Bolya, Cheng-Yang Fu, Xiaoliang Dai +2

While transformers have begun to dominate many tasks in vision, applying them to large images is still computationally difficult. A large reason for this is that self-attention sca…

cs.CV20211 cited

Data-Efficient Language-Supervised Zero-Shot Learning with Self-Distillation

Ruizhe Cheng, Bichen Wu, Peizhao Zhang +2

Traditional computer vision models are trained to predict a fixed set of predefined categories. Recently, natural language has been shown to be a broader and richer source of super…

cs.CV202128 cited

Unbiased Teacher for Semi-Supervised Object Detection

Yen-Cheng Liu, Chih-Yao Ma, Zijian He +6

Semi-supervised learning, i.e., training networks with both labeled and unlabeled data, has made significant progress recently. However, existing works have primarily focused on im…