activity
20202022
most citedHRFormer: High-Resolution Transformer for Dense Prediction

128 citations · 500 across the 27 of their papers we have counts for

collaborators

29 papers

cs.CV20228 cited

Augmentation Matters: A Simple-yet-Effective Approach to Semi-supervised Semantic Segmentation

Zhen Zhao, Lihe Yang, Sifan Long +3

Recent studies on semi-supervised semantic segmentation (SSS) have seen fast progress. Despite their promising performance, current state-of-the-art methods tend to increasingly co…

cs.CV2022

Masked Lip-Sync Prediction by Audio-Visual Contextual Exploitation in Transformers

Yasheng Sun, Hang Zhou, Kaisiyuan Wang +7

Previous studies have explored generating accurately lip-synced talking faces for arbitrary targets given audio conditions. However, most of them deform or generate the whole facia…

cs.CV20221 cited

Cyclically Disentangled Feature Translation for Face Anti-spoofing

Haixiao Yue, Keyao Wang, Guosheng Zhang +4

Current domain adaptation methods for face anti-spoofing leverage labeled source domain data and unlabeled target domain data to obtain a promising generalizable decision boundary.…

cs.CV202217 cited

Real-time Neural Radiance Talking Portrait Synthesis via Audio-spatial Decomposition

Jiaxiang Tang, Kaisiyuan Wang, Hang Zhou +6

While dynamic Neural Radiance Fields (NeRF) have shown success in high-fidelity 3D modeling of talking portraits, the slow training and inference speed severely obstruct their pote…

cs.CV20223 cited

Instance-specific and Model-adaptive Supervision for Semi-supervised Semantic Segmentation

Zhen Zhao, Sifan Long, Jimin Pi +2

Recently, semi-supervised semantic segmentation has achieved promising performance with a small fraction of labeled data. However, most existing studies treat all unlabeled data eq…

cs.CV2022

Beyond Attentive Tokens: Incorporating Token Importance and Diversity for Efficient Vision Transformers

Sifan Long, Zhen Zhao, Jimin Pi +2

Vision transformers have achieved significant improvements on various vision tasks but their quadratic interactions between tokens significantly reduce computational efficiency. Ma…