activity
20172022
most citedRTFormer: Efficient Design for Real-Time Semantic Segmentation with Transformer

76 citations · 368 across the 23 of their papers we have counts for

collaborators
Showing cs.CVShow all

31 papers · 1 filter

cs.CV20221 cited

Cyclically Disentangled Feature Translation for Face Anti-spoofing

Haixiao Yue, Keyao Wang, Guosheng Zhang +4

Current domain adaptation methods for face anti-spoofing leverage labeled source domain data and unlabeled target domain data to obtain a promising generalizable decision boundary.…

cs.CV20228 cited

CAE v2: Context Autoencoder with CLIP Target

Xinyu Zhang, Jiahui Chen, Junkun Yuan +10

Masked image modeling (MIM) learns visual representation by masking and reconstructing image patches. Applying the reconstruction supervision on the CLIP representation has been pr…

cs.CV202219 cited

Group DETR v2: Strong Object Detector with Encoder-Decoder Pretraining

Qiang Chen, Jian Wang, Chuchu Han +12

We present a strong object detector with encoder-decoder pretraining and finetuning. Our method, called Group DETR v2, is built upon a vision transformer encoder ViT-Huge~\cite{dos…

cs.CV202276 cited

RTFormer: Efficient Design for Real-Time Semantic Segmentation with Transformer

Jian Wang, Chenhui Gou, Qiman Wu +4

Recently, transformer-based networks have shown impressive results in semantic segmentation. Yet for real-time semantic segmentation, pure CNN-based approaches still dominate in th…

cs.CV20221 cited

StyleSwap: Style-Based Generator Empowers Robust Face Swapping

Zhiliang Xu, Hang Zhou, Zhibin Hong +7

Numerous attempts have been made to the task of person-agnostic face swapping given its wide applications. While existing methods mostly rely on tedious network and loss designs, t…

cs.CV2022

Few-Shot Head Swapping in the Wild

Changyong Shu, Hemao Wu, Hang Zhou +7

The head swapping task aims at flawlessly placing a source head onto a target body, which is of great importance to various entertainment scenarios. While face swapping has drawn m…