activity
20162022
most citedTowards Lightweight Transformer via Group-wise Transformation for Vision-and-Language Tasks

66 citations · 106 across the 6 of their papers we have counts for

collaborators

15 papers

cs.CV202266 cited

Towards Lightweight Transformer via Group-wise Transformation for Vision-and-Language Tasks

Gen Luo, Yiyi Zhou, Xiaoshuai Sun +5

Despite the exciting performance, Transformer is criticized for its excessive parameters and computation cost. However, compressing Transformer remains as an open problem due to it…

cs.CV20218 cited

HifiFace: 3D Shape and Semantic Prior Guided High Fidelity Face Swapping

Yuhan Wang, Xu Chen, Junwei Zhu +7

In this work, we propose a high fidelity face swapping method, called HifiFace, which can well preserve the face shape of the source face and generate photo-realistic results. Unli…

cs.CV20219 cited

Image-to-image Translation via Hierarchical Style Disentanglement

Xinyang Li, Shengchuan Zhang, Jie Hu +6

Recently, image-to-image translation has made significant progress in achieving both multi-label (\ie, translation conditioned on different labels) and multi-style (\ie, generation…

cs.CV2021

Network Pruning using Adaptive Exemplar Filters

Mingbao Lin, Rongrong Ji, Shaojie Li +4

Popular network pruning algorithms reduce redundant information by optimizing hand-crafted models, and may cause suboptimal performance and long time in selecting filters. We innov…

cs.CV2021

Dual-Level Collaborative Transformer for Image Captioning

Yunpeng Luo, Jiayi Ji, Xiaoshuai Sun +5

Descriptive region features extracted by object detection networks have played an important role in the recent advancements of image captioning. However, they are still criticized…

cs.CV202018 cited

Improving Image Captioning by Leveraging Intra- and Inter-layer Global Representation in Transformer Network

Jiayi Ji, Yunpeng Luo, Xiaoshuai Sun +5

Transformer-based architectures have shown great success in image captioning, where object regions are encoded and then attended into the vectorial representations to guide the cap…