activity
20182022
most citedTowards Lightweight Transformer via Group-wise Transformation for Vision-and-Language Tasks

66 citations · 128 across the 11 of their papers we have counts for

collaborators

23 papers

cs.LG202217 cited

Make Sharpness-Aware Minimization Stronger: A Sparsified Perturbation Approach

Peng Mi, Li Shen, Tianhe Ren +4

Deep neural networks often suffer from poor generalization caused by complex and non-convex loss landscapes. One of the popular solutions is Sharpness-Aware Minimization (SAM), whi…

cs.CV202266 cited

Towards Lightweight Transformer via Group-wise Transformation for Vision-and-Language Tasks

Gen Luo, Yiyi Zhou, Xiaoshuai Sun +5

Despite the exciting performance, Transformer is criticized for its excessive parameters and computation cost. However, compressing Transformer remains as an open problem due to it…

cs.CV2021

Dual-Level Collaborative Transformer for Image Captioning

Yunpeng Luo, Jiayi Ji, Xiaoshuai Sun +5

Descriptive region features extracted by object detection networks have played an important role in the recent advancements of image captioning. However, they are still criticized…

cs.CV202018 cited

Improving Image Captioning by Leveraging Intra- and Inter-layer Global Representation in Transformer Network

Jiayi Ji, Yunpeng Luo, Xiaoshuai Sun +5

Transformer-based architectures have shown great success in image captioning, where object regions are encoded and then attended into the vectorial representations to guide the cap…

cs.CV20201 cited

Fast Class-wise Updating for Online Hashing

Mingbao Lin, Rongrong Ji, Xiaoshuai Sun +4

Online image hashing has received increasing research attention recently, which processes large-scale data in a streaming fashion to update the hash functions on-the-fly. To this e…

cs.CV2020

Multi-task Collaborative Network for Joint Referring Expression Comprehension and Segmentation

Gen Luo, Yiyi Zhou, Xiaoshuai Sun +4

Referring expression comprehension (REC) and segmentation (RES) are two highly-related tasks, which both aim at identifying the referent according to a natural language expression.…