7 citations · 7 across the 2 of their papers we have counts for
3 papers
cs.CL2022
ProGen: Progressive Zero-shot Dataset Generation via In-context Feedback
Jiacheng Ye, Jiahui Gao, Jiangtao Feng +3
Recently, dataset-generation-based zero-shot learning has shown promising results by training a task-specific model with a dataset synthesized from large pre-trained language model…
cs.LG2022★ 7 cited
Revisiting Over-smoothing in BERT from the Perspective of Graph
Han Shi, Jiahui Gao, Hang Xu +5
Recently over-smoothing phenomenon of Transformer-based models is observed in both vision and language fields. However, no existing work has delved deeper to further investigate th…
cs.LG2021
SparseBERT: Rethinking the Importance Analysis in Self-attention
Han Shi, Jiahui Gao, Xiaozhe Ren +4
Transformer-based models are popularly used in natural language processing (NLP). Its core component, self-attention, has aroused widespread interest. To understand the self-attent…