88 citations · 145 across the 9 of their papers we have counts for
12 papers · 1 filter
Delta Tuning: A Comprehensive Study of Parameter Efficient Methods for Pre-trained Language Models
Ning Ding, Yujia Qin, Guang Yang +17
Despite the success, the process of fine-tuning large-scale PLMs brings prohibitive adaptation costs. In fact, fine-tuning all the parameters of a colossal model and retaining sepa…
Language Models are Good Translators
Shuo Wang, Zhaopeng Tu, Zhixing Tan +3
Recent years have witnessed the rapid advance in neural machine translation (NMT), the core of which lies in the encoder-decoder architecture. Inspired by the recent progress of la…
CPM-2: Large-scale Cost-effective Pre-trained Language Models
Zhengyan Zhang, Yuxian Gu, Xu Han +16
In recent years, the size of pre-trained language models (PLMs) has grown by leaps and bounds. However, efficiency issues of these large-scale PLMs limit their utilization in real-…
Alternated Training with Synthetic and Authentic Data for Neural Machine Translation
Rui Jiao, Zonghan Yang, Maosong Sun +1
While synthetic bilingual corpora have demonstrated their effectiveness in low-resource neural machine translation (NMT), adding more synthetic data often deteriorates translation…
On the Language Coverage Bias for Neural Machine Translation
Shuo Wang, Zhaopeng Tu, Zhixing Tan +3
Language coverage bias, which indicates the content-dependent differences between sentence pairs originating from the source and target languages, is important for neural machine t…
Transfer Learning for Sequence Generation: from Single-source to Multi-source
Xuancheng Huang, Jingfang Xu, Maosong Sun +1
Multi-source sequence generation (MSG) is an important kind of sequence generation tasks that takes multiple sources, including automatic post-editing, multi-source translation, mu…