activity
20162022
most citedNon-Autoregressive Neural Machine Translation with Enhanced Decoder Input

16 citations · 28 across the 5 of their papers we have counts for

collaborators

13 papers

cs.SD2022

Bridging Music and Text with Crowdsourced Music Comments: A Sequence-to-Sequence Framework for Thematic Music Comments Generation

Peining Zhang, Junliang Guo, Linli Xu +2

We consider a novel task of automatically generating text descriptions of music. Compared with other well-established text generation tasks such as image caption, the scarcity of w…

cs.CL20222 cited

Sequence-to-Action: Grammatical Error Correction with Action Guided Sequence Generation

Jiquan Li, Junliang Guo, Yongxin Zhu +4

The task of Grammatical Error Correction (GEC) has received remarkable attention with wide applications in Natural Language Processing (NLP) in recent years. While one of the key p…

cs.CL20212 cited

Towards Variable-Length Textual Adversarial Attacks

Junliang Guo, Zhirui Zhang, Linlin Zhang +4

Adversarial attacks have shown the vulnerability of machine learning models, however, it is non-trivial to conduct textual adversarial attacks on natural language processing tasks…

cs.CL2020

Incorporating BERT into Parallel Sequence Decoding with Adapters

Junliang Guo, Zhirui Zhang, Linli Xu +3

While large scale pre-trained language models such as BERT have achieved great success on various natural language understanding tasks, how to efficiently and effectively incorpora…

cs.LG2020

STL-SGD: Speeding Up Local SGD with Stagewise Communication Period

Shuheng Shen, Yifei Cheng, Jingchang Liu +1

Distributed parallel stochastic gradient descent algorithms are workhorses for large scale machine learning tasks. Among them, local stochastic gradient descent (Local SGD) has att…

cs.LG20198 cited

Fine-Tuning by Curriculum Learning for Non-Autoregressive Neural Machine Translation

Junliang Guo, Xu Tan, Linli Xu +3

Non-autoregressive translation (NAT) models remove the dependence on previous target tokens and generate all target tokens in parallel, resulting in significant inference speedup b…