11 citations · 20 across the 3 of their papers we have counts for
3 papers
cs.CL2022★ 9 cited
MDIA: A Benchmark for Multilingual Dialogue Generation in 46 Languages
Qingyu Zhang, Xiaoyu Shen, Ernie Chang +2
Owing to the lack of corpora for low-resource languages, current works on dialogue generation have mainly focused on English. In this paper, we present mDIA, the first large-scale…
cs.SE2022★ 11 cited
SPT-Code: Sequence-to-Sequence Pre-Training for Learning Source Code Representations
Changan Niu, Chuanyi Li, Vincent Ng +3
Recent years have seen the successful application of large pre-trained models to code representation learning, resulting in substantial improvements on many code-related downstream…
cs.CL2021
AST-Transformer: Encoding Abstract Syntax Trees Efficiently for Code Summarization
Ze Tang, Chuanyi Li, Jidong Ge +3
Code summarization aims to generate brief natural language descriptions for source code. As source code is highly structured and follows strict programming language grammars, its A…