21 citations · 39 across the 8 of their papers we have counts for
8 papers
EMMA-X: An EM-like Multilingual Pre-training Algorithm for Cross-lingual Representation Learning
Ping Guo, Xiangpeng Wei, Yue Hu +4
Expressing universal semantics common to all languages is helpful in understanding the meanings of complex and culture-specific sentences. The research theme underlying this scenar…
PolyLM: An Open Source Polyglot Large Language Model
Xiangpeng Wei, Haoran Wei, Huan Lin +15
Large language models (LLMs) demonstrate remarkable ability to comprehend, reason, and generate following nature language instructions. However, the development of LLMs has been pr…
Bridging the Domain Gaps in Context Representations for k-Nearest Neighbor Neural Machine Translation
Zhiwei Cao, Baosong Yang, Huan Lin +6
-Nearest neighbor machine translation (NN-MT) has attracted increasing attention due to its ability to non-parametrically adapt to new translation domains. By using an upstre…
From Statistical Methods to Deep Learning, Automatic Keyphrase Prediction: A Survey
Binbin Xie, Jia Song, Liangying Shao +6
Keyphrase prediction aims to generate phrases (keyphrases) that highly summarizes a given document. Recently, researchers have conducted in-depth studies on this task from various…
Towards Fine-Grained Information: Identifying the Type and Location of Translation Errors
Keqin Bao, Yu Wan, Dayiheng Liu +5
Fine-grained information on translation errors is helpful for the translation evaluation community. Existing approaches can not synchronously consider error position and type, fail…
Draft, Command, and Edit: Controllable Text Editing in E-Commerce
Kexin Yang, Dayiheng Liu, Wenqiang Lei +3
Product description generation is a challenging and under-explored task. Most such work takes a set of product attributes as inputs then generates a description from scratch in a s…