84 citations · 147 across the 41 of their papers we have counts for
45 papers
Soft Language Clustering for Multilingual Model Pre-training
Jiali Zeng, Yufan Jiang, Yongjing Yin +5
Multilingual pre-trained language models have demonstrated impressive (zero-shot) cross-lingual transfer abilities, however, their performance is hindered when the target language…
Findings of the WMT 2022 Shared Task on Translation Suggestion
Zhen Yang, Fandong Meng, Yingxue Zhang +2
We report the result of the first edition of the WMT shared task on Translation Suggestion (TS). The task aims to provide alternatives for specific words or phrases given the entir…
Token-Label Alignment for Vision Transformers
Han Xiao, Wenzhao Zheng, Zheng Zhu +2
Data mixing strategies (e.g., CutMix) have shown the ability to greatly improve the performance of convolutional neural networks (CNNs). They mix two images as inputs for training…
OPERA: Omni-Supervised Representation Learning with Hierarchical Supervisions
Chengkun Wang, Wenzhao Zheng, Zheng Zhu +2
The pretrain-finetune paradigm in modern computer vision facilitates the success of self-supervised learning, which tends to achieve better transferability than supervised learning…
Summer: WeChat Neural Machine Translation Systems for the WMT22 Biomedical Translation Task
Ernan Li, Fandong Meng, Jie Zhou
This paper introduces WeChat's participation in WMT 2022 shared biomedical translation task on Chinese to English. Our systems are based on the Transformer, and use several differe…
BJTU-WeChat's Systems for the WMT22 Chat Translation Task
Yunlong Liang, Fandong Meng, Jinan Xu +2
This paper introduces the joint submission of the Beijing Jiaotong University and WeChat AI to the WMT'22 chat translation task for English-German. Based on the Transformer, we app…