activity
20182022
most citedDAPPLE: A Pipelined Data Parallel Approach for Training Large Models

29 citations · 83 across the 13 of their papers we have counts for

collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL2021

M6: A Chinese Multimodal Pretrainer

Junyang Lin, Rui Men, An Yang +22

In this work, we construct the largest dataset for multimodal pretraining in Chinese, which consists of over 1.9TB images and 292GB texts that cover a wide range of domains. We pro…

cs.CL2020

EasyTransfer -- A Simple and Scalable Deep Transfer Learning Platform for NLP Applications

Minghui Qiu, Peng Li, Chengyu Wang +8

The literature has witnessed the success of leveraging Pre-trained Language Models (PLMs) and Transfer Learning (TL) algorithms to a wide range of Natural Language Processing (NLP)…

cs.CL2020

AdaBERT: Task-Adaptive BERT Compression with Differentiable Neural Architecture Search

Daoyuan Chen, Yaliang Li, Minghui Qiu +7

Large pre-trained language models such as BERT have shown their effectiveness in various natural language processing tasks. However, the huge parameter size makes them difficult to…

cs.CL20193 cited

RPM-Oriented Query Rewriting Framework for E-commerce Keyword-Based Sponsored Search

Xiuying Chen, Daorui Xiao, Shen Gao +5

Sponsored search optimizes revenue and relevance, which is estimated by Revenue Per Mille (RPM). Existing sponsored search models are all based on traditional statistical models, w…

cs.CL2018

Transfer Learning for Context-Aware Question Matching in Information-seeking Conversations in E-commerce

Minghui Qiu, Liu Yang, Feng Ji +6

Building multi-turn information-seeking conversation systems is an important and challenging research topic. Although several advanced neural text matching models have been propose…