45 citations · 124 across the 5 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2021
LV-BERT: Exploiting Layer Variety for BERT
Weihao Yu, Zihang Jiang, Fei Chen +2
Modern pre-trained language models are mostly built upon backbones stacking self-attention and feed-forward layers in an interleaved order. In this paper, beyond this stereotyped l…
cs.CL2020
ConvBERT: Improving BERT with Span-based Dynamic Convolution
Zihang Jiang, Weihao Yu, Daquan Zhou +3
Pre-trained language models like BERT and its variants have recently achieved impressive performance in various natural language understanding tasks. However, BERT heavily relies o…
cs.CL2020
ReClor: A Reading Comprehension Dataset Requiring Logical Reasoning
Weihao Yu, Zihang Jiang, Yanfei Dong +1
Recent powerful pre-trained language models have achieved remarkable performance on most of the popular datasets for reading comprehension. It is time to introduce more challenging…