activity
20162023
most citedBaichuan 2: Open Large-scale Language Models

127 citations · 300 across the 37 of their papers we have counts for

collaborators
Showing 2022Show all

5 papers · 1 filter

cs.CL2022

Solution of DeBERTaV3 on CommonsenseQA

Letian Peng, Zuchao Li, Hai Zhao

We report the performance of DeBERTaV3 on CommonsenseQA in this report. We simply formalize the answer selection as a text classification for DeBERTaV3. The strong natural language…

cs.CL2022★ 1 cited

Adversarial Self-Attention for Language Understanding

Hongqiu Wu, Ruixue Ding, Hai Zhao +3

Deep neural models (e.g. Transformer) naturally learn spurious features, which create a ``shortcut'' between the labels and inputs, thus impairing the generalization and robustness…

cs.CL2022★ 1 cited

Back to the Future: Bidirectional Information Decoupling Network for Multi-turn Dialogue Modeling

Yiyang Li, Hai Zhao, Zhuosheng Zhang

Multi-turn dialogue modeling as a challenging branch of natural language understanding (NLU), aims to build representations for machines to understand human dialogues, which provid…

cs.CL2022

Lite Unified Modeling for Discriminative Reading Comprehension

Yilin Zhao, Hai Zhao, Libin Shen +1

As a broad and major category in machine reading comprehension (MRC), the generalized goal of discriminative MRC is answer prediction from the given materials. However, the focuses…

cs.LG2022

Distinguishing Non-natural from Natural Adversarial Samples for More Robust Pre-trained Language Model

Jiayi Wang, Rongzhou Bao, Zhuosheng Zhang +1

Recently, the problem of robustness of pre-trained language models (PrLMs) has received increasing research interest. Latest studies on adversarial attacks achieve high attack succ…