37 citations · 44 across the 3 of their papers we have counts for
3 papers
cs.CR2023★ 4 cited
East: Efficient and Accurate Secure Transformer Framework for Inference
Yuanchao Ding, Hua Guo, Yewei Guan +4
Transformer has been successfully used in practical applications, such as ChatGPT, due to its powerful advantages. However, users' input is leaked to the model provider during the…
cs.CL2023★ 3 cited
ERNIE 3.0 Tiny: Frustratingly Simple Method to Improve Task-Agnostic Distillation Generalization
Weixin Liu, Xuyi Chen, Jiaxiang Liu +4
Task-agnostic knowledge distillation attempts to address the problem of deploying large pretrained language model in resource-constrained scenarios by compressing a large pretraine…
cs.CL2021★ 37 cited
ERNIE 3.0 Titan: Exploring Larger-scale Knowledge Enhanced Pre-training for Language Understanding and Generation
Shuohuan Wang, Yu Sun, Yang Xiang +26
Pre-trained language models have achieved state-of-the-art results in various Natural Language Processing (NLP) tasks. GPT-3 has shown that scaling up pre-trained language models c…