37 citations · 42 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 2 cited
Rethinking the Metric in Few-shot Learning: From an Adaptive Multi-Distance Perspective
Jinxiang Lai, Siqian Yang, Guannan Jiang +9
Few-shot learning problem focuses on recognizing unseen classes given a few labeled images. In recent effort, more attention is paid to fine-grained feature embedding, ignoring the…
cs.DC2022★ 3 cited
Large-scale Knowledge Distillation with Elastic Heterogeneous Computing Resources
Ji Liu, Daxiang Dong, Xi Wang +5
Although more layers and more parameters generally improve the accuracy of the models, such big models generally have high computational complexity and require big memory, which ex…
cs.CL2021★ 37 cited
ERNIE 3.0 Titan: Exploring Larger-scale Knowledge Enhanced Pre-training for Language Understanding and Generation
Shuohuan Wang, Yu Sun, Yang Xiang +26
Pre-trained language models have achieved state-of-the-art results in various Natural Language Processing (NLP) tasks. GPT-3 has shown that scaling up pre-trained language models c…