7 citations · 11 across the 12 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Adversarial Moment-Matching Distillation of Large Language Models
Chen Jia
Knowledge distillation (KD) has been shown to be highly effective in guiding a student model with a larger teacher model and achieving practical benefits in improving the computati…
cs.CL2024★ 1 cited
BLADE: Enhancing Black-box Large Language Models with Small Domain-Specific Models
Haitao Li, Qingyao Ai, Jia Chen +5
Large Language Models (LLMs) like ChatGPT and GPT-4 are versatile and capable of addressing a diverse range of tasks. However, general LLMs, which are developed on open-domain data…