5 citations · 7 across the 3 of their papers we have counts for
3 papers
Selective Pre-training for Private Fine-tuning
Da Yu, Sivakanth Gopi, Janardhan Kulkarni +5
Text prediction models, when used in applications like email clients or word processors, must protect user data privacy and adhere to model size constraints. These constraints are…
Planting and Mitigating Memorized Content in Predictive-Text Language Models
C. M. Downey, Wei Dai, Huseyin A. Inan +3
Language models are widely deployed to provide automatic text completion services in user products. However, recent research has revealed that language models (especially large one…
LiteTransformerSearch: Training-free Neural Architecture Search for Efficient Language Models
Mojan Javaheripi, Gustavo H. de Rosa, Subhabrata Mukherjee +6
The Transformer architecture is ubiquitously used as the building block of large-scale autoregressive language models. However, finding architectures with the optimal trade-off bet…