7 citations · 8 across the 3 of their papers we have counts for
3 papers
cs.CL2022
SPARTAN: Sparse Hierarchical Memory for Parameter-Efficient Transformers
Ameet Deshpande, Md Arafat Sultan, Anthony Ferritto +3
Fine-tuning pre-trained language models (PLMs) achieves impressive performance on a range of downstream tasks, and their sizes have consequently been getting bigger. Since a differ…
cs.CL2022★ 1 cited
ALIGN-MLM: Word Embedding Alignment is Crucial for Multilingual Pre-training
Henry Tang, Ameet Deshpande, Karthik Narasimhan
Multilingual pre-trained models exhibit zero-shot cross-lingual transfer, where a model fine-tuned on a source language achieves surprisingly good performance on a target language.…
cs.CL2022★ 7 cited
Linking Emergent and Natural Languages via Corpus Transfer
Shunyu Yao, Mo Yu, Yang Zhang +3
The study of language emergence aims to understand how human languages are shaped by perceptual grounding and communicative intent. Computational approaches to emergent communicati…