2 citations · 2 across the 1 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Pula: Training Large Language Models for Setswana
Nathan Brown, Vukosi Marivate
In this work we present Pula, a suite of bilingual language models proficient in both Setswana and English. Leveraging recent advancements in data availability and efficient fine-t…
cs.CL2023★ 2 cited
Efficient Transformer Knowledge Distillation: A Performance Review
Nathan Brown, Ashton Williamson, Tahj Anderson +1
As pretrained transformer language models continue to achieve state-of-the-art performance, the Natural Language Processing community has pushed for advances in model compression a…