activity
20192022
most citedUnsung Challenges of Building and Deploying Language Technologies for Low Resource Language Communities

14 citations · 26 across the 5 of their papers we have counts for

collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL20221 cited

TyDiP: A Dataset for Politeness Classification in Nine Typologically Diverse Languages

Anirudh Srinivasan, Eunsol Choi

We study politeness phenomena in nine typologically diverse languages. Politeness is an important facet of communication and is sometimes argued to be cultural-specific, yet existi…

cs.CL202210 cited

CALCS 2021 Shared Task: Machine Translation for Code-Switched Data

Shuguang Chen, Gustavo Aguilar, Anirudh Srinivasan +2

To date, efforts in the code-switching literature have focused for the most part on language identification, POS, NER, and syntactic parsing. In this paper, we address machine tran…

cs.CL2021

Predicting the Performance of Multilingual NLP Models

Anirudh Srinivasan, Sunayana Sitaram, Tanuja Ganu +3

Recent advancements in NLP have given us models like mBERT and XLMR that can serve over 100 languages. The languages that these models are evaluated on, however, are very few in nu…

cs.CL2020

GLUECoS : An Evaluation Benchmark for Code-Switched NLP

Simran Khanuja, Sandipan Dandapat, Anirudh Srinivasan +2

Code-switching is the use of more than one language in the same conversation or utterance. Recently, multilingual contextual embedding models, trained on multiple monolingual corpo…

cs.CL201914 cited

Unsung Challenges of Building and Deploying Language Technologies for Low Resource Language Communities

Pratik Joshi, Christain Barnes, Sebastin Santy +7

In this paper, we examine and analyze the challenges associated with developing and introducing language technologies to low-resource language communities. While doing so, we bring…