5 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CL2023★ 1 cited
Bhasha-Abhijnaanam: Native-script and romanized Language Identification for 22 Indic languages
Yash Madhani, Mitesh M. Khapra, Anoop Kunchukuttan
We create publicly available language identification (LID) datasets and models in all 22 Indian languages listed in the Indian constitution in both native-script and romanized text…
cs.CL2022★ 5 cited
Aksharantar: Open Indic-language Transliteration datasets and models for the Next Billion Users
Yash Madhani, Sushane Parthan, Priyanka Bedekar +5
Transliteration is very important in the Indian language context due to the usage of multiple scripts and the widespread use of romanized inputs. However, few training and evaluati…