activity
20192022
most citedData Augmentation and Terminology Integration for Domain-Specific Sinhala-English-Tamil Statistical Machine Translation

13 citations · 44 across the 7 of their papers we have counts for

collaborators
Showing cs.CLShow all

8 papers · 1 filter

cs.CL20225 cited

Some Languages are More Equal than Others: Probing Deeper into the Linguistic Disparity in the NLP World

Surangika Ranathunga, Nisansa de Silva

Linguistic disparity in the NLP world is a problem that has been widely acknowledged recently. However, different facets of this problem, or the reasons behind this disparity are s…

cs.CL20226 cited

Data Augmentation to Address Out-of-Vocabulary Problem in Low-Resource Sinhala-English Neural Machine Translation

Aloka Fernando, Surangika Ranathunga

Out-of-Vocabulary (OOV) is a problem for Neural Machine Translation (NMT). OOV refers to words with a low occurrence in the training data, or to those that are absent from the trai…

cs.CL2022

Pre-Trained Multilingual Sequence-to-Sequence Models: A Hope for Low-Resource Language Translation?

En-Shiun Annie Lee, Sarubi Thillainathan, Shravan Nayak +4

What can pre-trained multilingual sequence-to-sequence models like mBART contribute to translating low-resource languages? We conduct a thorough empirical experiment in 10 language…

cs.CL20218 cited

Neural Machine Translation for Low-Resource Languages: A Survey

Surangika Ranathunga, En-Shiun Annie Lee, Marjana Prifti Skenduli +3

Neural Machine Translation (NMT) has seen a tremendous spurt of growth in less than ten years, and has already entered a mature phase. While considered as the most widely used solu…

cs.CL20211 cited

Exploiting Parallel Corpora to Improve Multilingual Embedding based Document and Sentence Alignment

Dilan Sachintha, Lakmali Piyarathna, Charith Rajitha +1

Multilingual sentence representations pose a great advantage for low-resource languages that do not have enough data to build monolingual models on their own. These multilingual se…

cs.CL202011 cited

Sentiment Analysis for Sinhala Language using Deep Learning Techniques

Lahiru Senevirathne, Piyumal Demotte, Binod Karunanayake +2

Due to the high impact of the fast-evolving fields of machine learning and deep learning, Natural Language Processing (NLP) tasks have further obtained comprehensive performances f…