3 papers
cs.CL2025
Beyond Vanilla Fine-Tuning: Leveraging Multistage, Multilingual, and Domain-Specific Methods for Low-Resource Machine Translation
Sarubi Thillainathan, Songchen Yuan, En-Shiun Annie Lee +2
Fine-tuning multilingual sequence-to-sequence large language models (msLLMs) has shown promise in developing neural machine translation (NMT) systems for low-resource languages (LR…
cs.CL2022
BERTifying Sinhala -- A Comprehensive Analysis of Pre-trained Language Models for Sinhala Text Classification
Vinura Dhananjaya, Piyumal Demotte, Surangika Ranathunga +1
This research provides the first comprehensive analysis of the performance of pre-trained language models for Sinhala text classification. We test on a set of different Sinhala tex…
cs.LG2021
Neural Mixture Models with Expectation-Maximization for End-to-end Deep Clustering
Dumindu Tissera, Kasun Vithanage, Rukshan Wijesinghe +4
Any clustering algorithm must synchronously learn to model the clusters and allocate data to those clusters in the absence of labels. Mixture model-based methods model clusters wit…