50 citations · 65 across the 18 of their papers we have counts for
3 papers · 1 filter
Adapting the Tesseract Open-Source OCR Engine for Tamil and Sinhala Legacy Fonts and Creating a Parallel Corpus for Tamil-Sinhala-English
Charangan Vasantharajan, Laksika Tharmalingam, Uthayasanker Thayasivam
Most low-resource languages do not have the necessary resources to create even a substantial monolingual corpus. These languages may often be found in government proceedings but ma…
Towards Offensive Language Identification for Tamil Code-Mixed YouTube Comments and Posts
Charangan Vasantharajan, Uthayasanker Thayasivam
Offensive Language detection in social media platforms has been an active field of research over the past years. In non-native English spoken countries, social media users mostly u…
Data-Driven Simulation of Ride-Hailing Services using Imitation and Reinforcement Learning
Haritha Jayasinghe, Tarindu Jayatilaka, Ravin Gunawardena +1
The rapid growth of ride-hailing platforms has created a highly competitive market where businesses struggle to make profits, demanding the need for better operational strategies.…