activity
20202022
most citedImproving Sentiment Analysis By Emotion Lexicon Approach on Vietnamese Texts

8 citations · 9 across the 2 of their papers we have counts for

collaborators

10 papers

cs.CL20228 cited

Improving Sentiment Analysis By Emotion Lexicon Approach on Vietnamese Texts

An Long Doan, Son T. Luu

The sentiment analysis task has various applications in practice. In the sentiment analysis task, words and phrases that represent positive and negative emotions are important. Fin…

cs.CL20221 cited

UIT-ViCoV19QA: A Dataset for COVID-19 Community-based Question Answering on Vietnamese Language

Triet Minh Thai, Ngan Ha-Thao Chu, Anh Tuan Vo +1

For the last two years, from 2020 to 2021, COVID-19 has broken disease prevention measures in many countries, including Vietnam, and negatively impacted various aspects of human li…

cs.CL2021

Conversational Machine Reading Comprehension for Vietnamese Healthcare Texts

Son T. Luu, Mao Nguyen Bui, Loi Duc Nguyen +3

Machine reading comprehension (MRC) is a sub-field in natural language processing that aims to assist computers understand unstructured texts and then answer questions related to t…

cs.CL2021

UIT-ISE-NLP at SemEval-2021 Task 5: Toxic Spans Detection with BiLSTM-CRF and ToxicBERT Comment Classification

Son T. Luu, Ngan Luu-Thuy Nguyen

We present our works on SemEval-2021 Task 5 about Toxic Spans Detection. This task aims to build a model for identifying toxic words in whole posts. We use the BiLSTM-CRF model com…

cs.CL2021

A Large-scale Dataset for Hate Speech Detection on Vietnamese Social Media Texts

Son T. Luu, Kiet Van Nguyen, Ngan Luu-Thuy Nguyen

In recent years, Vietnam witnesses the mass development of social network users on different social platforms such as Facebook, Youtube, Instagram, and Tiktok. On social medias, ha…

cs.CL2020

Empirical Study of Text Augmentation on Social Media Text in Vietnamese

Son T. Luu, Kiet Van Nguyen, Ngan Luu-Thuy Nguyen

In the text classification problem, the imbalance of labels in datasets affect the performance of the text-classification models. Practically, the data about user comments on socia…