activity
20242026
most citedViWikiFC: Fact-Checking for Vietnamese Wikipedia-Based Textual Knowledge Source

1 citations · 1 across the 1 of their papers we have counts for

collaborators

5 papers

cs.CL20261 cited

ViWikiFC: Fact-Checking for Vietnamese Wikipedia-Based Textual Knowledge Source

Hung Tuan Le, Long Truong To, Manh Trong Nguyen +1

Fact-checking is essential due to the explosion of misinformation in the media ecosystem. Although false information exists in every language and country, most research to solve th…

cs.CL2025

ViSoLex: An Open-Source Repository for Vietnamese Social Media Lexical Normalization

Anh Thi-Hoang Nguyen, Dung Ha Nguyen, Kiet Van Nguyen

ViSoLex is an open-source system designed to address the unique challenges of lexical normalization for Vietnamese social media text. The platform provides two core services: Non-S…

cs.CL2024

Evaluating Large Language Model Capability in Vietnamese Fact-Checking Data Generation

Long Truong To, Hung Tuan Le, Dat Van-Thanh Nguyen +4

Large Language Models (LLMs), with gradually improving reading comprehension and reasoning capabilities, are being applied to a range of complex language tasks, including the autom…

cs.CL2024

A Weakly Supervised Data Labeling Framework for Machine Lexical Normalization in Vietnamese Social Media

Dung Ha Nguyen, Anh Thi Hoang Nguyen, Kiet Van Nguyen

This study introduces an innovative automatic labeling framework to address the challenges of lexical normalization in social media texts for low-resource languages like Vietnamese…

cs.CL2024

Automatic Textual Normalization for Hate Speech Detection

Anh Thi-Hoang Nguyen, Dung Ha Nguyen, Nguyet Thi Nguyen +2

Social media data is a valuable resource for research, yet it contains a wide range of non-standard words (NSW). These irregularities hinder the effective operation of NLP tools. C…