3 papers
cs.CL2026
VietMix: A Naturally-Occurring Parallel Corpus and Augmentation Framework for Vietnamese-English Code-Mixed Machine Translation
Hieu Tran, Phuong-Anh Nguyen-Le, Huy Nghiem +3
Machine translation (MT) systems universally degrade when faced with code-mixed text. This problem is more acute for low-resource languages that lack dedicated parallel corpora. Th…
cs.IR2025
Extracting Research Instruments from Educational Literature Using LLMs
Jiseung Yoo, Curran Mahowald, Meiyu Li +1
Large Language Models (LLMs) are transforming information extraction from academic literature, offering new possibilities for knowledge management. This study presents an LLM-based…
cs.LG2024
Large Language Models Meet Graph Neural Networks: A Perspective of Graph Mining
Yuxin You, Zhen Liu, Xiangchao Wen +2
Graph mining is an important area in data mining and machine learning that involves extracting valuable information from graph-structured data. In recent years, significant progres…