3 papers
cs.CL2024
Can General-Purpose Large Language Models Generalize to English-Thai Machine Translation ?
Jirat Chiaranaipanich, Naiyarat Hanmatheekuna, Jitkapat Sawatphol +6
Large language models (LLMs) perform well on common tasks but struggle with generalization in low-resource and low-computation settings. We examine this limitation by testing vario…
cs.CL2024
Addressing Topic Leakage in Cross-Topic Evaluation for Authorship Verification
Jitkapat Sawatphol, Can Udomcharoenchaikit, Sarana Nutanong
Authorship verification (AV) aims to identify whether a pair of texts has the same author. We address the challenge of evaluating AV models' robustness against topic shifts. The co…
cs.CL2024
ThaiCoref: Thai Coreference Resolution Dataset
Pontakorn Trakuekul, Wei Qi Leong, Charin Polpanumas +3
While coreference resolution is a well-established research area in Natural Language Processing (NLP), research focusing on Thai language remains limited due to the lack of large a…