80 citations · 90 across the 4 of their papers we have counts for
4 papers
WangchanThaiInstruct: An instruction-following Dataset for Culture-Aware, Multitask, and Multi-domain Evaluation in Thai
Peerat Limkonchotiwat, Pume Tuchinda, Lalita Lowphansirikul +5
Large language models excel at instruction-following in English, but their performance in low-resource languages like Thai remains underexplored. Existing benchmarks often rely on…
PyThaiNLP: Thai Natural Language Processing in Python
Wannaphong Phatthiyaphaibun, Korakot Chaovavanich, Charin Polpanumas +6
We present PyThaiNLP, a free and open-source natural language processing (NLP) library for Thai language implemented in Python. It provides a wide range of software, models, and da…
An Efficient Self-Supervised Cross-View Training For Sentence Embedding
Peerat Limkonchotiwat, Wuttikorn Ponwitayarat, Lalita Lowphansirikul +3
Self-supervised sentence representation learning is the task of constructing an embedding space for sentences without relying on human annotation efforts. One straightforward appro…
WangchanBERTa: Pretraining transformer-based Thai Language Models
Lalita Lowphansirikul, Charin Polpanumas, Nawat Jantrakulchai +1
Transformer-based language models, more specifically BERT-based architectures have achieved state-of-the-art performance in many downstream tasks. However, for a relatively low-res…