12 citations · 35 across the 17 of their papers we have counts for
3 papers · 1 filter
Optimizing Large Language Models for Turkish: New Methodologies in Corpus Selection and Training
H. Toprak Kesgin, M. Kaan Yuce, Eren Dogan +7
In this study, we develop and assess new corpus selection and training methodologies to improve the effectiveness of Turkish language models. Specifically, we adapted Large Languag…
Scaling BERT Models for Turkish Automatic Punctuation and Capitalization Correction
Abdulkader Saoud, Mahmut Alomeyr, Himmet Toprak Kesgin +1
This paper investigates the effectiveness of BERT based models for automated punctuation and capitalization corrections in Turkish texts across five distinct model sizes. The model…
Transformers as Neural Augmentors: Class Conditional Sentence Generation via Variational Bayes
M. Şafak Bilici, Mehmet Fatih Amasyali
Data augmentation methods for Natural Language Processing tasks are explored in recent years, however they are limited and it is hard to capture the diversity on sentence level. Be…