Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
Approaches to Analysing Historical Newspapers Using LLMs
Filip Dobranić, Tina Munda, Oliver Pejić +8
This study presents a computational analysis of the Slovene historical newspapers \textit{Slovenec} and \textit{Slovenski narod} from the sPeriodika corpus, combining topic modelli…
cs.CL2019
The FRENK Datasets of Socially Unacceptable Discourse in Slovene and English
Nikola Ljubešić, Darja Fišer, Tomaž Erjavec
In this paper we present datasets of Facebook comment threads to mainstream media posts in Slovene and English developed inside the Slovene national project FRENK which cover two t…
cs.CL2019
KAS-term: Extracting Slovene Terms from Doctoral Theses via Supervised Machine Learning
Nikola Ljubešić, Darja Fišer, Tomaž Erjavec
This paper presents a dataset and supervised learning experiments for term extraction from Slovene academic texts. Term candidates in the dataset were extracted via morphosyntactic…