2 papers
cs.CL2019
The FRENK Datasets of Socially Unacceptable Discourse in Slovene and English
Nikola Ljubešić, Darja Fišer, Tomaž Erjavec
In this paper we present datasets of Facebook comment threads to mainstream media posts in Slovene and English developed inside the Slovene national project FRENK which cover two t…
cs.CL2019
KAS-term: Extracting Slovene Terms from Doctoral Theses via Supervised Machine Learning
Nikola Ljubešić, Darja Fišer, Tomaž Erjavec
This paper presents a dataset and supervised learning experiments for term extraction from Slovene academic texts. Term candidates in the dataset were extracted via morphosyntactic…