1 paper · 1 filter
Taja Kuzman, Peter Rupnik, Nikola Ljubešić
This paper presents a new training dataset for automatic genre identification GINCO, which is based on 1,125 crawled Slovenian web documents that consist of 650 thousand words. Eac…