3 citations · 3 across the 3 of their papers we have counts for
9 papers
Vector of Locally-Aggregated Word Embeddings (VLAWE): A Novel Document-level Representation
Radu Tudor Ionescu, Andrei M. Butnaru
In this paper, we propose a novel representation for text documents based on aggregating word embedding vectors into document embeddings. Our approach is inspired by the Vector of…
MOROCO: The Moldavian and Romanian Dialectal Corpus
Andrei M. Butnaru, Radu Tudor Ionescu
In this work, we introduce the MOldavian and ROmanian Dialectal COrpus (MOROCO), which is freely available for download at https://github.com/butnaruandrei/MOROCO. The corpus conta…
Transductive Learning with String Kernels for Cross-Domain Text Classification
Radu Tudor Ionescu, Andrei M. Butnaru
For many text classification tasks, there is a major problem posed by the lack of labeled data in a target domain. Although classifiers for a target domain can be trained on labele…
Improving the results of string kernels in sentiment analysis and Arabic dialect identification by adapting them to your test set
Radu Tudor Ionescu, Andrei M. Butnaru
Recently, string kernels have obtained state-of-the-art results in various text classification tasks such as Arabic dialect identification or native language identification. In thi…
UnibucKernel Reloaded: First Place in Arabic Dialect Identification for the Second Year in a Row
Andrei M. Butnaru, Radu Tudor Ionescu
We present a machine learning approach that ranked on the first place in the Arabic Dialect Identification (ADI) Closed Shared Tasks of the 2018 VarDial Evaluation Campaign. The pr…
Automated essay scoring with string kernels and word embeddings
Mădălina Cozma, Andrei M. Butnaru, Radu Tudor Ionescu
In this work, we present an approach based on combining string kernels and word embeddings for automatic essay scoring. String kernels capture the similarity among strings based on…