5 papers
SaRoCo: Detecting Satire in a Novel Romanian Corpus of News Articles
Ana-Cristina Rogoz, Mihaela Gaman, Radu Tudor Ionescu
In this work, we introduce a corpus for satire detection in Romanian news. We gathered 55,608 public news articles from multiple real and satirical news sources, composing one of t…
UnibucKernel: Geolocating Swiss German Jodels Using Ensemble Learning
Mihaela Gaman, Sebastian Cojocariu, Radu Tudor Ionescu
In this work, we describe our approach addressing the Social Media Variety Geolocation task featured in the 2021 VarDial Evaluation Campaign. We focus on the second subtask, which…
Clustering Word Embeddings with Self-Organizing Maps. Application on LaRoSeDa -- A Large Romanian Sentiment Data Set
Anca Maria Tache, Mihaela Gaman, Radu Tudor Ionescu
Romanian is one of the understudied languages in computational linguistics, with few resources available for the development of natural language processing tools. In this paper, we…
Combining Deep Learning and String Kernels for the Localization of Swiss German Tweets
Mihaela Gaman, Radu Tudor Ionescu
In this work, we introduce the methods proposed by the UnibucKernel team in solving the Social Media Variety Geolocation task featured in the 2020 VarDial Evaluation Campaign. We a…
Automatically Identifying Complaints in Social Media
Daniel Preotiuc-Pietro, Mihaela Gaman, Nikolaos Aletras
Complaining is a basic speech act regularly used in human and computer mediated communication to express a negative mismatch between reality and expectations in a particular situat…