7 papers
RoD-TAL: A Benchmark for Answering Questions in Romanian Driving License Exams
Andrei Vlad Man, RÄzvan-Alexandru SmÄdu, Cristian-George Craciun +3
The intersection of AI and legal systems presents a growing need for tools that support legal education, particularly in under-resourced languages such as Romanian. In this work, w…
Air Pollution Forecasting in Bucharest
DragoÅ-Andrei Åerban, RÄzvan-Alexandru SmÄdu, Dumitru-Clementin Cercel
Air pollution, especially the particulate matter 2.5 (PM2.5), has become a growing concern in recent years, primarily in urban areas. Being exposed to air pollution is linked to de…
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
RÄzvan-Alexandru SmÄdu, Andreea Iuga, Dumitru-Clementin Cercel +1
Satire, irony, and sarcasm are techniques typically used to express humor and critique, rather than deceive; however, they can occasionally be mistaken for factual reporting, akin…
GRAF: Graph Retrieval Augmented by Facts for Romanian Legal Multi-Choice Question Answering
Cristian-George CrÄciun, RÄzvan-Alexandru SmÄdu, Dumitru-Clementin Cercel +1
Pre-trained Language Models (PLMs) have shown remarkable performances in recent years, setting a new paradigm for NLP research and industry. The legal domain has received some atte…
MuSaRoNews: A Multidomain, Multimodal Satire Dataset from Romanian News Articles
RÄzvan-Alexandru SmÄdu, Andreea Iuga, Dumitru-Clementin Cercel
Satire and fake news can both contribute to the spread of false information, even though both have different purposes (one if for amusement, the other is to misinform). However, it…
Investigating Large Language Models for Complex Word Identification in Multilingual and Multidomain Setups
RÄzvan-Alexandru SmÄdu, David-Gabriel Ion, Dumitru-Clementin Cercel +2
Complex Word Identification (CWI) is an essential step in the lexical simplification task and has recently become a task on its own. Some variations of this binary classification t…