papers

Publications (31)

cs.CL2021

skweak: Weak Supervision Made Easy for NLP

Pierre Lison, Jeremy Barnes, Aliaksandr Hubin

We present skweak, a versatile, Python-based software toolkit enabling NLP developers to apply weak supervision to a wide range of NLP tasks. Weak supervision is an emerging machin…

cs.CL2024

English Prompts are Better for NLI-based Zero-Shot Emotion Classification than Target-Language Prompts

Patrick Bareiß, Roman Klinger, Jeremy Barnes

Emotion classification in text is a challenging task due to the processes involved when interpreting a textual description of a potential emotion stimulus. In addition, the set of…

cs.CL2020

A Fine-Grained Sentiment Dataset for Norwegian

Lilja Øvrelid, Petter Mæhlum, Jeremy Barnes +1

We introduce NoReC_fine, a dataset for fine-grained sentiment analysis in Norwegian, annotated with respect to polar expressions, targets and holders of opinion. The underlying tex…

cs.CL2025

HiTZ at VarDial 2025 NorSID: Overcoming Data Scarcity with Language Transfer and Automatic Data Annotation

Jaione Bengoetxea, Mikel Zubillaga, Ekhi Azurmendi +4

In this paper we present our submission for the NorSID Shared Task as part of the 2025 VarDial Workshop (Scherrer et al., 2025), consisting of three tasks: Intent Detection, Slot F…

cs.CL2018

MultiBooked: A Corpus of Basque and Catalan Hotel Reviews Annotated for Aspect-level Sentiment Classification

Jeremy Barnes, Patrik Lambert, Toni Badia

While sentiment analysis has become an established field in the NLP community, research into languages other than English has been hindered by the lack of resources. Although much…

cs.CL2024

XNLIeu: a dataset for cross-lingual NLI in Basque

Maite Heredia, Julen Etxaniz, Muitze Zulaika +3

XNLI is a popular Natural Language Inference (NLI) benchmark widely used to evaluate cross-lingual Natural Language Understanding (NLU) capabilities across languages. In this paper…