1 citations · 2 across the 6 of their papers we have counts for
3 papers · 1 filter
Winter Soldier: Backdooring Language Models at Pre-Training with Indirect Data Poisoning
Wassim Bouaziz, Mathurin Videau, Nicolas Usunier +1
The pre-training of large language models (LLMs) relies on massive text datasets sourced from diverse and difficult-to-curate origins. Although membership inference attacks and hid…
On Monotonicity in AI Alignment
Gilles Bareilles, Julien Fageot, Lê-Nguyên Hoang +4
Comparison-based preference learning has become central to the alignment of AI models with human preferences. However, these methods may behave counterintuitively. After empiricall…
Targeted Data Poisoning for Black-Box Audio Datasets Ownership Verification
Wassim Bouaziz, El-Mahdi El-Mhamdi, Nicolas Usunier
Protecting the use of audio datasets is a major concern for data owners, particularly with the recent rise of audio deep learning models. While watermarks can be used to protect th…