193 citations
- Université de LilleFR199 papers
- Centre National de la Recherche ScientifiqueFR190 papers
- Centre de Recherche en Informatique, Signal et Automatique de LilleFR140 papers
- Université Paris-SaclayFR62 papers
- Centre de Recherche en InformatiqueFR57 papers
- Institut d'Astrophysique SpatialeFR56 papers
- Sorbonne UniversitéFR56 papers
- Université Paris CitéFR55 papers
- Commissariat à l'Énergie Atomique et aux Énergies AlternativesFR52 papers
- Université Grenoble AlpesFR51 papers
- CEA Paris-SaclayFR50 papers
- Institut de Recherche en Astrophysique et PlanétologieFR50 papers
6 papers · 1 filter
Reasoning Core: A Scalable RL Environment for LLM Symbolic Reasoning
Valentin Lacombe, Valentin Quesnel, Damien Sileo
We introduce Reasoning Core, a new scalable environment for Reinforcement Learning with Verifiable Rewards (RLVR), designed to advance foundational symbolic reasoning in Large Lang…
The Fair Game: Auditing & Debiasing AI Algorithms Over Time
Debabrota Basu, Udvas Das
An emerging field of AI, namely Fair Machine Learning (ML), aims to quantify different types of bias (also known as unfairness) exhibited in the predictions of ML algorithms, and t…
Bridging the Data Provenance Gap Across Text, Speech and Video
Shayne Longpre, Nikhil Singh, Manuel Cherep +40
Progress in AI is driven largely by the scale and quality of training data. Despite this, there is a deficit of empirical analysis examining the attributes of well-established data…
Applying Ising Machines to Multi-objective QUBOs
Mayowa Ayodele, Richard Allmendinger, Manuel López-Ibáñez +2
Multi-objective optimisation problems involve finding solutions with varying trade-offs between multiple and often conflicting objectives. Ising machines are physical devices that…
gym-DSSAT: a crop model turned into a Reinforcement Learning environment
Romain Gautron, Emilio J. Padrón, Philippe Preux +3
Addressing a real world sequential decision problem with Reinforcement Learning (RL) usually starts with the use of a simulated environment that mimics real conditions. We present…
Indexed Minimum Empirical Divergence for Unimodal Bandits
Hassan Saber, Pierre Ménard, Odalric-Ambrym Maillard
We consider a multi-armed bandit problem specified by a set of one-dimensional family exponential distributions endowed with a unimodal structure. We introduce IMED-UB, a algorithm…