2 papers
cs.CL2025
Frontier LLMs Still Struggle with Simple Reasoning Tasks
Alan Malek, Jiawei Ge, Nevena Lazic +3
While state-of-the-art large language models (LLMs) demonstrate advanced reasoning capabilities-achieving remarkable performance on challenging competitive math and coding benchmar…
cs.LG2024
Mind the Graph When Balancing Data for Fairness or Robustness
Jessica Schrouff, Alexis Bellot, Amal Rannen-Triki +5
Failures of fairness or robustness in machine learning predictive settings can be due to undesired dependencies between covariates, outcomes and auxiliary factors of variation. A c…