2 papers
cs.CL2026
Alignment Reduces Expressed but Not Encoded Gender Bias: A Unified Framework and Study
Nour Bouchouchi, Thibault Laugel, Xavier Renard +3
During training, Large Language Models (LLMs) learn social regularities that can lead to gender bias in downstream applications. Most mitigation efforts focus on reducing bias in g…
cs.AI2025
Metric assessment protocol in the context of answer fluctuation on MCQ tasks
Ekaterina Goliakova, Xavier Renard, Marie-Jeanne Lesot +3
Using multiple-choice questions (MCQs) has become a standard for assessing LLM capabilities efficiently. A variety of metrics can be employed for this task. However, previous resea…