From the 1 of 8 linked papers with an AI index.
8 papers
Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems
Marylou Fauchard, Florian Carichon, Margarida Carvalho +1
The paper studies how misaligned objectives in large‑language‑model powered multi‑agent systems affect performance in the social deduction game Werewolf, showing that hidden object…
IDEAFix: Evaluation Framework for Creative Defixation Prompting in LLMs
F. Carichon, S. Sharma, M. Girard +2
Large language models (LLMs) are increasingly used for tasks involving creative problem solving and idea generation. However, there is a lack of consensus concerning their creative…
Mechanics of Bias and Reasoning: Interpreting the Impact of Chain-of-Thought Prompting on Gender Bias in LLMs
Edie Pearman, Sophia Osborne, Mira Kandlikar-Bloch +3
Large language models (LLMs) are increasingly deployed in socially sensitive settings despite substantial documentation that they encode gender biases. Chain-of-Thought (CoT) promp…
Can LLMs Cook Jamaican Couscous? A Study of Cultural Novelty in Recipe Generation
F. Carichon, R. Rampa, G. Farnadi
Large Language Models (LLMs) are increasingly used to generate and shape cultural content, ranging from narrative writing to artistic production. While these models demonstrate imp…
Let's Unlearn Stereotypes Before Decision-Making: Assessing the Impact of Intrinsic Bias Mitigation on Downstream Fairness in LLMs
Mina Arzaghi, 'Mina Arzaghi', Alireza Dehghanpour Farashah +6
Large Language Models (LLMs) are increasingly used in high-stakes decision-making systems, where biased predictions can reinforce social and economic disparities. Although prior wo…
Reasoning with Preference Constraints: A Benchmark for Language Models in Many-to-One Matching Markets
Marylou Fauchard, Florian Carichon, Margarida Carvalho +1
Recent advances in reasoning with large language models (LLMs) have demonstrated strong performance on complex mathematical tasks, including combinatorial optimization. Techniques…