2 papers
cs.LG2026
StabilityBench: Benchmarking Instability in LLMs
Emma Kondrup, Zachary Yang, Anne Imouza +1
AI Assistants are increasingly deployed in high-stakes settings, such as healthcare or government services. Yet their real-world behavior remains poorly understood due to strong co…
cs.AI2025
Dr. Bias: Social Disparities in AI-Powered Medical Guidance
Emma Kondrup, Anne Imouza
With the rapid progress of Large Language Models (LLMs), the general public now has easy and affordable access to applications capable of answering most health-related questions in…