3 papers
cs.LG2026
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability
Yash Aggarwal, Atmika Gorti, Vinija Jain +3
Large language models (LLMs) are increasingly deployed in settings that require nuanced ethical reasoning, yet existing bias evaluations treat model outputs as simply "biased" or "…
cs.CL2025
Mental Health Equity in LLMs: Leveraging Multi-Hop Question Answering to Detect Amplified and Silenced Perspectives
Batool Haider, Atmika Gorti, Aman Chadha +1
Large Language Models (LLMs) in mental healthcare risk propagating biases that reinforce stigma and harm marginalized groups. While previous research identified concerning trends,…
cs.CL2024
Unboxing Occupational Bias: Grounded Debiasing of LLMs with U.S. Labor Data
Atmika Gorti, Manas Gaur, Aman Chadha
Large Language Models (LLMs) are prone to inheriting and amplifying societal biases embedded within their training data, potentially reinforcing harmful stereotypes related to gend…