1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.CL2025★ 1 cited
Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective
Bhavik Chandna, Zubair Bashir, Procheta Sen
Large Language Models (LLMs) are known to exhibit social, demographic, and gender biases, often as a consequence of the data on which they are trained. In this work, we adopt a mec…
cs.IR2024
A Counterfactual Explanation Framework for Retrieval Models
Bhavik Chandna, Procheta Sen
Explainability has become a crucial concern in today's world, aiming to enhance transparency in machine learning and deep learning models. Information retrieval is no exception to…