3 papers
cs.CL2026
A Representation-Level Assessment of Bias Mitigation in Foundation Models
Svetoslav Nizhnichenkov, Rahul Nair, Elizabeth Daly +1
We investigate how successful bias mitigation reshapes the embedding space of encoder-only and decoder-only foundation models, offering an internal audit of model behaviour through…
cs.CL2025
Interpreting LLM-as-a-Judge Policies via Verifiable Global Explanations
Jasmina Gajcin, Erik Miehling, Rahul Nair +3
Using LLMs to evaluate text, that is, LLM-as-a-judge, is increasingly being used at scale to augment or even replace human annotations. As such, it is imperative that we understand…
cs.LG2025
Humble AI in the real-world: the case of algorithmic hiring
Rahul Nair, Inge Vejsbjerg, Elizabeth Daly +2
Humble AI (Knowles et al., 2023) argues for cautiousness in AI development and deployments through scepticism (accounting for limitations of statistical learning), curiosity (accou…