Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Visualizing token importance for black-box language models
Paulius Rauba, Qiyao Wei, Mihaela van der Schaar
We consider the problem of auditing black-box large language models (LLMs) to ensure they behave reliably when deployed in production settings, particularly in high-stakes domains…
cs.CL2025
Statistical Hypothesis Testing for Auditing Robustness in Language Models
Paulius Rauba, Qiyao Wei, Mihaela van der Schaar
Consider the problem of testing whether the outputs of a large language model (LLM) system change under an arbitrary intervention, such as an input perturbation or changing the mod…