1 citations · 1 across the 3 of their papers we have counts for
3 papers
Statistical Hypothesis Testing for Auditing Robustness in Language Models
Paulius Rauba, Qiyao Wei, Mihaela van der Schaar
Consider the problem of testing whether the outputs of a large language model (LLM) system change under an arbitrary intervention, such as an input perturbation or changing the mod…
Quantifying perturbation impacts for large language models
Paulius Rauba, Qiyao Wei, Mihaela van der Schaar
We consider the problem of quantifying how an input perturbation impacts the outputs of large language models (LLMs), a fundamental task for model reliability and post-hoc interpre…
Defining Expertise: Applications to Treatment Effect Estimation
Alihan Hüyük, Qiyao Wei, Alicia Curth +1
Decision-makers are often experts of their domain and take actions based on their domain knowledge. Doctors, for instance, may prescribe treatments by predicting the likely outcome…