3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CL2024
Predictions from language models for multiple-choice tasks are not robust under variation of scoring methods
Polina Tsvilodub, Hening Wang, Sharon Grosch +1
This paper systematically compares different methods of deriving item-level predictions of language models for multiple-choice tasks. It compares scoring methods for answer options…
cs.CL2023
Evaluating Pragmatic Abilities of Image Captioners on A3DS
Polina Tsvilodub, Michael Franke
Evaluating grounded neural language model performance with respect to pragmatic qualities like the trade off between truthfulness, contrastivity and overinformativity of generated…
cs.CL2023★ 3 cited
Overinformative Question Answering by Humans and Machines
Polina Tsvilodub, Michael Franke, Robert D. Hawkins +1
When faced with a polar question, speakers often provide overinformative answers going beyond a simple "yes" or "no". But what principles guide the selection of additional informat…