361 citations · 364 across the 4 of their papers we have counts for
4 papers
DELPHI: Data for Evaluating LLMs' Performance in Handling Controversial Issues
David Q. Sun, Artem Abzaliev, Hadas Kotek +3
Controversy is a reflection of our zeitgeist, and an important aspect to any discourse. The rise of large language models (LLMs) as conversational systems has increased public reli…
Gender bias and stereotypes in Large Language Models
Hadas Kotek, Rikker Dockum, David Q. Sun
Large Language Models (LLMs) have made substantial progress in the past several months, shattering state-of-the-art benchmarks in many domains. This paper investigates LLMs' behavi…
Intelligent Assistant Language Understanding On Device
Cecilia Aas, Hisham Abdelsalam, Irina Belousova +20
It has recently become feasible to run personal digital assistants on phones and other personal devices. In this paper we describe a design for a natural language understanding sys…
Feedback Effect in User Interaction with Intelligent Assistants: Delayed Engagement, Adaption and Drop-out
Zidi Xiu, Kai-Chen Cheng, David Q. Sun +7
With the growing popularity of intelligent assistants (IAs), evaluating IA quality becomes an increasingly active field of research. This paper identifies and quantifies the feedba…