3 papers
cs.HC2025
Examining the Expanding Role of Synthetic Data Throughout the AI Development Pipeline
Shivani Kapania, Stephanie Ballard, Alex Kessler +1
Alongside the growth of generative AI, we are witnessing a surge in the use of synthetic data across all stages of the AI development pipeline. It is now common practice for resear…
cs.HC2024
"I'm Not Sure, But...": Examining the Impact of Large Language Models' Uncertainty Expression on User Reliance and Trust
Sunnie S. Y. Kim, Q. Vera Liao, Mihaela Vorvoreanu +2
Widely deployed large language models (LLMs) can produce convincing yet incorrect outputs, potentially misleading users who may rely on them as if they were correct. To reduce such…
cs.LG2023
Open Datasheets: Machine-readable Documentation for Open Datasets and Responsible AI Assessments
Anthony Cintron Roman, Jennifer Wortman Vaughan, Valerie See +4
This paper introduces a no-code, machine-readable documentation framework for open datasets, with a focus on responsible AI (RAI) considerations. The framework aims to improve comp…