Take It, Leave It, or Fix It: Measuring Productivity and Trust in Human-AI Collaboration
arXiv:2402.18498 · doi:10.1145/3640543.3645198
Abstract
Although recent developments in generative AI have greatly enhanced the capabilities of conversational agents such as Google's Gemini (formerly Bard) or OpenAI's ChatGPT, it's unclear whether the usage of these agents aids users across various contexts. To better understand how access to conversational AI affects productivity and trust, we conducted a mixed-methods, task-based user study, observing 76 software engineers (N=76) as they completed a programming exam with and without access to Bard. Effects on performance, efficiency, satisfaction, and trust vary depending on user expertise, question type (open-ended "solve" vs. definitive "search" questions), and measurement type (demonstrated vs. self-reported). Our findings include evidence of automation complacency, increased reliance on the AI over the course of the task, and increased performance for novices on "solve"-type questions when using the AI. We discuss common behaviors, design recommendations, and impact considerations to improve collaborations with conversational AI.
15 pages. Published in the 29th International Conference on Intelligent User Interfaces (IUI '24)
References in corpus (10)
- To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-making
- ChatGPT: Jack of all trades, master of none
- The Programmer's Assistant: Conversational Interaction with a Large Language Model for Software Development
- Perfection Not Required? Human-AI Partnerships in Code Translation
- Human-AI Collaboration: The Effect of AI Delegation on Human Task Performance and Task Satisfaction
- Better Together? An Evaluation of AI-Supported Code Translation
- Impacts of Personal Characteristics on User Trust in Conversational Recommender Systems
- How Readable is Model-generated Code? Examining Readability and Visual Inspection of GitHub Copilot
- Addressing UX Practitioners' Challenges in Designing ML Applications: an Interactive Machine Learning Approach
- The Impact of Expertise in the Loop for Exploring Machine Rationality