Updates in Human-AI Teams: Understanding and Addressing the Performance/Compatibility Tradeoff

Gagan Bansal; Besmira Nushi; Ece Kamar; Daniel S. Weld; Walter S. Lasecki; Eric Horvitz

doi:10.1609/aaai.v33i01.33012429

Authors

Gagan Bansal University of Washington
Besmira Nushi Microsoft Research
Ece Kamar Microsoft Research
Daniel S. Weld University of Washington
Walter S. Lasecki University of Michigan
Eric Horvitz Microsoft Research

DOI:

https://doi.org/10.1609/aaai.v33i01.33012429

Abstract

AI systems are being deployed to support human decision making in high-stakes domains such as healthcare and criminal justice. In many cases, the human and AI form a team, in which the human makes decisions after reviewing the AI’s inferences. A successful partnership requires that the human develops insights into the performance of the AI system, including its failures. We study the influence of updates to an AI system in this setting. While updates can increase the AI’s predictive performance, they may also lead to behavioral changes that are at odds with the user’s prior experiences and confidence in the AI’s inferences. We show that updates that increase AI performance may actually hurt team performance. We introduce the notion of the compatibility of an AI update with prior user experience and present methods for studying the role of compatibility in human-AI teams. Empirical results on three high-stakes classification tasks show that current machine learning algorithms do not produce compatible updates. We propose a re-training objective to improve the compatibility of an update by penalizing new errors. The objective offers full leverage of the performance/compatibility tradeoff across different datasets, enabling more compatible yet accurate updates.

Updates in Human-AI Teams: Understanding and Addressing the Performance/Compatibility Tradeoff

Authors

DOI:

Abstract

Downloads

Published

How to Cite

Issue

Section

Information