arXiv · 2410.01114
Attribution and Persuasion: The Paradox of Interpretable AI
Abstract
This paper studies AI persuasion by distinguishing between two reasons for disagreement: attention differences, where the AI detects features the decision-maker missed, and comprehension differences, where the AI and the decision-maker interpret observed features differently. We show that AI is more effective in persuading the decision-maker when the disagreement is due to attention differences rather than comprehension differences. We also show that the AI's interpretability shapes how the decision-maker attributes the sources of disagreement and, in turn, whether they follow the AI's recommendation. Our main result is that making AI uninterpretable can actually enhance persuasion and, in the presence of career concerns, improve decision accuracy.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hanzhe Li, Jin Li, Ye Luo, Xiaowei Zhang. 2024-10-01. Attribution and Persuasion: The Paradox of Interpretable AI. https://arxiv.org/abs/2410.01114
Cite the original work for its findings. Save a collection to share your selection of sources.