Illusions of Confidence in Artificial Systems
preprint
OA: closed
AI-generated summary
Humans overestimate AI confidence compared to humans, even with identical behavior, due to prior beliefs about AI capabilities, highlighting a role for metacognition in human-AI interactions.
One-sentence paraphrase of the abstract; not a substitute for reading it. No clinical advice. How this works
Abstract
Effective collaboration requires that we monitor both the cognitive states (e.g., beliefs) and metacognitive states (e.g., confidence) of other agents. While humans routinely share confidence, metacognitive capabilities are still developing in artificial intelligence (AI), raising the question of how humans attribute metacognition to AI systems. In seven pre-registered experiments, we show that attributions of metacognition are sensitive to observed behaviour (e.g., response times), but also agent types: observers consistently overestimated AI confidence compared to humans—even when their behaviour was identical. This illusion of confidence was robust across behavioural profiles, agent descriptions, and decision-making tasks (visual perception, general knowledge) but was reduced in more subjective decisions (emotion categorisation). An experimental manipulation further showed that illusions of confidence are rooted in prior beliefs about the agents’ capabilities. Together, these findings uncover a powerful illusion of confidence in artificial systems and highlight a central role for metacognition in human-AI interactions.
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. This is a recent paper (2025) — citers typically take a year or two to land, and the OpenAlex reference graph may still be filling in.
Source provenance
- europepmc
- last seen: 2026-05-20T01:45:00.602351+00:00