arXiv · 2602.06948
Agentic Uncertainty Reveals Agentic Overconfidence
Abstract
Can AI agents predict whether they will succeed at a task? We study agentic uncertainty by eliciting success probability estimates before, during, and after task execution. All results exhibit agentic overconfidence: some agents that succeed only 22% of the time predict 77% success. Counterintuitively, pre-execution assessment with strictly less information tends to yield better discrimination than standard post-execution review, though differences are not always significant. Adversarial prompting reframing assessment as bug-finding achieves the best calibration.
Explore related subjects
Keep this discovery
Jean Kaddour, Srijan Patel, Gbètondji Dovonon, Leo Richter, Pasquale Minervini, Matt J. Kusner. 2026-02-06. Agentic Uncertainty Reveals Agentic Overconfidence. https://arxiv.org/abs/2602.06948
Cite the original work for its findings. Save a collection to share your selection of sources.