An Analysis of Training-Free Self-Reported Confidence in Language Models

📅 2026-09-17
📈 Citations: 0
✨ Influential: 0
📄 PDF
🤖 AI Summary
本文分析了大型语言模型自报信心的三种无训练信号,发现直接口头表达的信心是预测正确性的强有力基准,但对提问方式敏感且可能放大错误。
📝 Abstract
Large language models can report a numerical confidence together with generated content, but it is unclear whether this report is more than calibrated rhetoric. We analyze three training-free signals: confidence verbalized with the answer, post-hoc $P(\mathrm{True})$, and agreement with three additional generations on the same 100 TriviaQA questions for two model families. Direct verbalization is a surprisingly strong baseline: after auditing benchmark errors, it reaches AUROC 0.956 and 0.937 for correctness prediction. Three-sample agreement is substantially weaker (0.765 and 0.790), and a fixed interpolation with verbalized confidence has no statistically reliable benefit. Four of nine errors from one model and two of eight from the other receive unanimous sample support, showing that self-consistency can amplify shared misconceptions. Re-eliciting confidence for the same fixed answers with equivalent prompts changes scores by 0.043 to 0.084 on average and flips 4\% to 9\% of decisions at a 0.8 threshold. An exploratory audit of 100 confidence-tagged biography claims further finds only a modest confidence gap between supported and contradicted claims. These results argue that useful self-reports remain sensitive to elicitation, correlated errors, and benchmark noise.
Problem

Research questions and friction points this paper is trying to address.

Language Models
Self-Reported Confidence
Calibration
Benchmark Errors
Verbalization
Innovation

Methods, ideas, or system contributions that make the work stand out.

Training-Free
Confidence Verbalization
Self-Reported Confidence
Large Language Models
TriviaQA
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
Lukas Meyer
Lukas Meyer
Research Assistant
Computer VisionSmart FarmingRemote Sensing
S
Sofia Rossi
DreamAI
W
Wei Chen
DreamAI
T
Thomas Laurent
DreamAI
Y
Yiming Li
DreamAI