Skip to content
AI Atlas
PaperActive

Do Agents Know When They Succeed? Calibrating Agent Confidence from Internal Representations

arxiv.org/abs/2609.09448

quality89

Updated 8 h ago · first seen 11 Sept 2026

paper_01M294GK41Q0R1AMAKX5QG3N90

Published
11 Sept 2026
T1 · 8 h ago
arXiv
2609.09448
T1 · 8 h ago
Category
cs.AI
T1 · 8 h ago

As of

Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.

Claim history · arXiv id

1 claims · 1 propertiesShow all properties

arXiv idarxiv_id1

Claim history for arXiv id
ValueValid from → toStatusSourceConfidenceExtractor
2609.09448currentcurrentarXiv (Atom API + RSS)T1highdeterministic

Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →