AIIC AI Intelligence Centre

SOURCE-LINKED INTELLIGENCE

Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search

arXiv · AI, language, vision and robotics · article · Sep 1, 2026 · UTC

Language model agents increasingly propose actions, observe external feedback, and explain their own behavior. Their confidence and rationales are convenient monitoring signals, but convenience is not verification. We introduce an environment-grounded audit in which every intermediate proposal receives an exact outcome. A language model operates an evolutionary Contexto search whose feedback function assigns every valid guess an exact rank without human annotation. Across 200 runs spanning five configurations and three model families, four reporting configurations produce 12,249 self-reports.

Read original source ↗ Open in workspace

recordType
paper
region
Global

Evidence & attribution

First collected: 2026-09-21T06:21:59.299Z. This is not the publication date.