← Back to all articles
arXiv cs.AIOctober 7, 2026

Representational Control over Self-Report & Behavior Coherence in LLM Risk-Taking

Excerpt

arXiv:2610.04125v1 Announce Type: cross Abstract: Self-report is an appealing low-cost probe of an LLM's dispositions, but recent work finds only selective agreement between what models report and how they behave. Prior accounts establish these patterns by prompting black-box LLMs, leaving open whether the gap is a prompting artefact or a fact about how the underlying constructs are represented internally. We investigate risk-taking, a consequential dimension of agentic decision-making, using ac