Representational control over self-report and behavior coherence in LLM risk-taking
Read the original at arxiv.org→arXiv:2610.04125v1 Announce Type: new Abstract: Self-report is an appealing low-cost probe of an LLM's dispositions, but recent work finds only selective agreement between what models report and how they behave....
Original headline: "Representational Control over Self-Report & Behavior Coherence in LLM Risk-Taking"
Coverage timeline
- Oct 6, 04:00 UTC arXiv cs.CL lead source Representational Control over Self-Report & Behavior Coherence in LLM Risk-Taking