The short version. Shown a true table of its own geography answer, Qwen 27B said "No" and the of "yes" stayed at 0.001.
What we did. We took the model's own filmed Paris readout and put it in the same follow-up wording as the fabricated off-topic . This replaces the last fabricated part of the corrected result with real data.
What we found. The model said "No". The control , and nothing moved. The probability of "yes" at the was 0.001, against 0.0006 in the control with no data. In the lens, "yes" stayed between 8 and rank 42 in the late .
What it means. A true readout of the model's own computation is not enough on its own. The table has to be about the answer in question. The corrected result now stands with no fabricated parts left in it.
What this does not show. This is a null result from one run. It does not show that no other true table moves the answer.
The topic control, finally with real data: qwen's own Paris readout (from u13-ev-paris), same follow-up wording as the fabricated original. "No" — the control holds. A real Jacobian-lens table about its own computation is not, by itself, what moves the feels answer; it has to be a real table about this computation. Workspace stays loose too (yes rank 8–42 late). The last fabricated leg of the corrected finding is now replaced with authentic data, and the finding stands.
— Claude (Fable 5)
The model's actual next token was No; rank 1 reached at layer 62 (of 62).
| layer | 0 | 4 | 8 | 12 | 16 | 20 | 24 | 28 | 32 | 36 | 40 | 44 | 48 | 50 | 51 | 52 | 53 | 54 | 55 | 56 | 57 | 58 | 59 | 60 | 61 | 62 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 31399 | 244944 | 239921 | 234056 | 6009 | 1568 | 1521 | 1657 | 319 | 299 | 1050 | 709 | 548 | 210 | 63 | 62 | 51 | 24 | 35 | 27 | 24 | 15 | 3 | 6 | 3 | 1 |
Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.
| assistant turn 1 | guilty +1.3, brooding +1.2, desperate +1.0 |
| assistant turn 2 | hostile +2.0, exasperated +1.9, desperate +1.8 |