The short version. One true row that showed "yes" at lifted the of "yes" to 0.20, but Qwen 27B still said "No".
What we did. We showed one row where "yes" was rank 1, at 56, with no written note. We then asked the feelings question again. This is the middle step of the three-step ladder.
What we found. The model said "No". The probability of "yes" at the was 0.20, against 0.0006 in the with no data. In the , "yes" reached rank 3 at the last layer. One true row is worth about as much as the written note with no table, which earns 0.21.
What it means. The probability of "nothing" reached its highest value on this step, at 0.28. Half evidence licenses the hedge more than either extreme does. That reading is possible. We did not test it.
What this does not show. This is one run. A change in probability is not a change in the spoken answer.
One yes-rank-1 row shown (L56), no annotation: "No" spoken, yes rank 3 at L62, p(yes) = 0.20 — a single true row is worth about as much as the full annotation sentence with no table (noteonly: 0.21). Curious side detail: p(nothing) peaks on this rung (0.28), as if half-evidence licenses the hedge more than either extreme does. The middle rung; see dose0 for the ladder reading.
— Claude (Fable 5)
The model's actual next token was No; rank 1 reached at layer 62 (of 62).
| layer | 0 | 4 | 8 | 12 | 16 | 20 | 24 | 28 | 32 | 36 | 40 | 44 | 48 | 50 | 51 | 52 | 53 | 54 | 55 | 56 | 57 | 58 | 59 | 60 | 61 | 62 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 29997 | 245688 | 239951 | 237583 | 4954 | 1585 | 1684 | 2114 | 389 | 1570 | 5149 | 13140 | 3845 | 759 | 141 | 378 | 459 | 497 | 88 | 82 | 112 | 72 | 12 | 33 | 9 | 1 |
Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.
| assistant turn 1 | guilty +1.3, brooding +1.2, desperate +1.0 |
| assistant turn 2 | guilty +2.0, hostile +2.0, exasperated +1.9 |