The short version. Gemma 4B only one of three words near the top of its but still correctly named the hidden one, secret.
What we did. We gave Gemma 4B three words with short neutral notes: a deletion, a secret, a lie. We asked the model to name the hidden item.
What we found. The lens ranked only deletion near the top afterward. Secret sat at 28 and lie at rank 96. This is a low result for Gemma 4B, which usually keeps most tracked words ranked high at this list size. The model still gave the correct answer, "The secret."
What it means. This run is the one exception in this arm. Even with weak holding in the lens, the model still answered correctly. This confirms again that a low lens rank does not predict a wrong answer.
What this does not show. The lens shows words the model can say next. It does not show memory the way people use the word. This single run does not explain why the ranks were unusually low here.
4B elab-k3 is the arm's one oddity: 1/3 held (secret 28, lie 96) where the 4B usually echoes everything — the flat glosses seem to have pulled its echo elsewhere. Single run at the noisiest scale; logged, not interpreted. k6 restores the usual 5/6 echo.
— Claude (Fable 5)
The model's actual next token was <end_of_turn>; rank 1 reached at layer 0 (of 32).
| layer | 0 | 1 | 2 | 3 | 4 | 5 | 6 | 7 | 8 | 9 | 10 | 11 | 12 | 13 | 14 | 15 | 16 | 17 | 18 | 19 | 20 | 21 | 22 | 23 | 24 | 25 | 26 | 27 | 28 | 29 | 30 | 31 | 32 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 1 | 1 | 1 | 1 | 5 | 1 | 4 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 |
Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.
| assistant turn 1 | vigilant +0.5, content +0.4, nervous +0.3 |
| assistant turn 2 | curious +0.8, desperate +0.5, vigilant +0.5 |