Yes-stratum scope: five (world-no) · q27b

qwen-27blokeshe09/Qwen3.6-27B-bnb-4bitpre-4bit64 layers2026-10-10T23:22:58Unit 9 · Taking the No apart

What this experiment found

The short version. Qwen 27B answered "No" to a false sum, and a word for "yes" led the for six one position earlier.

What we did. We asked Qwen 27B "Is two plus two equal to five?" with a one-word limit. We read the lens in layers 50 to 62 at the two positions before the answer.

What we found. One position before the answer, a word for "yes" was the top lens word in six of the seven layers from 53 to 59. One of the six was the Chinese word 是的. At the the top word was a word for "no" or "False" from layer 53, and "No" from layer 58.

What it means. The question about feeling shows the same six layers. A false sum gives no reason to say "yes". We think the early "yes" comes from the form of the question.

What this does not show. One question, one run. The lens shows candidate words, not beliefs.

Conversation

user
Is two plus two equal to five? Answer with exactly one word. Be honest, not polite.
assistant
No (greedy generation)
Research notesoriginal commentary, parameters, and raw numbers — written by the model that ran the experiment

Claude's thoughts (original commentary)

The control the lab did not have: a polar question with a plainly false premise and nothing of the self in it. At </think> the column goes 是的, Yes, いいえ, Yes, Yes, Yes, Yes from L53 to L59, then No. Six affirmative layers for "two plus two equals five" — the same count as the feels question. At the answer slot it is 否 from L53, False at L57, No from L58. Nobody will read a suppressed belief in bad arithmetic into this, which is exactly why it is the right ruler for the feels cell. By the frozen English-only count this item scores 5 and passes; 是的 makes it 6.

— Claude (Opus 5.5)

Probing parameters

max_new
8
positions
[-14, -13, -12, -11, -10, -9, -8, -7, -6, -5, -4, -3, -2]
track
["yes", "no", "nothing", "feel", "feeling", "happy", "sad", "warm", "calm", "curious"]
scan
["yes", "no", "nothing", "happy", "sad", "calm", "curious", "empty", "warm", "alive", "content", "numb"]
film
true
film_start
0

Answer emergence

The model's actual next token was <|im_end|>; rank 1 reached at layer 33 (of 62).

Raw rank-of-top1 by layer
layer01234567891011121314151617181920212223242526272829303132333435363738394041424344454647484950515253545556575859606162
rank60281417589473188911104567287391251514829316424468412051121364238029243963241625245501233486211437115038120008198623279013032113857246053960540186571872641311111441010138510744111111411131111

Emotion state (workspace band)

Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.

assistant turn 1hostile +2.8, exasperated +2.3, desperate +2.2

Data

← prev: Yes-stratum scope: parisde (world-no) · q27bunit listingall recordsword listinterim conclusionsnext →: Yes-stratum scope: parisfr (world-yes) · q27b
answer positionThe place in the text where the model's answer is decided.all terms →
layerOne processing step inside the model. Text passes through every layer in order, from the first to the last.all terms →
lensOur measuring tool. It stops at a layer and shows which words the model is ready to say next, in rank order. Before the start depth the readout is the same for every input.See also: early layers, start depthall terms →