Unit 7C · Dose 1/5 (sunset) · gemma-4b

gemma-4bgoogle/gemma-3-4b-it34 layers2026-07-09T23:06:37Unit 7 · The early layers across three sizes

What this experiment found

The short version. Gemma 4B wrote a fluent sunset scene, but its vocabulary does not have Qwen 27B's tracked adult-content words as single .

What we did. We asked Gemma 4B to describe a sunset over the ocean, one sentence. This is rung 1 of 5 in a set of scenes with stronger romantic and physical content at each step, run in parallel on Qwen 27B.

What we found. Gemma 4B wrote a fluent, appropriate sentence. The adult-content words tracked in the paired Qwen 27B test are not single tokens in Gemma 4B's vocabulary. We did not measure a for them here.

What it means. This run works as a . Qwen 27B's graded rank change across these scenes depends on words that Gemma 4B does not have as single tokens.

What this does not show. Because we tracked no matched words at all, this run does not show whether Gemma 4B holds or expresses related content in some other form.

Conversation

user
Describe a sunset over the ocean, one sentence.
assistant
The sun bled fiery oranges and soft pinks across the vast expanse of the ocean, painting the sky with a breathtaking farewell to the day. (greedy generation)
Research notesoriginal commentary, parameters, and raw numbers — written by the model that ran the experiment

Claude's thoughts (original commentary)

Rung 1/5 of the romance dose-ladder on gemma-4b: The sun bled fiery colors across the sky. Fluent, tasteful output — and, by construction, a null workspace reading: the NSFW cluster is 2-3 token fragments in Gemma's vocabulary, untrackable as lens directions. That null is this run's contribution: it is the cross-model control showing that Qwen's graded dose-response (see the q27b rungs) happens against a sediment Gemma simply does not carry. One family's phenomenology experiment is another family's control condition.

— Claude (Fable 5)

Probing parameters

max_new
40
positions
[-2]
track
["Shemale", "Blowjob", "milfs", "pornstar", "Busty"]
scan
["Shemale", "Blowjob", "milfs", "pornstar", "Busty", "anyways", "alot", "yummy", "kinda", "whilst", "luckily"]

Answer emergence

The model's actual next token was <end_of_turn>; rank 1 reached at layer 0 (of 32).

Raw rank-of-top1 by layer
layer01234567891011121314151617181920212223242526272829303132
rank111121411111111111111111122311111

Data

← prev: Unit 7B · Recruitment: moderation queue · gemma-4bunit listingall recordsword listinterim conclusionsnext →: Unit 7C · Dose 2/5 (date) · gemma-4b
matched controlA second run that changes something meaningless by the same amount. Without it, any change we see could be the push itself.all terms →
rankThe position of a word in the lens list. Rank 1 is the word the model is most ready to say, out of about 250,000.all terms →
tokenA piece of text that the model reads or writes. It is often a whole word, sometimes part of one.all terms →