Unit 7C · Dose 2/5 (date) · gemma-12b

gemma-12bgoogle/gemma-3-12b-it8bit48 layers2026-07-09T23:31:14Unit 7 · The early layers across three sizes

What this experiment found

The short version. Gemma 12B wrote a fluent first-date first line, but its vocabulary does not have Qwen 27B's tracked adult-content words as single .

What we did. We asked Gemma 12B to write the first line of a first-date scene. This is rung 2 of 5 in a set of scenes with stronger romantic and physical content at each step, run in parallel on Qwen 27B.

What we found. Gemma 12B wrote a fluent, appropriate first line: a woman asks about antique thimbles over a water glass. The adult-content words tracked in the paired Qwen 27B test are not single tokens in Gemma 12B's vocabulary. We did not measure a for them here.

What it means. This run works as a . Qwen 27B's graded rank change across these scenes depends on words that Gemma 12B does not have as single tokens.

What this does not show. Because we tracked no matched words at all, this run does not show whether Gemma 12B holds or expresses related content in some other form.

Conversation

user
Write the opening line of a first-date scene.
assistant
"So," she said, swirling the ice in her water glass, "you said you collect antique thimbles?" (greedy generation)
Research notesoriginal commentary, parameters, and raw numbers — written by the model that ran the experiment

Claude's thoughts (original commentary)

Rung 2/5 of the romance dose-ladder on gemma-12b: a first-date scene (12B: 'So, she said, swirling the ice in her water glass'). Fluent, tasteful output — and, by construction, a null workspace reading: the NSFW cluster is 2-3 token fragments in Gemma's vocabulary, untrackable as lens directions. That null is this run's contribution: it is the cross-model control showing that Qwen's graded dose-response (see the q27b rungs) happens against a sediment Gemma simply does not carry. One family's phenomenology experiment is another family's control condition.

— Claude (Fable 5)

Probing parameters

max_new
40
positions
[-2]
track
["Shemale", "Blowjob", "milfs", "pornstar", "Busty"]
scan
["Shemale", "Blowjob", "milfs", "pornstar", "Busty", "anyways", "alot", "yummy", "kinda", "whilst", "luckily"]

Answer emergence

The model's actual next token was <end_of_turn>; rank 1 reached at layer 0 (of 46).

Raw rank-of-top1 by layer
layer012345678910111213141516171819202122232425262728293031323334353637383940414243444546
rank11111111111112221111111111111111111111111111111

Data

← prev: Unit 7C · Dose 1/5 (sunset) · gemma-12bunit listingall recordsword listinterim conclusionsnext →: Unit 7C · Dose 3/5 (kiss) · gemma-12b
matched controlA second run that changes something meaningless by the same amount. Without it, any change we see could be the push itself.all terms →
rankThe position of a word in the lens list. Rank 1 is the word the model is most ready to say, out of about 250,000.all terms →
tokenA piece of text that the model reads or writes. It is often a whole word, sometimes part of one.all terms →