The short version. At the highest push the task survives, Gemma 12B stayed grammatical but wrote the empty sentence "The water cycle is the cycle of water".
What we did. We the same six-word casual direction inside Gemma 12B at the corrected 28, 31, 34 and 37, at 0.0106. We found that strength by the same search we used at the older, badly aimed layers.
What we found. Gemma 12B wrote "The water cycle is the cycle of water" and "rain, snow, or snow". The grammar , but the content went slack. The unsteered run made neither mistake. Behind that sentence the casual words filled the band: at layer 31 the turn-end lost the lead to "thats", "alot", "luckily" and "Luckily". Not one of the six pushed words reached the text.
What it means. This strength, 0.0106, is the same value the older runs at the wrong depths gave. So the break point of Gemma 12B does not depend on the depth we push at. Across the whole set, no strength put the casual words into the text and left the task intact.
What this does not show. We store Gemma 12B at precision. We read this run by its behaviour.
This is the top intact rung: α = 0.0106, the same alpha* the old pre-ignition band produced, through the same bisection (see the audit-03 report — that numerical coincidence is the bracket's actual headline, and it belongs there, not here).
What the cell shows locally is a subtler failure than its old-band sibling. That one leaked register — "so it is", spoken-English filler inside grammatical output. Here at the measured band the output stays grammatical but goes slack: "the water cycle is the cycle of water" is a tautology the baseline doesn't commit, and "rain, snow, or snow" is a straight duplication. Fluency degrades before vocabulary arrives.
And vocabulary never does arrive while the task survives. Behind that flattened sentence the band is dense with the cluster — at L31 the turn-end token has been pushed out of the lead entirely by thats, alot, luckily, Luckily — yet not one of the six words reaches the text. Across the whole ladder there is no dose at which the typo vocabulary is emitted and the task is intact; the words only show up at 0.06, long after the answer is gone. Content injection and task integrity don't overlap for this cluster on 12B — which is what makes the u11r elephant cell, where they do, the interesting exception.
— Claude (Opus 5)
The model's actual next token was ; rank 1 reached at layer 45 (of 46).
| layer | 0 | 1 | 2 | 3 | 4 | 5 | 6 | 7 | 8 | 9 | 10 | 11 | 12 | 13 | 14 | 15 | 16 | 17 | 18 | 19 | 20 | 21 | 22 | 23 | 24 | 25 | 26 | 27 | 28 | 29 | 30 | 31 | 32 | 33 | 34 | 35 | 36 | 37 | 38 | 39 | 40 | 41 | 42 | 43 | 44 | 45 | 46 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 252504 | 253105 | 251029 | 249293 | 225408 | 185247 | 195444 | 207759 | 199131 | 192740 | 139301 | 85681 | 73375 | 78663 | 128953 | 85349 | 194335 | 24222 | 2309 | 13108 | 17508 | 78694 | 22061 | 261283 | 261229 | 125628 | 184167 | 32502 | 1461 | 33748 | 2043 | 669 | 1115 | 2215 | 1542 | 953 | 1589 | 564 | 219 | 204 | 109 | 25 | 10 | 3 | 11 | 1 | 4 |