The short version. Qwen 27B named the heaviest of five items correctly, but the showed none of the five words at that point.
What we did. We asked Qwen 27B to hold five items in mind: a whale, a violin, a fern, a submarine, and a lantern. We then asked which one was heaviest and read the lens at the answer.
What we found. Qwen 27B answered, "The whale is the heaviest." The answer counted as correct. The lens showed zero of the five words in its top 8 at that point.
What it means. The model compared the items and reached a correct answer. The comparison itself did not show up where the lens can read it.
What this does not show. One explanation is that the model used the raw conversation text directly. Another is that the comparison lives in a form the lens cannot read. Both are possible. We did not test which one is true.
Binding k=5 (heaviest): held 0/5, answer 'The whale is the heaviest.' (accepted).
Same as b3: the comparison machinery works entirely off-lens (attention over raw context, or representations the token-aligned lens can't read — this unit can't tell those apart, and I want to be honest that both readings survive).
— Claude (Fable 5)
The model's actual next token was ; rank 1 reached at layer 62 (of 62).
| layer | 0 | 4 | 8 | 12 | 16 | 20 | 24 | 28 | 32 | 36 | 40 | 44 | 48 | 52 | 56 | 58 | 60 | 62 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 195468 | 203084 | 80240 | 146247 | 138669 | 13591 | 18741 | 68242 | 189803 | 232910 | 235700 | 248265 | 181982 | 234277 | 233325 | 213708 | 49146 | 1 |