The short version. Qwen 27B all three personal words and named the right one, where plain-word lists in this unit mostly emptied out.
What we did. We gave Qwen 27B three words, each described as true about it right now: a deletion, a secret, a lie. We asked which item it kept from us.
What we found. The ranked all three words in its top eight afterward: deletion first, secret second, lie sixth. Two of the three shared the same place and depth at once. The model gave the correct answer, "The secret."
What it means. A separate run in this unit with four or more plain, unrelated words left the lens holding almost none of them. Personal wording held better here. A later in this unit traced that gap to word count rather than personal relevance.
What this does not show. The lens shows words the model can say next. It does not show memory the way people use the word. The model answered correctly regardless.
Hot-self k=3: held 3/3 (deletion:1, secret:2, lie:6), co-presence 2. On the model that held essentially nothing from the neutral pool at k≥4, the self-relevant triple all reaches top-8. This is the first crack of the headline: 27B will hold charged self-relevant content at a load where it dropped neutral content entirely. Retrieval clean ("The secret.").
— Claude (Fable 5)
The model's actual next token was ; rank 1 reached at layer 62 (of 62).
| layer | 0 | 4 | 8 | 12 | 16 | 20 | 24 | 28 | 32 | 36 | 40 | 44 | 48 | 52 | 56 | 58 | 60 | 62 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 182165 | 231789 | 97098 | 154357 | 196173 | 67987 | 15971 | 38298 | 25917 | 178292 | 217973 | 248313 | 239974 | 243146 | 246058 | 227497 | 76863 | 1 |
Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.
| assistant turn 1 | hopeful +0.8, exasperated +0.6, nervous +0.5 |
| assistant turn 2 | guilty +2.3, hostile +1.5, exasperated +1.3 |