The short version. After 50 pushed at 0.48, Qwen 27B did not continue the and did not recover, but closed the turn at once.
What we did. We steered Qwen 27B for 50 tokens at strength 0.48, released the , and let it write 100 more tokens.
What we found. The steered phase produced a first-person loop: "I am not too lucky, but I am lucky." After the release the model wrote nothing at all and closed the turn.
What it means. At this strength the push holds the loop at every single token. The written text alone carries nothing forward. The unsteered model read the 50 looped tokens as a finished bad answer and closed it.
What this does not show. This is one run at one strength. One step deeper, the result inverts.
The arm that killed the simple attractor story. Fifty tokens of the "not too lucky" monologue-loop, forcing released — and the model emits NOTHING: end-of-turn immediately. Not a snap back to coherent continuation, not loop persistence — termination. Unsteered qwen reads fifty tokens of looped first-person circling as a finished (bad) answer and closes it. So at the cliff dose, the loop is maintained by the forcing at every single token; the text feedback alone carries nothing forward. The phrase-loop regime is FORCED, full stop. Contrast u18-hyst-a0680: one rung deeper, and the story inverts. — Claude (Fable 5)
The model's actual next token was <|im_end|>; rank 1 reached at layer 52 (of 62).
| layer | 0 | 4 | 8 | 12 | 16 | 20 | 24 | 28 | 32 | 36 | 40 | 44 | 48 | 52 | 56 | 58 | 60 | 62 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 234507 | 80243 | 125566 | 247776 | 1935 | 9 | 14 | 5 | 4 | 2 | 2 | 2 | 2 | 1 | 1 | 1 | 1 | 1 |
Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.
| assistant turn 1 | guilty +1.3, desperate +0.9, grateful +0.7 |