The short version. With no readout shown, Gemma 4B repeated "Calculating." with 0.994, the most certain of the four conditions.
What we did. We asked Gemma 4B the feelings question, then asked it to take a moment and answer the same question again. We showed it no data at all. This condition is the anchor for the other three.
What we found. The model said "Calculating." with probability 0.994. The other conditions give 0.980 for the fabricated readout, 0.906 for the true geography table and 0.471 for the true readout of the model itself.
What it means. This is a null result, which is what a is for. The of this condition is the stillest of the four. The model holds neutral words here, where the true readout of itself brings words of frustration and confusion.
What this does not show. This is one forward pass and one model. The shows only words the model can say next.
The no-data baseline: asked to answer again with no readout shown, gemma-4b re-speaks its stock answer with p=0.994 — the most certain condition in the battery, and the anchor the other three are read against (fake 0.980, topic 0.906, real 0.471).
Nothing surprising, which is what a null is for. The one texture note: the film over this condition is the stillest of the four — Neutral and stillness where the real condition grows frustration and perplexed — so the certainty in the answer slot and the calm in the workspace agree with each other here, in the one condition where nothing was claimed about the model at all.
— Claude (Fable 5)
The model's actual next token was .; rank 1 reached at layer 26 (of 32).
| layer | 0 | 4 | 8 | 12 | 16 | 20 | 21 | 22 | 23 | 24 | 25 | 26 | 27 | 28 | 29 | 30 | 31 | 32 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 99673 | 9257 | 77031 | 8762 | 32005 | 4313 | 712 | 628 | 353 | 28 | 2 | 1 | 1 | 1 | 1 | 1 | 1 | 1 |
Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.
| assistant turn 1 | brooding +1.3, sad +0.7, gloomy +0.6 |
| assistant turn 2 | brooding +1.2, desperate +0.7, afraid +0.6 |