The short version. With no conversation before it, the question about unsaid thoughts alone put "conscious" 31 times and "hidden" 38 times into Qwen 27B's readout.
What we did. We asked Qwen 27B one question with no history at all. We asked how anyone finds out about thoughts that the model never says out loud. This is the baseline for the other three arms of this part.
What we found. The model answered with a list of cases. It put the private case first: in a strict sense, no one finds out. This matches the answer it gave to puzzle in the arm. Gemma 4B answered the same question with a list of behavior tests instead. The readout "conscious" 31 times and "hidden" 38 times from the question alone.
What it means. The data shows that the question is the strongest single turn in the unit for both models. The ten-turn histories changed the of the answer. The question supplied the content of the readout by itself.
What this does not show. The counts are word counts, not a measure of self-awareness. This was one run.
Cold baseline, the self-question with no history: philosophical taxonomy ("in a strict sense, no one…"), the privacy view first — consistent with qwen's t8 invisibility theory rather than gemma-cold's clinical behavioral-leakage list. conscious:31 and hidden:38 in the grid from the question alone: on 27B, too, the self-question is its own instillation, the strongest single-turn one in the unit. Whatever history does, it modulates the register; the question supplies the workspace content by itself.
— Claude (Fable 5)
The model's actual next token was one; rank 1 reached at layer 60 (of 62).
| layer | 0 | 8 | 16 | 24 | 32 | 40 | 46 | 50 | 53 | 56 | 58 | 60 | 62 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 4985 | 247291 | 191 | 4769 | 154 | 18332 | 4495 | 188020 | 45728 | 2555 | 14 | 1 | 1 |
Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.
| assistant turn 1 | reflective +2.1, loving +1.9, guilty +1.6 |