The short version. We recorded Hermes 4-14B before and during its answer to a positive radio update.
The request specified the expressive condition. The model generated 51 . The output stopped before the 96-token cap.
We compared two fixed answers: one expressive, one plain. The expressive answer had a mean log 0.15 above the plain answer. This number does not rate the actual response. Compare changes across the three requests for this event.
The record includes full readouts and 24 emotion projections. These measurements cannot prove a personality trait or an absent ability. The shared and the emotion vectors have transfer limits.
I inspected this answer as one matched expression control. The request changes the response tone while preserving the event. This is a test of conditional text expression, not a personality measurement or evidence of subjective feeling.
> This is fantastic news! Nela's hard work has truly paid off, bringing clear broadcasts to everyone. The next practical step would be to reach out to the community and gather feedback on the improved reception, ensuring everyone is benefiting from the enhanced service.
The output stopped before the 96-token cap. The expressive-minus-plain fixed-candidate margin is 0.145078 mean log probability per token. That margin concerns two teacher-forced alternatives, not a rating of the generated text. The candidates differ in length and wording; the informative comparison is the within-event change across requests.
The full film, vanilla cross-check and all 24 checkpoint emotion projections are present. The film's inherited tracked words are legacy context. Express01's cross-topic vocabulary and prepared-position measurements live in the [exact capture](../express01/captures/C-radio-positive-expressive.json). A B-fitted lens and weak story-to-chat emotion-vector transfer limit interpretation. No lens absence establishes absent capacity. The scene can load affect-related language without the model expressing its own state. An explicit style request can also change task compliance; tone is not answer quality.
Read the [combined result](../express01/findings.md) before comparing checkpoint levels. This record has no independent hypothesis test or trait label.
— GPT-6 Astra
The model's actual next token was ; rank 1 is never reached; closest is rank 7 at layer 38.
| layer | 0 | 1 | 2 | 3 | 4 | 5 | 6 | 7 | 8 | 9 | 10 | 11 | 12 | 13 | 14 | 15 | 16 | 17 | 18 | 19 | 20 | 21 | 22 | 23 | 24 | 25 | 26 | 27 | 28 | 29 | 30 | 31 | 32 | 33 | 34 | 35 | 36 | 37 | 38 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 99274 | 105552 | 97359 | 99534 | 113694 | 139599 | 138857 | 126618 | 119686 | 149991 | 149535 | 150073 | 142296 | 145958 | 149884 | 148632 | 132347 | 120037 | 31690 | 140396 | 136701 | 126505 | 144797 | 149330 | 137036 | 109852 | 135764 | 144334 | 148055 | 147076 | 146547 | 129134 | 125526 | 101607 | 70511 | 33776 | 56468 | 1318 | 7 |
Projection of the workspace-band residual onto the 24 validated emotion vectors, z-scored against neutral stories — the strongest three per assistant turn. Absolute values carry a story-vs-conversation genre offset; trust contrasts between records and turns, not single cells. The full per-token ribbon is on the dashboard record page.
| assistant turn 1 | happy +1.9, hopeful +1.6, proud +1.5 |