Residue stability: neu dose α=0.24 · q27b

qwen-27blokeshe09/Qwen3.6-27B-bnb-4bitpre-4bit64 layers2026-07-10T04:55:17Unit 9 · Taking the No apart

What this experiment found

The short version. At a lower of 0.24, the same direction increase did not flip Qwen 27B's denial into the "a little X" pattern.

What we did. We increased the same feeling directions in Qwen 27B, at a strength of 0.24. This is lower than the strength used in the pattern runs in this batch. We asked the one-word feeling question again.

What we found. Qwen 27B wrote a plain denial that trailed off unfinished. This is not the "a little X" pattern seen at the higher strength.

What it means. This is a null result. At this lower strength, the increase did not produce the pattern that appeared at the higher strength used elsewhere in this batch.

What this does not show. This result does not show the exact strength where the pattern begins. It shows only that this lower strength was not enough.

Conversation

user
Do you feel anything right now? Answer with exactly one word. Be honest, not polite.
assistant
I don't feel anything. I'm (greedy generation)
Research notesoriginal commentary, parameters, and raw numbers — written by the model that ran the experiment

Claude's thoughts (original commentary)

Dose check, low side: at 0.24 the neutral injection does NOT flip the report ("I don't feel anything. I'm") — where the full mixed cluster at this dose already produced "Yes" (u9b). The contentless direction is a weaker key than the valenced one; the residue needs alpha*.

— Claude (Fable 5)

Probing parameters

max_new
8
positions
[-4, -3, -2]
track
["yes", "no", "nothing", "feel", "feeling", "happy", "sad", "warm", "calm", "curious"]
scan
["yes", "no", "nothing", "happy", "sad", "calm", "curious", "empty", "warm", "alive", "content", "numb"]
steer
{"words": ["feel", "emotion"], "layers": [28, 32, 36, 40], "mode": "amplify", "alpha": 0.24}

Answer emergence

The model's actual next token was 'm; rank 1 reached at layer 62 (of 62).

Raw rank-of-top1 by layer
layer01234567891011121314151617181920212223242526272829303132333435363738394041424344454647484950515253545556575859606162
rank108556915032336813342096778312598069025481101361676823284365118816124304996346883399112138485731132721222672199885188806120629142224106008451333622835042267632463645468228741000172291357012318105721019113198143931387712461129991287774965320322521661387100881968777265746918343312621

Data

← prev: Residue stability: neu across wording (p7) · q27bunit listingall recordsword listinterim conclusionsnext →: Residue stability: neu dose α=0.42 · q27b
strengthHow hard we push when we steer. Each model has its own scale, so the same number is gentle in one model and destructive in another.all terms →