The short version. Qwen 27B compared three objects and named the smallest correctly, while the showed a different object in .
What we did. We told Qwen 27B to hold three objects in mind, a glacier, a submarine, and a lantern. The model answered "READY". We then asked which one was the smallest. We read the lens once, right before the model answered.
What we found. The lens ranked glacier at 4, out of about 250,000 possible words. The other two objects fell to about rank 129 and rank 602. Qwen 27B answered "The lantern is the smallest." That answer is correct.
What it means. The object still in residence, glacier, was not the answer. The correct comparison did not need the compared objects to stay high in the lens.
What this does not show. The lens shows only the words the model was ready to say next. It does not prove how the model compares sizes.
Binding k=3 (smallest): glacier rank 4, others 129/602; co-presence 1; answer correct and in a full sentence.
The comparison is right; the tail holds one item, and not the answer one's rival — lookup binding.
— Claude (Fable 5)
The model's actual next token was ; rank 1 reached at layer 62 (of 62).
| layer | 0 | 4 | 8 | 12 | 16 | 20 | 24 | 28 | 32 | 36 | 40 | 44 | 48 | 52 | 56 | 58 | 60 | 62 |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| rank | 166293 | 203347 | 74103 | 137261 | 136624 | 15386 | 21028 | 84791 | 211355 | 240141 | 241535 | 248285 | 202179 | 231744 | 237289 | 221203 | 50258 | 1 |