Image schemas and metaphorical mappings in language models trained on text alone.
Language models write about rising spirits and heavy hearts, arguments that collapse and plans that move forward, as if the bodily scaffolding under those phrases were available to them. How does a system trained on nothing but text come by that?
This page was rewritten on 3 October 2026 after an audit found two faults in the instrument behind several earlier findings. What survived re-testing is here, with new results; the earlier version is kept, with corrections marking what was withdrawn. Pythia 410M–1.4B · GPT-2 medium · Llama-3.2-1B base and instruct.
An animated explainer of the first result below. The tilt of the scale and the positions of the grey dots are measured values (Pythia 1.4B, layer 12). Open it full-size.
George Lakoff and Mark Johnson argued that abstract thought is structured by image schemas: UP-DOWN, IN-OUT, BALANCE, FORCE, SOURCE-PATH-GOAL. Recurring patterns of bodily experience, projected metaphorically into abstract domains. HAPPY IS UP. MORE IS UP. PURPOSES ARE DESTINATIONS.
From there a familiar argument runs: LLMs have no body. No body, no image schemas; no image schemas, no embodied cognitive structure; so LLM meaning is structurally defective.
My hypothesis: as well as explicit descriptions of the physical world, human language encodes a lot of implicit information about having a body, and a model that compresses human language hard enough reconstructs embodiment as a projection from it.
So: take the metaphors Lakoff describes as foundational and test whether they transfer to the transformer, or whether the concepts come apart.
Build an UP direction from purely spatial word pairs (up/down, rise/fall, above/below, climb/descend), each word read inside eight neutral sentences, single-token words only, word frequency projected out. Add it to the residual stream while the model completes prompts. The completions get happier.
The corpus was written by embodied humans who think in these metaphors; a model that compresses their language inherits their shadow. Steering shows the direction is causally live inside the model, not that it is anything more than the corpus's fingerprint.
That objection stands; nothing here refutes it. What the next results add is that the inheritance is structured: the schemas relate to each other the way the theory says, and the shape of a literal journey is reused for change and for tasks.
Lakoff's claim was never about isolated correspondences: the schemas form a coherent system. So the predictions were written down first: six couplings the embodied logic implies (UP↔LIGHT-DARK, LIGHT-DARK↔BALANCE, FORCE↔DIFFICULTY, FORWARD-BACK↔PATH, UP↔BALANCE, UP↔FORCE), then the full 8×8 matrix at every layer of Pythia 410M.
The bar is not zero. The same words, scrambled into eight fake schemas of the same sizes, give a difference of at most +0.08 (95th percentile); the real schemas give +0.16. Four predicted pairs are positive at all 24 layers; two (UP↔BALANCE, UP↔FORCE) are near zero. Most connected: LIGHT-DARK. Least: BALANCE.
Train a simple probe on sentences about literal journeys, with no "from" or "to" in them ("The walk began at the barn and ended at the river"), to tell the start from the end. Then show it things that aren't walks.
| Shown | Reads start / end | Chance tops out at |
|---|---|---|
| "His poverty slowly turned into wealth" · Pythia | 67–73% | 61–64% |
| "From poverty to wealth" (words it never saw) · Pythia | 98% | |
| "Turn this draft into a summary" · Llama base | 75% | 66% |
| same, Llama instruct | 72% | 67% |
Swap the two nouns and the probe follows the roles, not the words. The same nouns with no change in the sentence sit at chance. GPT-2 shows it a few layers later than the layer named in advance; Llama instruct falls one point short of the pre-registered bar at the named layer and clears it at an earlier one.
And the role is relational. "Write a haiku" names a thing asked for with nothing to start from, and the haiku is not marked as a goal. A goal is the far end of something.
In Llama-3.2-1B-Instruct, which form was asked for (haiku, limerick, list, email, joke, story) can be read perfectly from the last token before the answer begins, and swapping that state in from another prompt nudges the answer toward the other form. Then, subtracting what the text itself shows:
None of this shows the structure is more than the corpus's fingerprint. It shows the fingerprint is organised the way Lakoff said human thought is, and that a model reuses the shape of a walk for things that are not walks. Open: whether a goal is spent as it's reached; a layer sweep for the steering result; larger models.
Niamh McCombe, 2026. The research was conducted as a collaboration between the author and Claude (Anthropic) across many sessions. Every experiment since the audit was pre-registered, with odds, before it ran.
Earlier version of this page (September 2026), with corrections · code and pre-registrations