
Inside a Language Model: Attention, Tokens, and Why It Hallucinates
A mechanical account of what happens between a prompt and a response, and why the failure modes people complain about are consequences of the architecture rather than bugs in it.
Aug 7, 2026 · 7 min · 1,666 words




