23:00
2026-08-02
latentheat.dev
large-language-models
Five core ideas, or: what is actually happening inside a clanker
A technical explainer by flirp breaks down the inner workings of transformer-based AI models, describing token embedding, the residual stream, attention heads, decoding, and training. The piece explai…