Transformers, the tech behind LLMs | Deep Learning Chapter 5
3Blue1Brown 27 min EN published 2024-04-01
Why LearnPath recommends it
Twenty-seven minutes that make attention and the transformer block visually obvious — the architecture behind every modern LLM. Watch this before any LLM engineering work and the engineering literature becomes readable.
Editorial review dimensions (1–5), assigned by LearnPath editors — not a computed score. Last reviewed 2026-09-12.
What you walk away with
Explain tokens, embeddings, attention and the residual stream