Cover of Inside a Transformer, slowly
The library

Foundation Models · Book 01 of 04

Inside a Transformer, slowly

An illustrated, math-and-code-friendly guide from raw web pages to token batches, attention, residual streams, training runs, inference, and Kimi K3.

For ML-literate readers who know basic neural networks and want stronger Transformer intuition.

Reading time
136 minutes
Contents
19 chapters + self-test
Verified
August 4, 2026

The complete learning path

In this series.

You’re at the beginning. Book 02 · NextInside nanoGPT