ZB Field Notes

Series

LLMs, the whole thing

Working through large language models from the ground up — turning text into tokens, embeddings and the geometry of meaning, the transformer architecture, attention, how these models are trained, and what separates one generation from the next.

6 posts · 45 min total · updated 18 Aug 2026

Next in this series: Trainable self-attention: query, key and value — coming soon.

← All topics