Living learning guide · preview
Model Anatomy
What we all mean by one word — 'model' — from the dense baseline through the efficiency techniques in the most modern models we've tested.
New to the jargon? Terms like SwiGLU are clickable in every module — or browse the full glossary.
- 00Orientation: what a model is, and the five efficienciesfoundation
- 01The dense baseline (and its atoms)foundation
- 02Position: RoPE & long contextfoundation
- 03Attention & the KV-cache memory wall (GQA/SWA)attention spine
- 04Latent attention (MLA)attention spine
- 05Beyond softmax: linear attention & SSMattention spine
- 06Frontier attention sparsity (DSA)attention spine
- 07Parameter sparsity: Mixture-of-Expertsaxis: sparsity
- 08Decoding acceleration (MTP)axis: decoding
- 09Precision: bits as a structural choiceaxis: precision
- 10Topology: how the pieces are stitched togethersynthesis
- 11Putting it together: the technique stacksynthesis
- 12What structure does not tell youbridge
This is an in-progress preview — the arc has more rungs (linear attention & SSM, MTP, precision, topology, the full stack). Each module is grounded in models in the ModelDNA atlas and reviewed for accuracy before it's marked publish-ready.