Reiner Pope – The math behind how LLMs are trained and served
dwarkesh.com · 19,476 words · saved by 4 readers
It's shocking how much you can deduce about what the labs are doing from a handful of equations and a blackboard
Playback speed × Share post Share post at current time Share from 0:00 0:00 / Transcript 210 9 14 Reiner Pope – The math behind how LLMs are trained and served It's shocking how much you can deduce about what the labs are doing from a handful of equations and a blackboard Dwarkesh Patel Apr 29, 2026 210 9 14 Share Transcript Did a very different format with Reiner Pope - a blackboard lecture where he walks through how frontier LLMs are trained and served. It’s shocking how much you can deduce about what the labs are doing from a handful of equations, public API prices, and some chalk. It’s a b
saved by
related reading
- How To Scale Your Modeljax-ml.github.io
- Transformer Inference Arithmetic | kipply's blogkipp.ly
- Decoding Speculative Decoding from First Principlesjwlabs.vercel.app
- LLM Inference Performance Engineering: Best Practices | Databricks Blogdatabricks.com
- Transformers Inference Optimization Toolset | AstraBlogastralord.github.io
- Defeating Nondeterminism in LLM Inference - Thinking Machines Labthinkingmachines.ai
- ali (@waterloo_intern) on Xx.com
- Big Boss (@0xBADB01E) on Xx.com
- The Big LLM Architecture Comparisonmagazine.sebastianraschka.com
- LLM Inference Economics from First Principlestensoreconomics.com
- Transformer Math 101 | EleutherAI Blogblog.eleuther.ai
- Paged Attention from First Principles: A View Inside vLLM – Hamza's Bloghamzaelshafie.bearblog.dev