✳flâneur — a map of the web's best reading
Reiner Pope – The math behind how LLMs are trained and served
dwarkesh.com · 19,476 words · saved by 3 readers
It's shocking how much you can deduce about what the labs are doing from a handful of equations and a blackboard
Playback speed × Share post Share post at current time Share from 0:00 0:00 / Transcript 210 9 14 Reiner Pope – The math behind how LLMs are trained and served It's shocking how much you can deduce about what the labs are doing from a handful of equations and a blackboard Dwarkesh Patel Apr 29, 2026 210 9 14 Share Transcript Did a very different format with Reiner Pope - a blackboard lecture where he walks through how frontier LLMs are trained and served. It’s shocking how much you can deduce about what the labs are doing from a handful of equations, public API prices, and some chalk. It’s a b
Explore this link on the map →saved by
related reading
- How To Scale Your Modeljax-ml.github.io
- Transformer Inference Arithmetic | kipply's blogkipp.ly
- LLM Inference Performance Engineering: Best Practices | Databricks Blogdatabricks.com
- Transformers Inference Optimization Toolset | AstraBlogastralord.github.io
- Defeating Nondeterminism in LLM Inference - Thinking Machines Labthinkingmachines.ai
- LLM Inference Economics from First Principlestensoreconomics.com
- Transformer Math 101 | EleutherAI Blogblog.eleuther.ai
- LoRA Without Regret - Thinking Machines Labthinkingmachines.ai
- How is LLaMa.cpp possible?finbarr.ca
- Paged Attention from First Principles: A View Inside vLLM – Hamza's Bloghamzaelshafie.bearblog.dev
- A guide to LLM inference and performancebaseten.co
- Real-time LLM Inference on Standard Datacenter GPUs (3,000 tokens/s per request)blog.kog.ai