✳flâneur — a map of the web's best reading
tinker-nomics
tinker-nomics.vercel.app · 829 words · saved by 1 readers
Interactive RLVR cost model for Tinker
tinker-nomics tinker-nomics Tinker-nomics The economics of post-training By Simon Guo · March 2026 🧮 Tinker vs Self-hosted RL Cost Calculator Pick model, workload, steps, batch size — get a side-by-side breakdown. → Lately, I have been Tinkering a lot and learning about post-training: I can easily build RL pipelines, train on big models that I would have never been able to fit on my tiny Stanford GPUs, plus I got some free credits through the research grant program. . I keep getting the same questions when I try to get my labmates to use it or convince advisors to pay for it: “ Why not just r
Explore this link on the map →saved by
related reading
- Composer2.pdfcursor.com
- RLHF & Post-Training Course by Nathan Lambertrlhfbook.com
- Anatomy of a Modern Finetuning APIbenanderson.work
- Transformer Math 101 | EleutherAI Blogblog.eleuther.ai
- Reinforcement learning is an infrastructure problemmodal.com
- Cheap RL tasks will waste compute | Mechanize, Inc.mechanize.work
- Navigating the High Cost of AI Compute | Andreessen Horowitza16z.com
- State of RL for reasoning LLMs | A. Weersaweers.de
- Navigating the High Cost of AI Compute | Andreessen Horowitza16z.com
- RL Scaling Laws for LLMs - by Cameron R. Wolfe, Ph.D.cameronrwolfe.substack.com
- How Well Does RL Scale? - Toby Ordtobyord.com
- H100 vs GB200 NVL72 Training Benchmarks – Power, TCO, and Reliability Analysis, Software Improvement Over Time – SemiAnalysissemianalysis.com