✳flâneur — a map of the web's best reading
AI infrastructure in the "Era of experience"
tensoreconomics.com · 10,328 words · saved by 4 readers
Intelligence involution, economies of scale in RL, everything async and multi-turn.
AI infrastructure in the "Era of experience" Intelligence involution, economies of scale in RL, everything async and multi-turn. Piotr Mazurek Nov 26, 2025 27 4 2 Share In the famous essay from May 2025, “ Welcome to the Era of Experience ,” Rich Sutton and David Silver proposed a new paradigm of training AI models - models that learn not through predicting the next word against text scraped from Common Crawl, but through gaining experience via interaction with environments. As we approach the exhaustion of easily scrapable text data , we predict we’ll observe a shift toward AI models increasi
Explore this link on the map →saved by
related reading
- LoRA Without Regret - Thinking Machines Labthinkingmachines.ai
- State of RL for reasoning LLMs | A. Weersaweers.de
- Reiner Pope – The math behind how LLMs are trained and serveddwarkesh.com
- How We Build Trillion Parameter Reasoning RL with 10% GPUsmacaron.im
- Composer2.pdfcursor.com
- RL Scaling Laws for LLMs - by Cameron R. Wolfe, Ph.D.cameronrwolfe.substack.com
- RL at 1T Scale: prime-rl Performance Deep Diveprimeintellect.ai
- LLM Inference Performance Engineering: Best Practices | Databricks Blogdatabricks.com
- Orbit - Ultra-efficient RL Pipelinespherelab.ai
- Reinforcement learning is an infrastructure problemmodal.com
- How Well Does RL Scale? - Toby Ordtobyord.com
- IsoCompute Playbook: Optimally Scaling Sampling Compute for RL Training of LLMscompute-optimal-rl-llm-scaling.github.io