The Mismanaged Geniuses Hypothesis | Alex L. Zhang
We propose the mismanaged geniuses hypothesis, which posits that existing frontier language models are severely underutilized due to sub-optimal use of individual language model calls.
Alex Zhang, Zhening (Zed) Li, and Omar Khattab. tldr; AI models are already good enough for the next leap in capabilities. For the last decade, scaling the size and data of AI models has led to groundbreaking, super-human achievements in the capabilities of these systems. The recent success of RL and reasoning in particular implies that models can be trained to generalize on tasks we have never even solved ourselves. It is natural to believe that continuing this trend of scaling across a single neural model will be the recipe that gets us to the next jump in AI capabilities. We have an…
saved by
related reading
- alex zhang (@a1zhang) on Xx.com
- Language model harnesses are compositional generalizersalexzhang13.github.io
- AI in 2025: gestalt — LessWronglesswrong.com
- As Rocks May Think | Eric Jangevjang.com
- machine learning imindslice.substack.com
- The Extreme Inefficiency of RL for Frontier Models - Toby Ordtobyord.com
- Language Models will be Scaffoldsalexzhang13.github.io
- GenAI Handbookgenai-handbook.github.io
- Scaling: The State of Play in AIoneusefulthing.org
- the-illusion-of-thinking.pdfml-site.cdn-apple.com
- To Understand Language is to Understand Generalization | Eric Jangevjang.com
- Artificial General Intelligence Is Already Herenoemamag.com