Autoresearch Is Not About Training Models. It Is About What Happens When Agents Get a Scoreboard | by Krish | Mar, 2026 | AI Sutra
AISutra delivers quality articles on how machine learning and artificial intelligence is shaping our society. This site will help both business leaders and concerned citizens understand its impact across many dimensions Follow publication 3 Listen Share Andrej Karpathy released autoresearch earlier this month. A 630-line Python script. MIT licensed. One GPU, one file, one metric. You point an AI agent at a training setup, go to sleep, and wake up to roughly 100 completed experiments. The agent modified the code, trained for 5 minutes, checked if the result improved, kept or discarded, and repeated. That is the surface description. It is accurate but incomplete. Because what autoresearch actually demonstrates is a pattern that will reshape how foundational models get built, how applications get optimized on top of them, and how enterprise teams think about the cost of experimentation itself. I am writing this post after using Autoresearch to implement recursive self learning to two of m
Autoresearch Is Not About Training Models. It Is About What Happens When Agents Get a Scoreboard Krish 9 min read · Mar 23, 2026 -- Listen Share Press enter or click to view image in full size Andrej Karpathy released autoresearch earlier this month. A 630-line Python script. MIT licensed. One GPU, one file, one metric. You point an AI agent at a training setup, go to sleep, and wake up to roughly 100 completed experiments. The agent modified the code, trained for 5 minutes, checked if the result improved, kept or discarded, and repeated. That is the surface description. It is accurate but inc
Explore this link on the map →saved by
related reading
- Automated Weak-to-Strong Researcheralignment.anthropic.com
- An Apple-Picking Model of AI R&D | Tom Cunningham – Tom Cunninghamtecunningham.github.io
- When AI builds itself \ Anthropicanthropic.com
- AI 2027ai-2027.com
- I Let AI Agents Train Their Own Models. Here's What Actually Happened. | Hamza Mostafahamzamostafa.com
- Demystifying evals for AI agents \ Anthropicanthropic.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- Composer2.pdfcursor.com
- AIDE²: First Evidence of Recursive Self-Improvement | Weco AIweco.ai
- Harness Engineering for Self-Improvement | Lil'Loglilianweng.github.io
- [2603.08640] PostTrainBench: Can LLM Agents Automate LLM Post-Training?arxiv.org
- Lossy self-improvement - by Nathan Lambertinterconnects.ai