Is there a Half-Life for the Success Rates of AI Agents? — Toby Ord
Building on the recent empirical work of Kwa et al. (2025), I show that within their suite of research-engineering tasks the performance of AI agents on longer-duration tasks can be explained by an extremely simple mathematical model — a constant rate of failing during each minute a human would take
Is there a Half-Life for the Success Rates of AI Agents? May 7, 2025 Toby Ord Building on the recent empirical work of Kwa et al. (2025), I show that within their suite of research-engineering tasks the performance of AI agents on longer-duration tasks can be explained by an extremely simple mathematical model — a constant rate of failing during each minute a human would take to do the task. This implies an exponentially declining success rate with the length of the task and that each agent could be characterised by its own half-life. This empirical regularity allows us to estimate the success
Explore this link on the map →saved by
related reading
- Measuring AI Ability to Complete Long Tasks - METRmetr.org
- A (Slightly) Mechanistic Theory for Exponentially Increasing AI Time Horizons? — LessWronglesswrong.com
- My picture of the present in AI — LessWronglesswrong.com
- AI 2027ai-2027.com
- We spent 2 hours working in the future - METRmetr.org
- How Does Time Horizon Vary Across Domains? - METRmetr.org
- I underestimated AI capabilities (again) - by Ajeya Cotraplanned-obsolescence.org
- Clarifying and predicting AGI — LessWronglesswrong.com
- An Apple-Picking Model of AI R&D | Tom Cunningham – Tom Cunninghamtecunningham.github.io
- Humans Still Beat AI in the Long Horizon: Revisiting Test-Time Scaling in the Agent Era | Qiuyang Mangjoyemang33.github.io
- After Automation | Everyevery.to
- AI 2027ai-2027.com