✳flâneur — a map of the web's best reading
AI excels at code competitions, struggles with real work
blog.peterwildeford.com · 1,204 words · saved by 1 readers
What CodeForces rankings reveal about AI capabilities
AI excels at code competitions, struggles with real work What CodeForces rankings reveal about AI capabilities Peter Wildeford Feb 17, 2025 12 3 Share Photo by Chris Ried on Unsplash Author’s note: This analysis was originally part of a previous post, but it was buried at the bottom and I wanted to feature it more prominently. So I removed it from there, expanded it, and put it here. But sorry if you’re seeing this twice. ~ When IBM's Deep Blue beat chess champion Garry Kasparov in 1997, it was a notable milestone in the development of AI. Today, AI systems are beginning to achieve similar fea
Explore this link on the map →saved by
related reading
- [2501.01257] CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratingsarxiv.org
- MirrorCode: Evidence AI can already do some weeks-long coding tasks | Epoch AIepoch.ai
- [2203.07814] Competition-Level Code Generation with AlphaCodearxiv.org
- Competitive programming with AlphaCode — Google DeepMinddeepmind.com
- Composer2.pdfcursor.com
- My picture of the present in AI — LessWronglesswrong.com
- Don't fall into the anti-AI hype - <antirez>antirez.com
- pradyuprasad.compradyuprasad.com
- The Future of Programming: Copilots vs. Agents (Part I)eastwind.substack.com
- AI progress is about to speed up | Epoch AIepoch.ai
- After Automation | Everyevery.to
- Measuring the Impact of Early-2025 AI on Experienced Open-Source Developer Productivity - METRmetr.org