Hidden Technical Debt of AI Systems: Agent Runtime
Everyone talks about the model, the prompts, the evaluations. Almost nobody talks about the runtime the agent actually runs in. The runtime is the agent. Scu...
Eleven years ago, Sculley et al. drew the diagram everyone in MLOps has seen: a tiny black box labeled “ML Code” surrounded by a sprawl of much larger boxes — data collection, feature extraction, configuration, monitoring, serving infrastructure. The point of the diagram was that the model code is the smallest piece of a real ML system, and that everything else is where the technical debt accumulates. The same diagram is being redrawn for agents. The agentic model call is the small box. The largest box to the right that is currently driving most of the spend and influencing how system architec
saved by
related reading
- How to Harness Coding Agents with the Right Infrastructure | Blogalexlavaee.me
- Demystifying evals for AI agents \ Anthropicanthropic.com
- Together AI | The AI Native Cloudtogether.ai
- AddyOsmani.com - Long-running Agentsaddyosmani.com
- What is an AI code sandbox?modal.com
- Why We Built Our Own Background Agentbuilders.ramp.com
- Effective harnesses for long-running agents \ Anthropicanthropic.com
- Sail Researchsailresearch.com
- The Age of Async Agents — Cognition's Walden Yan & OpenInspect's Cole Murraylatent.space
- Low Latency and Model Training at Modalrhea24.github.io
- My picture of the present in AI — LessWronglesswrong.com
- Scaling Managed Agents: Decoupling the brain from the hands \ Anthropicanthropic.com