Continually improving our agent harness · Cursor
Models need a harness to become fully useful, and improving that harness means continuously iterating on context, evaluation, and model-specific tuning.
Blog / research We approach building the Cursor agent harness the way we'd approach any ambitious software product. Much of the work is vision-driven, where we start with an opinion about what the ideal agent experience should look like. From there, we form hypotheses about how to get closer to that vision, run experiments to test them, and iterate using quantitative and qualitative signals from evals and real usage. That process depends on having the right online and offline instrumentation, so we can tell when a change actually makes the harness better. When we get early access to new models
saved by
related reading
- The third era of AI software developmentcursor.com
- Best practices for coding with agents · Cursorcursor.com
- Arjun Virkarjunvirk.com
- Context Engineering for AI Agents: Lessons from Building Manusmanus.im
- Harness Engineering for Self-Improvement | Lil'Loglilianweng.github.io
- Cursor: AI coding agentcursor.com
- Effective context engineering for AI agents \ Anthropicanthropic.com
- What is an Agent Harness?rubriclabs.com
- Why We Built Our Own Background Agentbuilders.ramp.com
- Effective harnesses for long-running agents \ Anthropicanthropic.com
- Building a harness for open modelsparthsareen.com
- Dynamic context discovery · Cursorcursor.com