Agent behavior
A standard format for describing expected agent behavior for trace review and evaluation. Behavior specs live in .agents/behaviors/ and define what good agent conduct looks like before a reviewer, rubric, scorer, or eval checks for it.
Define what good agent behavior looks like Agent behavior is a format for writing down the behavior you expect an AI agent to follow across many interactions. Each behavior spec is a Markdown file that lives in your repo and describes the recurring conduct that makes the agent reliable. The spec captures that standard up front, so reviewers, rubrics, scorers, and evals have something concrete to measure against. .agents/behaviors/ .agents/behaviors/ └── financial-work-verification/ └── BEHAVIOR.md Written for review Specs speak to the people and agents who read traces, design evals,…
saved by
related reading
- Is Agentic: AI Agent Readiness Score for your Site and Appis-agentic.com
- Agentic Evals Pyramidrwilinski.ai
- Agentationagentation.dev
- Demystifying evals for AI agents \ Anthropicanthropic.com
- Agentationagentation.com
- Agent Observability and Tracingarize.com
- Cookbookcookbook.openai.com
- Amplitude Agent Analyticsamplitude.com
- LangChain: the open agent platform to own your intelligencelangchain.com
- Agent Evaluation: A Detailed Guidecameronrwolfe.substack.com
- Building reliable AI agents · parth sareenparthsareen.com
- GitHub - msitarzewski/agency-agents: A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes, and proven deliverables.github.com