General Agent: A Self-Evolving, Synthetic Agent Environment
We are open-sourcing general-agent, a fully synthetic environment that evolves its own task corpus across 1,040 domains, 4,504 tasks, and more than 8,000 unique tools for training and evaluating capable agents.
General Agent: A Self-Evolving, Synthetic Agent Environment Training capable agents requires exposure to diverse tasks and tools throughout the whole post-training pipeline. Yet, agentic environments with exposure to 1000s of tools remain scarce in the open-source community. Today, we are open-sourcing a first version of the general-agent environment ( Environments Hub ) — a fully synthetic environment capable of growing its task corpus to be more diverse and challenging over time. It formulates synthetic task creation as a 2-player game between two agents: Synthesizer — An agent tasked to syn
Explore this link on the map →related reading
- LLM Powered Autonomous Agents | Lil'Loglilianweng.github.io
- Demystifying evals for AI agents \ Anthropicanthropic.com
- Agent Evaluation: A Detailed Guidecameronrwolfe.substack.com
- Building reliable AI agents · parth sareenparthsareen.com
- Building Effective AI Agents \ Anthropicanthropic.com
- Building Effective AI Agents \ Anthropicanthropic.com
- A Taxonomy of RL Environments for LLM Agentsleehanchung.github.io
- Machine Studying | Jacob Xiaochen Lijacobxli.com
- PostTrainBenchposttrainbench.com
- Cookbookcookbook.openai.com
- Arjun Virkarjunvirk.com
- AI Agent Benchmark for Real-World Professional Workflowsagents-last-exam.org