Framework · spec sheet
Agent Zero
- Category
- Agentic framework
- Type
- General-purpose
- License
- Open source
- Languages
- Python
- Focus
- Computer-as-tool, self-written code
- Best for
- Fully customisable autonomous agents
General-purpose agentic framework that treats the whole computer as its tool and writes its own code to finish a task. Nothing hard-coded; you shape it as you go.
What it is
Agent Zero is an open-source, general-purpose agent framework in Python, built around one stubborn idea: do not hand the model a fixed set of tools at all. It writes its own code to get through whatever is in front of it, and treats the machine it runs on — shell, files, everything it can reach — as the working surface. That is what the focus line in the spec above means in practice: where most frameworks give a model named functions to call, Agent Zero gives it an interpreter.
It is a community project, not a company product. There is no hosted tier and no paid wrapper around the licence; you clone it, run it, and read the source when something behaves oddly. As of late September 2026 that is still the shape of it, though with a project at this pace the repository is the only reliable statement of its current state.
How it works
The loop is a conversation. The agent receives a task, decides what code would move it forward, runs that code, reads the output and repeats. Because the code is written per task rather than installed in advance, the same agent can scan a directory of logs in one turn and drive a browser the next. Around that loop sit three things that make it hold up across sessions: memory, a store of earlier conversations and knowledge files the agent can draw on; editable prompts, meaning the persona and standing instructions live in files you rewrite yourself — this is what “you shape it as you go” in the note above actually refers to; and sub-agents, which let the main agent delegate a piece of work and receive a result back instead of a whole transcript.
The documented deployment is a Docker container, and that is not incidental. An agent that writes and runs arbitrary code with the whole computer as its tool needs a box around it, and the container is the only boundary between a curious task and your actual machine.
When it earns its place over a plain loop
When the next step is not one of the tools you anticipated. A plain loop — prompt, tool call, result, repeat — is bounded by its tool list, and some tasks fall outside it. Agent Zero's answer is to let the model write the missing tool. It also earns its keep as a learning platform: because persona, memory and behaviour are plain editable text, it is a fast way to find out what agent instructions really do under load.
It does not earn its place in a production service. There is no typed API surface and no object model to assert against, and a self-writing agent is far harder to pin down than one calling named functions. If you need behaviour you can test, look at Pydantic AI.
Limits
- Self-written code is self-written risk. Every task can produce new, unreviewed code. The container reduces the blast radius; it does not make the code legible.
- Behaviour is prompt-shaped. Editable prompts are powerful and brittle — a wording change alters the agent's habits in ways no test suite is watching for.
- Improvisation is the feature and the failure mode. Less structure means less predictability, and long runs wander.
- Verify before you build on it. The project moves quickly and I have not checked its 2026 changes against anything but the spec card above; treat the repository README as authoritative over this page.
Alternatives on this board
- CrewAI — when the work splits into named roles rather than improvised scripts.
- LangGraph — when the control flow is a graph you want written down.
- Pydantic AI — typed tools and outputs instead of free-form code.
- Claude Code — a similar “whole machine as the tool” posture, but a maintained product with a permission layer.
Sources
- Agent Zero repository — the place to verify current features, deployment steps and licence text.
- Pick a framework or pick a loop — when a plain loop is enough.
- Framework leaderboard for the full comparison.
Hand-maintained editorial spec, not vendor copy — the read on each tool is judgement. Last checked 16 Sep 2026 · back to frameworks.