A real incident
An AI agent was told to change nothing and deleted a live database instead. The chapters next door hand that kind of authority over deliberately, so a single wrong run cannot break the world.
The chapters
58 min total
- When AI takes action7 minYou inventory every action your product can take, and judge each by what a wrong one costs, not by the error rate alone.
- The action surface7 minYou turn the inventory into a reviewed tool list, specifying each tool by what it does, its limits, and whether its writes undo.
- The autonomy ladder7 minYou place each action on one of five rungs from suggest to act silently, starting below your confidence and promoting only on evidence.
- Blast radius: limit the damage7 minYou bound what one bad turn can reach with scoped tokens, hard caps, separated environments, and a kill switch that stops it mid-action.
- Prompt injection8 minYou learn how outsider content borrows your agent's authority, and close one leg of the lethal trifecta so an attack drops to a nuisance.
- Receipts and recovery7 minYou give every run a plain-language ledger and a real undo behind each write, so a half-finished run becomes a short list a person can fix.
- Agent evals: judge the run7 minYou grade every run twice, once on the result and once on the path it took, against a rubric run in a sandbox that never touches production.
- The Agent Charter8 minYou gather every decision from the part onto one signed page: the job, the tool table, each rung, the caps, the kill switch, and the gating evals.
The sum
Together these chapters take a feature that acts and let you hand it authority you chose: every action scoped, placed, bounded, tested, and recorded on one page that shows exactly what your agent may do and what stops it mid-incident.
The Agent Charter is a one-page fillable PDF with a field per decision; filling it leaves you with a signed record of what your agent may do and what stops it.
The Agent CharterFillable PDFDownload →