Running agents in loops and pipelines with limited permissions, budgets and approval points, so their mistakes stay small.
- Agent Loops and Multi-Agent Orchestration
- Planners, workers and reviewers, and when adding agents improves reliability and when it only adds noise.
- Sandboxing and Least Privilege for Agents
- Restricting file, network and credential access so an agent's worst possible action is survivable.
- Separating Reading From Acting
- Designing workflows so an agent that reads untrusted content cannot take consequential actions on its own.
- Human Approval Points
- Choosing which actions, such as deploys, deletions, payments and external messages, need a person, and making that approval meaningful.
- Budgets, Timeouts and Kill Switches
- Limits on steps, time and spend, and a reliable way to stop everything at once.
- Audit Trails
- Recording what each agent did, with which permissions and on whose instruction.