Skip to content
Handbook navigation

Full-access lesson

Agent Safety

> Agent safety is defense in depth across untrusted input handling, least-privilege capabilities, per-call authorization, approval, sandboxing, budgets, monitoring, containment, and recovery. A toy risk score can demonstrate routing arithmetic; it cannot grant permission or prove an action safe.

The complete curriculum

All 12 volumes, companion resources, interviews, and architecture reviews.

Executable engineering practice

Subscriber-only Python, Java, TypeScript, and SQL labs in the isolated runner.

Agent Safety | KnowledgeOS