Operating an agent is an ongoing design activity. Use this workbook to connect traces, exceptions, security events, recovery outcomes, and cost observations to the next improvement.

1. Make the workflow observable

Record the workflow version, dependencies, owners, and escalation route. Define completion, intervention, failure, recovery, latency, and cost metrics for a stated observation window.

  • Define each metric’s denominator and time window.
  • Distinguish successful recovery from repeated actions.
  • Separate measured values from estimates.

2. Turn findings into decisions

For each material observation, capture evidence, impact, owner, action, and due date. Assess security findings against the permissions and threat model of the affected workflow.

  • Document containment decisions and affected access.
  • Track recovery state and unresolved work.
  • Route decisions to the responsible owner.

3. Feed operations back into design

Link each proposed improvement to a requirement, control, skill, context rule, or evaluation case. Review material changes through the lifecycle before promoting them.

  • Add a regression case for meaningful failures.
  • Review changes to tools, access, knowledge, and memory.
  • Verify the improvement against its original evidence.

Bring the decisions together.

Connect this record to your specifications, controls, evaluations, and operational evidence. Revisit it when the workflow changes.

Explore the complete lifecycle