Lessons

1Separate naming, routing, and application failures25 min read

Use layered network evidence to localize a service failure.

  • →Use layered network evidence to localize a service failure
Read lesson
2Plan maintenance without confusing it with failure tolerance25 min read

Calculate disruption allowance and surviving service capacity.

  • →Calculate disruption allowance and surviving service capacity
  • →Separate voluntary disruption policy from fault tolerance
Read lesson
3Design a safe self-service scaling API25 min read

Specify authorization and concurrency for a platform mutation.

  • →Specify authorization and concurrency for a platform mutation
  • →Design safe concurrent changes to a platform API
Read lesson
4Present an incident review that changes the next response25 min read

Build a causal incident account with measurable prevention.

  • →Build a causal incident account with measurable prevention
Read lesson

Skills in this course

  1. 01Use layered network evidence to localize a service failureUse layered network evidence to localize a service failure.
  2. 02Calculate disruption allowance and surviving service capacityCalculate disruption allowance and surviving service capacity.
  3. 03Specify authorization and concurrency for a platform mutationSpecify authorization and concurrency for a platform mutation.
  4. 04Build a causal incident account with measurable preventionBuild a causal incident account with measurable prevention.
  5. 05Separate voluntary disruption policy from fault toleranceSeparate voluntary disruption policy from fault tolerance.
  6. 06Design safe concurrent changes to a platform APIDesign safe concurrent changes to a platform API.