Lessons
1Separate naming, routing, and application failures25 min read
Use layered network evidence to localize a service failure.
- →Use layered network evidence to localize a service failure
2Plan maintenance without confusing it with failure tolerance25 min read
Calculate disruption allowance and surviving service capacity.
- →Calculate disruption allowance and surviving service capacity
- →Separate voluntary disruption policy from fault tolerance
3Design a safe self-service scaling API25 min read
Specify authorization and concurrency for a platform mutation.
- →Specify authorization and concurrency for a platform mutation
- →Design safe concurrent changes to a platform API
4Present an incident review that changes the next response25 min read
Build a causal incident account with measurable prevention.
- →Build a causal incident account with measurable prevention
Skills in this course
- 01Use layered network evidence to localize a service failureUse layered network evidence to localize a service failure.
- 02Calculate disruption allowance and surviving service capacityCalculate disruption allowance and surviving service capacity.
- 03Specify authorization and concurrency for a platform mutationSpecify authorization and concurrency for a platform mutation.
- 04Build a causal incident account with measurable preventionBuild a causal incident account with measurable prevention.
- 05Separate voluntary disruption policy from fault toleranceSeparate voluntary disruption policy from fault tolerance.
- 06Design safe concurrent changes to a platform APIDesign safe concurrent changes to a platform API.