00:00/00:00
Lecture 25 of 25
Module Summary
Course Content
0 / 25 completedSection 1: Course Overview1 videos
Course Overview
3m
Section 2: Introducing Site Reliability Engineering6 videos
What Is Site Reliability Engineering
8m
Comparing Traditional Ops and SRE
11m
Comparing DevOps and SRE
8m
Exploring the Key Tenets of SRE
5m
Understanding Why SRE Works
4m
Module Summary
2m
Section 3: Automation and Eliminating Toil7 videos
Identifying and Measuring Toil
6m
Engineering Away Toil
7m
Prioritising Toil-reducing Projects
7m
Dealing with the Remaining Toil
3m
Summary
2m
Understanding Toil
6m
Restricting Toil to 50%
4m
Section 4: Service Levels, Monitoring, and Alerting5 videos
Understanding Service Level Objectives and Error Budgets
8m
Defining Service Level Indicators and Service Level Objectives
6m
Monitoring Service Level Indicators
9m
Alerting on Service Level Objectives
7m
Module Summary
4m
Section 5: Incident Management On-call and Postmortems6 videos
What Does On-call Look Like
3m
Managing Incidents Control, Co-ordinate, and Communicate
4m
Working on Incidents Effectively
7m
Producing and Publishing Postmortems
4m
Avoiding Operational Overload
5m
Module Summary
2mNow Playing