Perform a search on the site.

Your currency

DevOps SRE Foundation℠: DevOps Engineering Fundamentals

Service reliability should be something you can discuss and monitor with development teams. Connect practices, automation and feedback through DevOps to organise a more consistent way of working. Develop a framework for reducing friction and supporting service improvement.

Duration
2 days 16 hours
Code
DEVOPSEF Code
Certification
Site Reliability Engineering Foundation℠ Certification
Site Reliability Engineering Foundation℠

Accredited training for the Site Reliability Engineering Foundation℠ certification.

Presentation

In a constantly evolving IT landscape, businesses face increasing risks. In this context, the site reliability engineer (SRE) plays a crucial role in improving system availability and resilience. Working with development teams within a DevOps approach, the SRE automates operational tasks, establishes monitoring systems and manages incidents. This rapidly growing role offers opportunities for IT professionals.

This DevOps SRE course introduces essential concepts and practices. You will explore SRE foundations and their relationship with DevOps and other methods, then learn to manage service level objectives (SLOs) and error budgets. You will also understand how to reduce toil, establish effective monitoring with SLIs and automate tasks using SRE tools. Finally, you will examine antifragility, SRE's organisational impact and future trends.

This 2-day program also prepares you to take PeopleCert's DevOps SRE Foundation certification exam, included in our offer (see the Certification tab for details). You will gain the skills and knowledge to advance your career in the expanding DevOps and SRE field.

Objectives

By the end of this DevOps SRE course, you will be able to:

  • understand the origins of site reliability engineering (SRE) and its emergence at Google LLC;
  • explain the relationship between SRE, DevOps and other IT engineering methodologies;
  • master site reliability engineering fundamentals;
  • understand service level objectives and their value to customers;
  • identify service level indicators;
  • implement a modern monitoring system;
  • establish error budgets and define error-related strategies;
  • apply good practices to reduce the toil budget;
  • assess the impact of complex tasks on business performance;
  • demonstrate that observability is a determining factor in service quality;
  • use modern site reliability engineering tools and automation good practices;
  • incorporate site antifragility engineering principles;
  • explain the organisational benefits of implementing SRE in a business;
  • take the exam and earn SRE Foundation℠ certification.
Last update: 24/09/2026

Site Reliability Engineering is a service mark of the DevOps Institute.