Site Reliability Engineering (SRE) Practitioner (SREP)

Course Overview

Introduces a range of practices for advancing service reliability engineering through a mixture of automation, organizational ways of working and business alignment. Tailored for those focused on large-scale service scalability and reliability.
The course fee includes the open book, online Proctored exam. Delegates will receive a voucher for this exam which they can sit, at their convenience, post course.

Moyens Pédagogiques :

Quiz pré-formation de vérification des connaissances (si applicable)
Réalisation de la formation par un formateur agréé par l’éditeur
Formation réalisable en présentiel ou en distanciel
Mise à disposition de labs distants/plateforme de lab pour chacun des participants (si applicable à la formation)
Distribution de supports de cours officiels en langue anglaise pour chacun des participants
- Il est nécessaire d'avoir une connaissance de l'anglais technique écrit pour la compréhension des supports de cours

Moyens d'évaluation :

Quiz pré-formation de vérification des connaissances (si applicable)
Évaluations formatives pendant la formation, à travers les travaux pratiques réalisés sur les labs à l’issue de chaque module, QCM, mises en situation…
Complétion par chaque participant d’un questionnaire et/ou questionnaire de positionnement en amont et à l’issue de la formation pour validation de l’acquisition des compétences

Who should attend

The target audience for the SRE Practitioner course are professionals including:

Anyone focused on large-scale service scalability and reliability
Anyone interested in modern IT leadership and organizational change approaches
Business Managers
Business Stakeholders
Change Agents
Consultants
DevOps Practitioners
IT Directors
IT Managers
IT Team Leaders
Product Owners
Scrum Masters
Software Engineers
Site Reliability Engineers
System Integrators
Tool Providers

Prerequisites

It is highly recommended that learners attend the SRE Foundation course with an accredited DevOps Institute Education Partner prior to attending the SRE Practitioner course. An understanding and knowledge of common SRE terminology, concepts, principles and related work experience are recommended. Please note: the DevOps Institute Site Reliability Engineering (SRE) Foundation (SREF) certification is a prerequisite to the SRE Practitioner exam.

Course Content

Course Introduction

Module 1: SRE Anti-patterns

Rebranding Ops or DevOps or Dev as SRE
Users notice an issue before you do
Measuring until my Edge
False positives are worse than no alerts
Configuration management trap for snowflakes
The Dogpile: Mob incident response
Point fixing
Production Readiness Gatekeeper
Fail-Safe really?

Module 2: SLO is a Proxy for Customer Happiness

Define SLIs that meaningfully measure the reliability of a service from a user’s perspective
Defining System boundaries in a distributed ecosystem for defining correct SLIs
Use error budgets to help your team have better discussions and make better data-driven decisions
Overall, Reliability is only as good as the weakest link on your service graph
Error thresholds when 3rd party services are used

Module 3: Building Secure and Reliable Systems

SRE and their role in Building Secure and Reliable systems
Design for Changing Architecture
Fault tolerant Design
Design for Security
Design for Resiliency
Design for Scalability
Design for Performance
Design for Reliability
Ensuring Data Security and Privacy

Module 4: Full-Stack Observability

Modern Apps are Complex & Unpredictable
Slow is the new down
Pillars of Observability
Implementing Synthetic and End user monitoring
Observability driven development
Distributed Tracing
What happens to Monitoring?
Instrumenting using Libraries an Agents

Module 5: Platform Engineering and AIOPs

Taking a Platform Centric View solves Organisational scalability challenges such as fragmentation, inconsistency and unpredictability.
How do you use AIOps to improve Resiliency
How can DataOps help you in the journey
A simple recipe to implement AIOps
Indicative measurement of AIOps

Module 6: SRE & Incident Response Management

SRE Key Responsibilities towards incident response
DevOps & SRE and ITIL
OODA and SRE Incident Response
Closed Loop Remediation and the Advantages
Swarming – Food for Thought
AI/ML for better incident management

Module 7: Chaos Engineering

Navigating Complexity
Chaos Engineering Defined
Quick Facts about Chaos Engineering
Chaos Monkey Origin Story
Who is adopting Chaos Engineering
Myths of Chaos
Chaos Engineering Experiments
GameDay Exercises
Security Chaos Engineering
Chaos Engineering Resources

Module 8: SRE is the Purest form of DevOps

Key Principles of SRE
SREs help increase Reliability across the product spectrum
Metrics for Success
Selection of Target areas
SRE Execution Model
Culture and Behavioral Skills are key
SRE Case study

Post-class assignments/exercises

Non-abstract Large Scale Design (after Day 1)
Observability and Monitoring (after Day 2)
Chaos Engineering Instrumentation

Prix & Delivery methods

Formation en ligne

Durée
3 jours

Prix

sur demande

Dates et Inscription

Demande de date

Formation en salle équipée

Durée
3 jours

Prix

sur demande

Dates et Inscription

Demande de date

Agenda

Délai d’accès – inscription possible jusqu’à la date de formation

Instructor-led Online Training : Cours en ligne avec instructeur If you have any questions about our online courses, feel free to contact us via phone or Email anytime.

Anglais

Fuseau horaire : Heure d'été d'Europe centrale (HAEC) ±1 heure

24.09. – 26.09.2025 Formation en ligne Fuseau horaire : Heure d'été d'Europe centrale (HAEC)

Allemand

Fuseau horaire : Heure d'été d'Europe centrale (HAEC) ±1 heure

09.07. – 11.07.2025 Formation en ligne Fuseau horaire : Heure d'été d'Europe centrale (HAEC)

10.12. – 12.12.2025 Formation en ligne Fuseau horaire : Heure normale d'Europe centrale (HNEC)

Modalités de financement

Handicap

Site Reliability Engineering (SRE) Practitioner (SREP)

Course Overview

Moyens Pédagogiques :

Moyens d'évaluation :

Who should attend

Prerequisites

Course Content

Course Introduction

Module 1: SRE Anti-patterns

Module 2: SLO is a Proxy for Customer Happiness

Module 3: Building Secure and Reliable Systems

Module 4: Full-Stack Observability

Module 5: Platform Engineering and AIOPs

Module 6: SRE & Incident Response Management

Module 7: Chaos Engineering

Module 8: SRE is the Purest form of DevOps

Post-class assignments/exercises

Prix & Delivery methods

Formation en ligne

Prix

Formation en salle équipée

Prix

Agenda

Anglais

Fuseau horaire : Heure d'été d'Europe centrale (HAEC) ±1 heure

Allemand

Fuseau horaire : Heure d'été d'Europe centrale (HAEC) ±1 heure

Modalités de financement

Handicap

Site Reliability Engineering (SRE) Practitioner (SREP)

Course Overview

Moyens Pédagogiques :

Moyens d'évaluation :

Who should attend

Prerequisites

Course Content

Course Introduction

Module 1: SRE Anti-patterns

Module 2: SLO is a Proxy for Customer Happiness

Module 3: Building Secure and Reliable Systems

Module 4: Full-Stack Observability

Module 5: Platform Engineering and AIOPs

Module 6: SRE & Incident Response Management

Module 7: Chaos Engineering

Module 8: SRE is the Purest form of DevOps

Post-class assignments/exercises

Prix & Delivery methods

Formation en ligne

Prix

Formation en salle équipée

Prix

Agenda

Anglais

Fuseau horaire : Heure d'été d'Europe centrale (HAEC) ±1 heure

Allemand

Fuseau horaire : Heure d'été d'Europe centrale (HAEC) ±1 heure

Fuseau horaire : Heure d'été d'Europe centrale (HAEC) ±1 heure

Fuseau horaire : Heure d'été d'Europe centrale (HAEC) ±1 heure