Thank you for sending your enquiry! One of our team members will contact you shortly.
Thank you for sending your booking! One of our team members will contact you shortly.
Duration 35 hours
Course Outline
SRE Anti-patterns
- Identifying counterproductive practices
- Understanding the impact of anti-patterns on system reliability
- Adopting best practices and corrective alternatives
SLOs as a Proxy for Customer Satisfaction
- Defining Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
- Managing error budgets to balance innovation with reliability
- Understanding the limits of distributed systems
Building Secure and Reliable Systems
- Designing for fault tolerance and resilience
- Integrating security into reliability engineering
- Implementing scalability and data protection strategies
Full-stack Observability
- Instrumentation and metrics collection
- Distributed tracing and synthetic monitoring
- Adopting an observability-driven development approach
Platform Engineering and AIOps
- Platform-centric engineering approaches
- Automation and orchestration within SRE
- Leveraging DataOps and operational intelligence
Incident Management in SRE
- Defining roles and responsibilities in incident response
- Applying frameworks such as OODA
- Utilising automated remediation and AI/ML-assisted resolution
Chaos Engineering
- Principles and strategies for resilience testing
- Planning and executing “game day” exercises
- Learning from controlled failure experiments
SRE as a Pure Form of DevOps
- Integrating SRE into DevOps workflows
- Aligning culture and collaboration practices
- Driving organisational transformation through SRE
Post-class Exercises
- Large-scale system design case studies
- Advanced instrumentation and monitoring scenarios
- Real-world reliability problem-solving
Review and Exam Preparation
- Final review of the DevOps Institute SRE Practitioner syllabus
- Sample questions and practice tests
- Exam-taking strategies and recommendations
Summary and Next Steps
Requirements
- A solid grasp of core Site Reliability Engineering principles
- Practical experience with DevOps practices and associated tooling
- Proficiency in system monitoring, incident management, and automation
Intended Audience
- SRE professionals preparing for the DevOps Institute SRE Practitioner certification
- DevOps engineers looking to transition into reliability-focused roles
- Operations leaders accountable for reliability strategy and execution
Testimonials (2)
Craig was extremely involved in the training, always making sure we are paying attention, adapted the examples to our day-to-day activities and always provided an answer when asked, even if the information was not added in the presentation.
Ecaterina Ioana Nicoale - BOOKING HOLDINGS ROMANIA SRL
Course - DevOps Foundation®
High level of commitment and knowledge of the trainer