Get in Touch

Course Outline

Introduction
  • How SRE integrates traditional IT with software development.
  • The necessity for automation and observability.
  • Comparing the roles of software engineers and system administrators.
  • Differences between Site Reliability Engineers and DevOps Engineers.

Overview of IT Systems
  • System architecture, both on-premise and cloud-based.

Overview of SRE Principles and Practices
  • Infrastructure as Code.
  • The role of containerisation and orchestration (Docker, Kubernetes, etc.).
  • Continuous Integration, Continuous Deployment, and Continuous Delivery.
  • Observability.

Evaluating an IT System
  • Assessing team and organisational resources.
  • Mapping out systems and processes.
  • Estimating the potential impact of SRE.
  • The role of the software engineering team.
  • The role of the operations team.
  • The role of management.

Maintaining System Reliability
  • Defining and measuring desired service reliability.
  • Understanding Service Level Objectives (SLOs).
  • Understanding Service Level Indicators (SLIs) and Service Level Agreements (SLAs).
  • Utilising Error Budgets.
  • Developing an SLO.

Optimising System Administration
  • Setting up a development environment.
  • Evaluating SRE tools.
  • Prioritising tasks for automation.
  • Writing software.

Deploying “Infrastructure as Code”
  • Testing and iterating on code.
  • Making systems anti-fragile.
  • Learning from failure.

Monitoring a System
  • Observing system performance.
  • SRE tools and techniques.

The Future of SRE

Requirements

  • A foundational understanding of IT infrastructure.
  • General familiarity with the software development lifecycle.
  • Experience in programming or scripting in any language.

Audience

  • Developers
  • System Administrators
  • Software Architects
  • DevOps Engineers
  • IT Managers
 21 Hours

Number of participants


Price per participant

Testimonials (7)

Provisional Upcoming Courses (Require 5+ participants)

Related Categories