Get in Touch

Course Outline

Introduction

  • How SRE integrates traditional IT with software development practices.
  • The necessity for automation and observability.
  • Distinguishing the roles of software engineers versus system administrators.
  • Comparing Site Reliability Engineers and DevOps engineers.

IT System Overview

  • System architecture across on-premise and cloud environments.

Foundational SRE Principles and Practices

  • Infrastructure as Code (IaC).
  • The impact of containerisation and orchestration (e.g., Docker, Kubernetes).
  • Continuous Integration, Continuous Deployment, and Continuous Delivery.
  • Observability.

Assessing IT Systems

  • Evaluating team and organisational resources.
  • Mapping out existing systems and processes.
  • Estimating the potential impact of SRE adoption.
  • Defining the role of the software engineering team.
  • Defining the role of the operational team.
  • The role of management.

Maintaining System Reliability

  • Defining and measuring desired service reliability.
  • Understanding Service Level Objectives (SLOs).
  • Understanding Service Level Indicators (SLIs) and Service Level Agreements (SLAs).
  • Managing Error Budgets.
  • Formulating an SLO.

Optimising System Administration

  • Setting up a development environment.
  • Evaluating SRE tools.
  • Prioritising tasks for automation.
  • Writing software.

Implementing "Infrastructure as Code"

  • Testing and iterating code.
  • Building anti-fragile systems.
  • Learning from failures.

System Monitoring

  • Observing system performance.
  • SRE tools and techniques.

The Future of SRE

Requirements

  • A foundational understanding of IT infrastructure.
  • A general grasp of the software development lifecycle.
  • Experience in programming or scripting with any language.

Target Audience

  • Developers
  • System Administrators
  • Software Architects
  • DevOps Engineers
  • IT Managers
 21 Hours

Testimonials (7)

Upcoming Courses

Related Categories