Get in Touch

Course Outline

Introduction

  • The integration of SRE with traditional IT and software development.
  • The importance of automation and observability.
  • Distinguishing the roles of software engineers and system administrators.
  • Comparing Site Reliability Engineers and DevOps Engineers.

IT System Overview

  • System architecture in on-premise and cloud environments.

SRE Principles and Practices

  • Infrastructure as Code.
  • The role of containerization and orchestration (e.g., Docker, Kubernetes).
  • Continuous Integration, Continuous Deployment, and Continuous Delivery.
  • Observability.

Evaluating an IT System

  • Assessing team and organizational resources.
  • Mapping existing systems and processes.
  • Estimating the potential impact of implementing SRE.
  • The role of the software engineering team.
  • The role of the operational team.
  • The role of management.

Maintaining System Reliability

  • Defining and measuring desired service reliability.
  • Understanding Service Level Objectives (SLOs).
  • Comprehending Service Level Indicators (SLIs) and Service Level Agreements (SLAs).
  • Utilizing Error Budgets.
  • Formulating an SLO.

Optimizing System Administration

  • Configuring the development environment.
  • Evaluating SRE tools.
  • Prioritizing tasks for automation.
  • Writing software.

Deploying Infrastructure as Code

  • Testing and iterating code.
  • Making systems resilient to failure.
  • Leveraging failures for learning.

System Monitoring

  • Observing system performance.
  • SRE tools and techniques.

The Future of SRE

Requirements

  • A foundational understanding of IT infrastructure.
  • Basic familiarity with the software development lifecycle.
  • Experience with programming or scripting in any language.

Target Audience

  • Developers
  • System Administrators
  • Software Architects
  • DevOps Engineers
  • IT Managers
 21 Hours

Number of participants


Price per participant

Testimonials (7)

Upcoming Courses

Related Categories