Get in Touch
 Duration 35 hours

Course Outline

SRE Anti-patterns

  • Spotting counterproductive operational practices
  • Assessing the impact of anti-patterns on system reliability
  • Adopting best practices and corrective measures

SLOs as a Proxy for Customer Satisfaction

  • Establishing Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
  • Overseeing error budgets while balancing innovation against reliability
  • Comprehending the inherent limits of distributed systems

Building Secure and Reliable Systems

  • Engineering for fault tolerance and resilience
  • Embedding security within reliability engineering processes
  • Strategies for scalability and data protection

Full-stack Observability

  • Implementation of instrumentation and metrics gathering
  • Utilizing distributed tracing and synthetic monitoring
  • Adopting observability-driven development lifecycles

Platform Engineering and AIOps

  • Employing platform-centric engineering models
  • Automation and orchestration within SRE contexts
  • Utilizing DataOps and operational intelligence

Incident Management in SRE

  • Clarifying roles and duties in incident response
  • Implementing frameworks such as OODA
  • Automated remediation and AI/ML-assisted problem resolution

Chaos Engineering

  • Core principles and approaches for resilience testing
  • Planning and running “game day” simulations
  • Deriving insights from controlled failure experiments

SRE as a Pure Form of DevOps

  • Integrating SRE principles into DevOps workflows
  • Fostering cultural alignment and collaborative practices
  • Leading organizational transformation via SRE

Post-class Exercises

  • Case studies focused on large-scale system architecture
  • Complex instrumentation and monitoring scenarios
  • Resolution of real-world reliability challenges

Review and Exam Preparation

  • Comprehensive review of the DevOps Institute SRE Practitioner syllabus
  • Working through sample questions and mock tests
  • Exam-taking tactics and strategic advice

Summary and Next Steps

Requirements

  • A solid grasp of fundamental Site Reliability Engineering principles
  • Practical experience with DevOps methodologies and associated tooling
  • Proficiency in system monitoring, incident management, and automation

Target Audience

  • SRE professionals pursuing the DevOps Institute SRE Practitioner certification
  • DevOps engineers looking to transition into reliability-centric roles
  • Operations leaders tasked with defining and executing reliability strategies

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories