Thank you for sending your enquiry! One of our team members will contact you shortly.
Thank you for sending your booking! One of our team members will contact you shortly.
Duration 35 hours
Course Outline
SRE Anti-patterns
- Spotting counterproductive operational practices
- Assessing the impact of anti-patterns on system reliability
- Adopting best practices and corrective measures
SLOs as a Proxy for Customer Satisfaction
- Establishing Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
- Overseeing error budgets while balancing innovation against reliability
- Comprehending the inherent limits of distributed systems
Building Secure and Reliable Systems
- Engineering for fault tolerance and resilience
- Embedding security within reliability engineering processes
- Strategies for scalability and data protection
Full-stack Observability
- Implementation of instrumentation and metrics gathering
- Utilizing distributed tracing and synthetic monitoring
- Adopting observability-driven development lifecycles
Platform Engineering and AIOps
- Employing platform-centric engineering models
- Automation and orchestration within SRE contexts
- Utilizing DataOps and operational intelligence
Incident Management in SRE
- Clarifying roles and duties in incident response
- Implementing frameworks such as OODA
- Automated remediation and AI/ML-assisted problem resolution
Chaos Engineering
- Core principles and approaches for resilience testing
- Planning and running “game day” simulations
- Deriving insights from controlled failure experiments
SRE as a Pure Form of DevOps
- Integrating SRE principles into DevOps workflows
- Fostering cultural alignment and collaborative practices
- Leading organizational transformation via SRE
Post-class Exercises
- Case studies focused on large-scale system architecture
- Complex instrumentation and monitoring scenarios
- Resolution of real-world reliability challenges
Review and Exam Preparation
- Comprehensive review of the DevOps Institute SRE Practitioner syllabus
- Working through sample questions and mock tests
- Exam-taking tactics and strategic advice
Summary and Next Steps
Requirements
- A solid grasp of fundamental Site Reliability Engineering principles
- Practical experience with DevOps methodologies and associated tooling
- Proficiency in system monitoring, incident management, and automation
Target Audience
- SRE professionals pursuing the DevOps Institute SRE Practitioner certification
- DevOps engineers looking to transition into reliability-centric roles
- Operations leaders tasked with defining and executing reliability strategies
Testimonials (2)
Craig was extremely involved in the training, always making sure we are paying attention, adapted the examples to our day-to-day activities and always provided an answer when asked, even if the information was not added in the presentation.
Ecaterina Ioana Nicoale - BOOKING HOLDINGS ROMANIA SRL
Course - DevOps Foundation®
High level of commitment and knowledge of the trainer