Thank you for sending your enquiry! One of our team members will contact you shortly.
Thank you for sending your booking! One of our team members will contact you shortly.
Duration 35 hours
Course Outline
SRE Anti-patterns
- Detecting counterproductive workflows
- Assessing the impact of anti-patterns on system reliability
- Establishing best practices and effective corrective alternatives
SLOs as a Metric for Customer Satisfaction
- Establishing Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
- Overseeing error budgets to balance innovation against reliability needs
- Comprehending the inherent limitations of distributed systems
Creating Secure and Reliable Systems
- Engineering for fault tolerance and resilience
- Embedding security within reliability engineering frameworks
- Implementing scalability and robust data protection strategies
Comprehensive Full-stack Observability
- Instrumentation techniques and metrics acquisition
- Utilizing distributed tracing and synthetic monitoring
- Adopting an observability-driven development approach
Platform Engineering and AIOps
- Embracing platform-centric engineering methodologies
- Enhancing automation and orchestration within SRE
- Harnessing DataOps and operational intelligence
Incident Management in SRE
- Clarifying roles and responsibilities in incident response
- Applying established frameworks such as OODA
- Implementing automated remediation and AI/ML-assisted resolution techniques
Chaos Engineering
- Core principles and strategies for resilience testing
- Organizing and executing "game day" scenarios
- Extracting insights from controlled failure experiments
SRE as a Refined Form of DevOps
- Seamlessly integrating SRE into DevOps workflows
- Fostering cultural alignment and collaborative practices
- Driving organizational transformation through SRE initiatives
Post-class Exercises
- Case studies focusing on large-scale system design
- Scenarios involving advanced instrumentation and monitoring
- Practical real-world reliability problem-solving
Revision and Exam Readiness
- Comprehensive final review of the DevOps Institute SRE Practitioner syllabus
- Engagement with sample questions and practice examinations
- Effective exam-taking strategies and expert recommendations
Summary and Future Pathways
Requirements
- Solid grasp of fundamental Site Reliability Engineering principles
- Proven experience with DevOps methodologies and associated toolchains
- Competence in system monitoring, incident management, and automation workflows
Target Audience
- SRE professionals aiming for the DevOps Institute SRE Practitioner certification
- DevOps engineers looking to transition into reliability-centric roles
- Operations leaders tasked with driving reliability strategy and execution
Testimonials (2)
Craig was extremely involved in the training, always making sure we are paying attention, adapted the examples to our day-to-day activities and always provided an answer when asked, even if the information was not added in the presentation.
Ecaterina Ioana Nicoale - BOOKING HOLDINGS ROMANIA SRL
Course - DevOps Foundation®
High level of commitment and knowledge of the trainer