Thank you for sending your enquiry! One of our team members will contact you shortly.
Thank you for sending your booking! One of our team members will contact you shortly.
Duration 35 hours
Course Outline
SRE Anti-patterns
- Spotting counterproductive practices
- Assessing the impact of anti-patterns on system reliability
- Applying best practices and corrective alternatives
SLOs as a Metric for Customer Satisfaction
- Establishing Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
- Overseeing error budgets to balance innovation with reliability
- Comprehending the limitations of distributed systems
Creating Secure and Reliable Systems
- Architecting for fault tolerance and resilience
- Weaving security into reliability engineering workflows
- Adopting scalability and data protection strategies
Comprehensive Observability
- Implementing instrumentation and metrics collection
- Utilizing distributed tracing and synthetic monitoring
- Promoting observability-driven development
Platform Engineering and AIOps
- Adopting platform-centered engineering methodologies
- Enhancing automation and orchestration in SRE contexts
- Harnessing DataOps and operational intelligence
Incident Management in SRE
- Defining roles and responsibilities in incident response
- Utilizing frameworks such as OODA
- Implementing automated remediation and AI/ML-assisted resolution
Chaos Engineering
- Applying principles and strategies for resilience testing
- Planning and conducting “game day” exercises
- Deriving insights from controlled failure experiments
SRE as a Pure Form of DevOps
- Weaving SRE into DevOps workflows
- Fostering cultural alignment and collaborative practices
- Propelling organizational transformation through SRE
Post-class Exercises
- Analyzing large-scale system design case studies
- Addressing advanced instrumentation and monitoring scenarios
- Solving real-world reliability challenges
Review and Exam Preparation
- Conducting a final review of the DevOps Institute SRE Practitioner syllabus
- Working through sample questions and practice tests
- Refining exam-taking strategies and recommendations
Summary and Next Steps
Requirements
- A solid grasp of fundamental Site Reliability Engineering principles
- Hands-on experience with DevOps methodologies and associated tools
- Working knowledge of system monitoring, incident management, and automation
Target Audience
- SRE professionals pursuing DevOps Institute SRE Practitioner certification
- DevOps engineers looking to broaden their scope into reliability-centric roles
- Operations leaders tasked with driving reliability strategy and execution
Testimonials (2)
The knowledge and experience of the consultant, as theoretical topics are addressed by applying them to the reality of processes. The course contains a highly valuable program in information technology management.
Luis Castro Gamboa - Cooperativa De Ahorro Y Credito Ande No. 1 R.L.
Course - Site Reliability Engineering (SRE) Foundation®
Machine Translated
That it was very clear in each specification
Ricardo Ramirez - AMX CONTENIDO
Course - DevOps Leader (DOL)®
Machine Translated