Get in Touch
 Duration 35 hours

Course Outline

SRE Anti-patterns

  • Identifying counterproductive practices.
  • Assessing the impact of anti-patterns on system reliability.
  • Exploring best practices and effective corrective alternatives.

SLOs as a Proxy for Customer Satisfaction

  • Defining Service Level Indicators (SLIs) and Service Level Objectives (SLOs).
  • Managing error budgets while balancing innovation with reliability.
  • Understanding the inherent limits of distributed systems.

Building Secure and Reliable Systems

  • Designing for fault tolerance and resilience.
  • Integrating security measures into reliability engineering.
  • Implementing scalability and data protection strategies.

Full-stack Observability

  • Instrumentation and metrics collection techniques.
  • Distributed tracing and synthetic monitoring.
  • Adopting observability-driven development methodologies.

Platform Engineering and AIOps

  • Platform-centered engineering approaches.
  • Automation and orchestration within SRE frameworks.
  • Leveraging DataOps and operational intelligence.

Incident Management in SRE

  • Defining roles and responsibilities in incident response.
  • Applying established frameworks such as OODA.
  • Utilizing automated remediation and AI/ML-assisted resolution.

Chaos Engineering

  • Principles and strategies for resilience testing.
  • Planning and executing “game day” exercises.
  • Deriving insights from controlled failure experiments.

SRE as a Pure Form of DevOps

  • Integrating SRE into existing DevOps workflows.
  • Aligning culture and collaboration practices.
  • Driving organizational transformation through SRE.

Post-class Exercises

  • Case studies involving large-scale system design.
  • Advanced instrumentation and monitoring scenarios.
  • Real-world reliability problem-solving.

Review and Exam Preparation

  • Comprehensive review of the DevOps Institute SRE Practitioner syllabus.
  • Sample questions and practice tests.
  • Strategies and recommendations for exam success.

Summary and Next Steps

Requirements

  • A solid grasp of fundamental Site Reliability Engineering principles.
  • Practical experience with DevOps practices and associated toolsets.
  • Proficiency in system monitoring, incident management, and automation.

Target Audience

  • SRE professionals aiming to obtain the DevOps Institute SRE Practitioner certification.
  • DevOps engineers looking to transition into reliability-focused roles.
  • Operations leaders tasked with overseeing reliability strategy and execution.

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories