Get in Touch
 Duration 35 hours

Course Outline

SRE Anti-patterns

  • Spotting counterproductive practices
  • Assessing the effect of anti-patterns on system reliability
  • Adopting best practices and corrective alternatives

SLOs as a Proxy for Customer Satisfaction

  • Establishing Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
  • Oversight of error budgets and the balance between innovation and reliability
  • Comprehending the boundaries of distributed systems

Constructing Secure and Reliable Systems

  • Architecture for fault tolerance and resilience
  • Embedding security within reliability engineering
  • Strategies for scalability and data protection

Full-stack Observability

  • Instrumentation and metrics aggregation
  • Distributed tracing and synthetic monitoring
  • Observability-centric development approaches

Platform Engineering and AIOps

  • Platform-focused engineering methodologies
  • Automation and orchestration in the context of SRE
  • Utilizing DataOps and operational intelligence

Incident Management in SRE

  • Defining roles and responsibilities in incident response
  • Application of frameworks such as OODA
  • Automated remediation and AI/ML-supported resolution

Chaos Engineering

  • Core principles and tactics for resilience testing
  • Planning and carrying out “game day” simulations
  • Deriving insights from controlled failure experiments

SRE as a Pure Form of DevOps

  • Merging SRE into DevOps workflows
  • Cultural alignment and collaborative practices
  • Facilitating organizational transformation via SRE

Post-class Exercises

  • Case studies on large-scale system design
  • Complex instrumentation and monitoring scenarios
  • Resolution of real-world reliability challenges

Review and Exam Preparation

  • Comprehensive review of the DevOps Institute SRE Practitioner syllabus
  • Sample questions and mock exams
  • Tactical advice and recommendations for taking the exam

Summary and Next Steps

Requirements

  • Foundational understanding of core Site Reliability Engineering principles
  • Practical experience with DevOps workflows and associated tooling
  • Proficiency in system monitoring, incident management, and automation

Target Audience

  • SRE professionals pursuing the DevOps Institute SRE Practitioner certification
  • DevOps engineers looking to broaden their skill set into reliability-focused roles
  • Operations leaders tasked with overseeing reliability strategy and execution

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories