Get in Touch

Course Outline

SRE Anti-patterns

  • Identification of detrimental practices
  • Assessing the effect of anti-patterns on system reliability
  • Recommended best practices and corrective measures

SLOs as an Indicator of Customer Satisfaction

  • Establishing Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
  • Governing error budgets to balance innovation with stability
  • Comprehending the boundaries of distributed systems

Constructing Secure and Reliable Systems

  • Designing for resilience and fault tolerance
  • Embedding security within reliability engineering frameworks
  • Strategies for scalability and data preservation

Comprehensive Observability

  • Instrumentation techniques and metrics acquisition
  • Distributed tracing and synthetic monitoring methods
  • Development driven by observability insights

Platform Engineering and AIOps

  • Adopting platform-centric engineering methodologies
  • Automation and orchestration within SRE contexts
  • Utilizing DataOps and operational intelligence

Incident Management within SRE

  • Defining roles and responsibilities during incident response
  • Utilizing frameworks such as OODA
  • Automated remediation and AI/ML-supported resolution

Chaos Engineering

  • Core principles and strategies for testing resilience
  • Planning and conducting “game day” simulations
  • Deriving insights from controlled failure scenarios

SRE as the Purest Form of DevOps

  • Weaving SRE principles into DevOps workflows
  • Fostering cultural alignment and collaborative practices
  • Propelling organizational change via SRE adoption

Post-Class Assignments

  • Case studies involving large-scale system design
  • Scenarios for advanced instrumentation and monitoring
  • Application of real-world reliability problem-solving

Revision and Examination Preparation

  • Final consolidation of the DevOps Institute SRE Practitioner curriculum
  • Review of sample inquiries and practice assessments
  • Strategies and advice for effective exam performance

Conclusion and Subsequent Actions

Requirements

  • A solid grasp of fundamental Site Reliability Engineering concepts
  • Hands-on experience with DevOps methodologies and associated toolsets
  • Working knowledge of system monitoring, incident management, and automation processes

Target Audience

  • SRE specialists pursuing the DevOps Institute SRE Practitioner credential
  • DevOps engineers looking to broaden their expertise into reliability-centric roles
  • Operations executives overseeing reliability strategies and operational execution
 35 Hours

Number of participants


Price per participant

Testimonials (2)

Upcoming Courses

Related Categories