Skip to main content
Garranto Academy
Cloud & DevOps Academy⭐ Be the first to rate this course

Certified DevSRE™ Practitioner Training

Certified DevSRE™ Practitioner: Advanced training in DevOps and Site Reliability Engineering (SRE).

2
Days
8
Hours/Day
Live
Training
Certified DevSRE™ Practitioner Training

Course Fee

S$2000

Up to 70% funding available

Course Information

What you'll learn

  • SRE Implementation Strategy
  • DevSRE Paradigm
  • Defining SRE
  • SLIs, SLOs, Error Budget
  • Observability & Resiliency
  • Golden Signals
  • Incident Troubleshooting
  • Toil Reduction
  • SRE Skill Set & Responsibilities
  • Automation vs. Engineering

Requirements

  • Minimum 1-3 years of experience with Agile software development and DevOps
  • Familiarity with IT operations work
  • Functional knowledge of infrastructure assets

Description

DevSRE™ Certified Practitioner

DevSRE™ Certified Practitioner

Exploring Site Reliability Engineering (SRE) - A Comprehensive Training Journey

In this training session, we will explore the world of Site Reliability Engineering (SRE) and its significance in modern software development and IT operations. We will begin with an overview of the use case, understanding its context and relevance. Next, we'll delve into the implementation strategy, identifying the necessary steps to ensure a successful SRE integration. The importance of SRE in the broader IT landscape will be covered in the "Why SRE-SRE Big Picture" section, where we'll discuss the benefits it brings to system reliability and overall business operations. To provide a practical learning experience, we will create a simulation and lab environment in the "Build The Environment" segment, where participants can apply SRE principles and practices. This hands-on approach will enable everyone to gain valuable insights and skills in SRE methodologies within a span of 3 hours.

Mastering Service Reliability: SLIs, SLOs, Monitoring, and Dashboards

In this section of the training, we will focus on understanding the critical concepts of SLIs, SLOs, and SLAs, which form the foundation of Site Reliability Engineering (SRE). Participants will gain insights into how to define, measure, and set objectives for service reliability. Following that, we will dive into the practical aspect of implementing monitoring solutions to collect essential data and metrics from the systems. This hands-on session will last for approximately 2 hours, enabling attendees to learn various monitoring tools and techniques. Additionally, we will spend an hour on building informative and user-friendly dashboards that provide real-time visibility into system performance and health. After the lunch break and two tea breaks, totaling one hour, participants will have ample time to recharge before continuing with the training.

Streamlining SRE: Instrumentation, Automation, and Incident Optimization

In this module, we will delve into the practical aspects of instrumenting the CR (Critical Resource) application, a key component of Site Reliability Engineering (SRE) implementation. This session, spanning 1.5 hours, will equip participants with the skills to effectively monitor and measure the performance and reliability of the CR App. Following that, we will focus on automating toil – the repetitive and manual tasks that can be streamlined through automation. This section will also take 1.5 hours and empower attendees to identify and implement automation opportunities to enhance efficiency. Lastly, we will spend 2 hours optimizing incident management, covering incident response protocols, post-mortems, and fostering a blameless culture for continuous improvement and learning.

SRE Implementation in Action: Auxon Case Study and Interactive Exercises

In the "Auxon Case Study" section, we will explore a real-world application of Site Reliability Engineering (SRE) principles to gain practical insights into its implementation. Participants will have one hour to delve into this case study and understand how SRE can be effectively applied in various scenarios. Following that, we will engage in "Use Cases & Tabletop Exercise," which will involve interactive problem-solving exercises based on SRE use cases. This hands-on approach, taking an additional hour, will enable participants to apply their knowledge and skills in real-time situations. A "Final Assessment" will then be conducted for two hours, assessing the participants' understanding of SRE concepts and their ability to tackle challenges effectively. We will also allocate an hour for lunch break and two tea breaks to ensure attendees have sufficient time to relax and recharge during this knowledge-packed training session.

Who this course is for

  • Head of Operations
  • Head of Operations Risk & Control
  • Innovation Manager
  • Customer Experience Manager
  • Business Process Improvement Executive
  • Innovation Manager
  • Head of Customer Experience
  • DevOps Engineers
  • Support Engineers

What You'll Learn

Practical hands-on experience
Industry-recognized certification
Real-world case studies
Expert-led live sessions
Comprehensive study materials
Post-training support

Facilities & Equipment

Virtual Training

  • Electronic materials
  • IT support for software & hardware
  • Administrative support

Face-to-Face Training

  • Air-conditioned classroom
  • Meals & refreshments provided
  • Projector & smart board
  • Stationery provided

What Learners Say

Real experiences, real results

Share Your Feedback

Attended this course? Log in to let other learners know what to expect.

By enrolling in this course, you agree to our terms and conditions.

Book Career Advisory