Certified DevSRE™ Practitioner Training
Certified DevSRE™ Practitioner: Advanced training in DevOps and Site Reliability Engineering (SRE).

Course Fee
S$2000
Up to 70% funding available
Course Information
What you'll learn
- SRE Implementation Strategy
- DevSRE Paradigm
- Defining SRE
- SLIs, SLOs, Error Budget
- Observability & Resiliency
- Golden Signals
- Incident Troubleshooting
- Toil Reduction
- SRE Skill Set & Responsibilities
- Automation vs. Engineering
Requirements
- Minimum 1-3 years of experience with Agile software development and DevOps
- Familiarity with IT operations work
- Functional knowledge of infrastructure assets
Description
DevSRE™ Certified Practitioner
DevSRE™ Certified Practitioner
Exploring Site Reliability Engineering (SRE) - A Comprehensive Training Journey
In this training session, we will explore the world of Site Reliability Engineering (SRE) and its significance in modern software development and IT operations. We will begin with an overview of the use case, understanding its context and relevance. Next, we'll delve into the implementation strategy, identifying the necessary steps to ensure a successful SRE integration. The importance of SRE in the broader IT landscape will be covered in the "Why SRE-SRE Big Picture" section, where we'll discuss the benefits it brings to system reliability and overall business operations. To provide a practical learning experience, we will create a simulation and lab environment in the "Build The Environment" segment, where participants can apply SRE principles and practices. This hands-on approach will enable everyone to gain valuable insights and skills in SRE methodologies within a span of 3 hours.
Mastering Service Reliability: SLIs, SLOs, Monitoring, and Dashboards
In this section of the training, we will focus on understanding the critical concepts of SLIs, SLOs, and SLAs, which form the foundation of Site Reliability Engineering (SRE). Participants will gain insights into how to define, measure, and set objectives for service reliability. Following that, we will dive into the practical aspect of implementing monitoring solutions to collect essential data and metrics from the systems. This hands-on session will last for approximately 2 hours, enabling attendees to learn various monitoring tools and techniques. Additionally, we will spend an hour on building informative and user-friendly dashboards that provide real-time visibility into system performance and health. After the lunch break and two tea breaks, totaling one hour, participants will have ample time to recharge before continuing with the training.
Streamlining SRE: Instrumentation, Automation, and Incident Optimization
In this module, we will delve into the practical aspects of instrumenting the CR (Critical Resource) application, a key component of Site Reliability Engineering (SRE) implementation. This session, spanning 1.5 hours, will equip participants with the skills to effectively monitor and measure the performance and reliability of the CR App. Following that, we will focus on automating toil – the repetitive and manual tasks that can be streamlined through automation. This section will also take 1.5 hours and empower attendees to identify and implement automation opportunities to enhance efficiency. Lastly, we will spend 2 hours optimizing incident management, covering incident response protocols, post-mortems, and fostering a blameless culture for continuous improvement and learning.
SRE Implementation in Action: Auxon Case Study and Interactive Exercises
In the "Auxon Case Study" section, we will explore a real-world application of Site Reliability Engineering (SRE) principles to gain practical insights into its implementation. Participants will have one hour to delve into this case study and understand how SRE can be effectively applied in various scenarios. Following that, we will engage in "Use Cases & Tabletop Exercise," which will involve interactive problem-solving exercises based on SRE use cases. This hands-on approach, taking an additional hour, will enable participants to apply their knowledge and skills in real-time situations. A "Final Assessment" will then be conducted for two hours, assessing the participants' understanding of SRE concepts and their ability to tackle challenges effectively. We will also allocate an hour for lunch break and two tea breaks to ensure attendees have sufficient time to relax and recharge during this knowledge-packed training session.
Who this course is for
- Head of Operations
- Head of Operations Risk & Control
- Innovation Manager
- Customer Experience Manager
- Business Process Improvement Executive
- Innovation Manager
- Head of Customer Experience
- DevOps Engineers
- Support Engineers
What You'll Learn
Facilities & Equipment
Virtual Training
- Electronic materials
- IT support for software & hardware
- Administrative support
Face-to-Face Training
- Air-conditioned classroom
- Meals & refreshments provided
- Projector & smart board
- Stationery provided
What Learners Say
Real experiences, real results
Share Your Feedback
Attended this course? Log in to let other learners know what to expect.
