The latest News and Information on DevOps, CI/CD, Automation and related technologies.
A guide to set practical Service Level Objectives (SLOs) & Service Level Indicators (SLIs) for your Site Reliability Engineering practices.
History will look back on this period of the 21st century as a pioneering, resilient, and excitingly disruptive time. We’re deep into a dynamic era as the cloud, Artificial Intelligence (AI), IT automation, and digital transformation converge to drive challenges and dazzling opportunities. The sheer force and potential of AI—coupled with unprecedented security risks and ongoing infrastructure advances will shape enterprises for years to come.
As a former incident responder and now as a responder advocate for FireHydrant, I’ve seen the “build vs. buy” debate play out many times. In fact, I even supported the tool that former employers used for managing incidents for years before they decided to buy (more on that in a future blog post).
In this podcast, our panellists discuss the foundations that any team needs to put in place when designing their incident management process. Starting from the basics of defining what we really mean by an incident, to how to set your severity levels, roles and statuses, Chris and Pete share their tips for building solid foundations to run your incidents.