The latest News and Information on Incident Management, On-Call, Incident Response and related technologies.
People working in IT support and incident management right now are faced with unusual difficulties supporting large remote workforces and managing unpredictable workloads. On Reddit, system admins and other IT pros are bemoaning the hiccups and hassles of working in isolation while trying to resolve issues and maintain high SLAs. You can’t go grab your indispensable SME for troubleshooting, because that person is also home and inundated with messages and alerts from many different tools.
As engineering teams shift from delivering services on monolithic architectures to microservices and even serverless environments, developers are no longer just responsible for creating and maintaining their code. Shared ownership has become the new normal (or at least trending towards) and so they are now responding to production incidents and in some cases in the on-call rotation. Of course incidents vary in terms of impact, but they do take time away from innovation and creating new capabilities.
As the Coronavirus crisis unfolds and all of us struggle to understand its implications and to adapt, many thoughts come to mind on many different levels – personal, business related, philosophical. This event is definitely a game changer, in the near future for sure – and many say in the long run as well.
Moogsoft Enterprise consolidates visibility and control of monitoring tools to help entire IT Ops and DevOps teams reduce noise, prioritize incidents, reduce escalations and ensure uptime. Working from anywhere, users can easily find and resolve the root cause of incidents before they become outages.
Today, the customer experience drives IT on all levels. In our digitally transformed world, we do everything online — transact, interact, purchase and more. This mandates constant change and zero downtime. Ironically, as enterprises adopt IT innovations, IT environments get harder to manage and impact the productivity and agility of DevOps and SRE teams — and as a result, the customer experience suffers.