Lack of Automation Hinders Speed of Response to IT Outages and Incidents
January 24, 2017

Vincent Geffray
Everbridge

Share this

It's an eye opener to see that while companies have implemented service management for the most part — more than 90 percent of companies reporting that they have an IT Service Management system (ITSM) — only 11 percent of companies stated that they have automated the process for organizing their response to IT outages and incidents, according to Everbridge's 2016 State of IT Incident Management report.

This finding is significant because 47 percent of the companies reported having a major IT incident at least 6 times a year, the average cost of downtime is $8,662 per minute, and companies take 27 minutes on average to assemble an IT response team. Automated solutions can reduce this average time to 5 minutes or less. Considering the average cost of $8,662 per minute, the savings realized could be higher than $190,000 per major IT incident.

Key findings from the research include:

Most Companies Have an ITSM or Ticketing System

Over 90 percent of companies reported using an ITSM or ticketing system.

Major IT Outages or Incidents Occur Quite Frequently

47 percent of companies experience a major IT outage or incident six times or more a year.

36 percent experience them close to monthly (11 or more times per year).

More than a quarter of respondents reported that their companies experienced more than 21 incidents last year — that's close to two per month.

Only 9 percent of respondents reported that their organization did not report a major IT outage or incident in the past year.

The most common sources of incidents are network outages (experienced by 61 percent of companies), hardware failure or capacity issues (58 percent), internal business application issues (51 percent ), and unplanned maintenance (41 percent).

Responding to IT Outages and Incidents is Complicated and Too Manual

Two thirds (66 percent) of companies have distributed IT organizations with people spread among multiple locations and multiple time zones.

39 percent have more than 25 people included in their IT response teams. 29 percent have more than 50 people who need to be coordinated to respond to an incident. 16 percent more than 100 people.

43 percent of respondents reported that at least part of their process relies on manually calling and reaching out to people to activate the incident response team. Only 11 percent reported using an IT alerting tool to automate the process. These systems can improve response by reaching people through multiple modalities; use schedules to see who is available; automatically escalate to additional people if designated primary contacts do not respond; automatically organize conference bridges; and provide an audit trail of performance.

Response Times Could be Significantly Reduced by Automation

The mean time to activate and assemble a response team was cited as 27 minutes. Automated solutions can reduce this response time to 5 minutes or less.

IT downtime is expensive and hurts productivity

The average cost of IT downtime was reported as $8,662 per minute.

63 percent of respondents stated that IT incidents or outages hurt employee productivity, 60 percent that it caused IT team disruption or distraction, and 34 percent that it decreased customer satisfaction.

13 percent reported that their organization had experienced bad press or publicity due to an IT incident or outage.

Methodology: The sample for the research was 152 IT professionals, including 86% of respondents from companies of 1000 employees or more, and 45% from companies with more than 10,000 employees.

Vincent Geffray is Senior Director, Product Marketing, at Everbridge
Share this

The Latest

November 24, 2020

Shoppers are heading into Black Friday with high expectations for digital experiences and are only willing to experience a service interruption of five minutes or less to get the best deal, according to the 2020 Black Friday and Cyber Monday eCommerce Trends Study, from xMatters ...

November 23, 2020

Digital Experience Monitoring (DEM) has become significant to businesses more than ever. Global events like Covid continue to disrupt best practices within IT to support business. The pandemic has already forced millions of employees to WFH and adopt a hybrid workspace. Network connectivity and cloud application issues in these environments will continue to impact productivity and slow progress. Even so, transparent migration and deployment of on-premise workloads across multi-cloud providers, by their very nature are complex ...

November 20, 2020

APMdigest posed the following question to the IT Operations community: How should ITOps adapt to the new normal? In response, industry experts offered their best recommendations for how ITOps can adapt to this new remote work environment. Part 5, the final installment in the series, covers open source and emerging technologies ...

November 19, 2020

APMdigest posed the following question to the IT Operations community: How should ITOps adapt to the new normal? In response, industry experts offered their best recommendations for how ITOps can adapt to this new remote work environment. Part 4 covers monitoring and visibility ...

November 18, 2020

APMdigest posed the following question to the IT Operations community: How should ITOps adapt to the new normal? In response, industry experts offered their best recommendations for how ITOps can adapt to this new remote work environment. Part 3 covers automation ...

November 17, 2020

APMdigest posed the following question to the IT Operations community: How should ITOps adapt to the new normal? In response, industry experts offered their best recommendations for how ITOps can adapt to this new remote work environment. Part 2 covers communication and collaboration ...

November 16, 2020

The "New Normal" in IT — the fact that most IT Operations personnel work from home (WFH) today — is here to stay. What started out as a reaction to the COVID-19 pandemic is now a way of life. Many experts agree that IT teams will not be going back to the office any time soon, even if the public health concerns are abated. How should ITOPs adapt to the new normal? That is the question APMdigest posed to the IT industry. ITOps experts — from analysts and consultants to the top vendors — offer their best recommendations for how ITOps can react to this new environment ...

November 12, 2020

The pandemic effectively "shocked" enterprises into pushing the gas on tech initiatives that, on the one hand, support a more flexible, decentralized workforce, but that were by-and-large already on the roadmap, regardless of whether businesses had been planning to support widespread work-from-home or not ...

November 10, 2020

Maintaining call quality with Microsoft Teams is a process, not a onetime event. Network engineers and Microsoft Teams application owners need to be vigilant in preserving optimal call quality to ensure audio, video, and screen-sharing always remain satisfactory for end-users. In this blog, we cover how the Microsoft Teams Call Quality Dashboard (CQD) combined with the audio/video synthetic transaction monitoring improves this maintenance process ...

November 09, 2020

For IT teams, catching errors in applications before they become detrimental to a project is critical. Wouldn't it be nice if there was someone standing over your shoulder, letting you know exactly when, where, and what the issue is so you can correct it immediately? Luckily, there are both application performance management (APM) and application stability management (ASM) solutions available that can do this for you, flagging errors in both the deployment and development stages of applications, before they can create larger issues down the line ...