Lack of Automation Hinders Speed of Response to IT Outages and Incidents
January 24, 2017

Vincent Geffray
Everbridge

Share this

It's an eye opener to see that while companies have implemented service management for the most part — more than 90 percent of companies reporting that they have an IT Service Management system (ITSM) — only 11 percent of companies stated that they have automated the process for organizing their response to IT outages and incidents, according to Everbridge's 2016 State of IT Incident Management report.

This finding is significant because 47 percent of the companies reported having a major IT incident at least 6 times a year, the average cost of downtime is $8,662 per minute, and companies take 27 minutes on average to assemble an IT response team. Automated solutions can reduce this average time to 5 minutes or less. Considering the average cost of $8,662 per minute, the savings realized could be higher than $190,000 per major IT incident.

Key findings from the research include:

Most Companies Have an ITSM or Ticketing System

Over 90 percent of companies reported using an ITSM or ticketing system.

Major IT Outages or Incidents Occur Quite Frequently

47 percent of companies experience a major IT outage or incident six times or more a year.

36 percent experience them close to monthly (11 or more times per year).

More than a quarter of respondents reported that their companies experienced more than 21 incidents last year — that's close to two per month.

Only 9 percent of respondents reported that their organization did not report a major IT outage or incident in the past year.

The most common sources of incidents are network outages (experienced by 61 percent of companies), hardware failure or capacity issues (58 percent), internal business application issues (51 percent ), and unplanned maintenance (41 percent).

Responding to IT Outages and Incidents is Complicated and Too Manual

Two thirds (66 percent) of companies have distributed IT organizations with people spread among multiple locations and multiple time zones.

39 percent have more than 25 people included in their IT response teams. 29 percent have more than 50 people who need to be coordinated to respond to an incident. 16 percent more than 100 people.

43 percent of respondents reported that at least part of their process relies on manually calling and reaching out to people to activate the incident response team. Only 11 percent reported using an IT alerting tool to automate the process. These systems can improve response by reaching people through multiple modalities; use schedules to see who is available; automatically escalate to additional people if designated primary contacts do not respond; automatically organize conference bridges; and provide an audit trail of performance.

Response Times Could be Significantly Reduced by Automation

The mean time to activate and assemble a response team was cited as 27 minutes. Automated solutions can reduce this response time to 5 minutes or less.

IT downtime is expensive and hurts productivity

The average cost of IT downtime was reported as $8,662 per minute.

63 percent of respondents stated that IT incidents or outages hurt employee productivity, 60 percent that it caused IT team disruption or distraction, and 34 percent that it decreased customer satisfaction.

13 percent reported that their organization had experienced bad press or publicity due to an IT incident or outage.

Methodology: The sample for the research was 152 IT professionals, including 86% of respondents from companies of 1000 employees or more, and 45% from companies with more than 10,000 employees.

Vincent Geffray is Senior Director, Product Marketing, at Everbridge
Share this

The Latest

March 21, 2019

Achieving audit compliance within your IT ecosystem can be an iterative process, and it doesn't have to be compressed into the five days before the audit is due. Following is a four-step process I use to guide clients through the process of preparing for and successfully completing IT audits ...

March 20, 2019

Network performance issues come in all shapes and sizes, and can require vast amounts of time and resources to solve. Here are three examples of painful network performance issues you're likely to encounter this year, and how NPMD solutions can help you overcome them ...

March 19, 2019

"Scale up" versus "scale out" doesn't just apply to hardware investments, it also has an impact on product features. "Scale up" promotes buying the feature set you think you need now, then adding "feature modules" and licenses as you discover additional feature requirements are needed. Often as networks grow in size they also grow in complexity ...

March 18, 2019

Network Packet Brokers play a critical role in gaining visibility into new complex networks. They deliver the packet data and information IT and security teams need to identify problems, recognize security issues, and ensure overall network performance. However, not all Packet Brokers are created equal when it comes to scalability. Simply "scaling up" your network infrastructure at every growth point is a more complex and more expensive endeavor over time. Let's explore three ways the "scale up" approach to infrastructure growth impedes NetOps and security professionals (and the business as a whole) ...

March 15, 2019

Loyal users are the key to your service desk's success. Happy users want to use your services and they recommend your services in the organization. It takes time and effort to exceed user expectations, but doing so means keeping the promises we make to our users and being careful not to do too much without careful consideration for what's best for the organization and users ...

March 14, 2019

What's the difference between user satisfaction and user loyalty? How can you measure whether your users are satisfied and will keep buying from you? How much effort should you make to offer your users the ultimate experience? If you're a service provider, what matters in the end is whether users will keep coming back to you and will stay loyal ...

March 13, 2019

What if I said that a 95% reduction in the amount of IT noise, 99% reduction in ticket volume and 99% L1 resolution rate are not only possible, but that some of the largest, most complex enterprises in the world see these metrics in their environments every day, thanks to Artificial Intelligence (AI) and Machine Learning (ML)? Would you dismiss that as belonging to the realm of science fiction? ...

March 12, 2019
As a consumer, when you order products online, how do you expect them to get delivered? Some key requirements are: the product must arrive on time, well-packed, and ultimately must give you an easy gateway to return it if it is not as per your expectations. All this has been made possible via a single application. But what if this application doesn't function the way you want or cracks down mid-way, or probably leaks off information about you to some potential hackers? Technical uncertainty and digital chaos are the two double-edged swords dangling over this billion-dollar ecommerce market. Can Quality Assurance and Software Testing save application developers from this endless juggle? ...
March 11, 2019

Of those surveyed, 96% of organizations have a digital transformation strategy, with 57% approaching it as an enterprise-wide priority, with a clear emphasis on speed of business, costs, risk, and customer satisfaction, according to IDC’s Aligning IT Strategies and Business Expectations for Digital Transformation Success, sponsored by EasyVista ...

March 08, 2019

One of my ongoing areas of focus is analytics, AIOps, and the intersection with AI and machine learning more broadly. Within this space, sad to say, semantic confusion surrounding just what these terms mean echoes the confusions surrounding ITSM ...