The New Normal for IT Ops Deepens Need for AI - Part 2
May 06, 2020

Will Cappelli
Moogsoft

Share this

The global pandemic has radically changed how enterprise IT services are consumed, both in the short and long term. Here's how AIOps can help IT Ops teams:

Start with The New Normal for IT Ops Deepens Need for AI - Part 1

Managing the New Normal

The new normal includes not only periodic recurrences of Covid-19 outbreaks but also the periodic emergence of new global pandemics. This means putting in place at least three layers of digital business continuity practice:

■ Continuity for illness-free periods

■ Continuity for periods marked by known pandemics

■ Continuity for periods marked by new pandemics

Rules-based, historical data analysis, and predictive analysis based on history become useless in this scenario. Instead, what's needed is technology that can anticipate outages without reliance on stable historical patterns, as AIOps does.

Significant economic contraction and resulting pressure on both capital and operational expenditures will lead to chronic understaffing of IT operations and NOC functions. IT Ops can leverage AIOps to achieve heightened levels of automation and to support radically deep cuts in the number of tools required to both monitor the digital infrastructure and respond to incidents that occur.

As remote work becomes default, it will become impossible to replicate the "monitoring cockpit" experience or the "service desk cockpit" experience. IT operations team members and first responders will need to get by with standard IT management software. That requires a significant increase in the number of signals that require observation on the one hand and the number of tickets which require response on the other hand. AIOps can help to manage this by reducing signals and tickets.

Optimizing the New Normal

The move to an almost entirely virtualized infrastructure and service portfolio will allow for maximum agility and the ability to reconfigure people, processes and technologies to meet emerging business needs (which will themselves likely be novel in the new normal.) To provide continuous assurance of service levels (even as the services themselves evolve), IT Ops teams can leverage AIOps and its ability to anticipate outages and brown-outs on the basis of data as it arrives, as opposed to pre-existing static models of topology and user behaviour.

The shift from an IT budget that, beyond labor commitments, is dominated by capital expenditures and maintenance, to one that is almost entirely dominated by renewable operational expenditures, will increase business resilience in the face of the three types of continuity issues outlined above. AIOps can help in this area as well by helping to anticipate short-term fluctuations in resource requirements based on the possibility of looming outages and brown-outs.

The economic contraction will accelerate digitalization and, in fact, lead to what may be called "maximum digitalization" with the consequence that, for the most part, business process events will be IT system state changes. One will not be able to manage business processes unless one simultaneously manages IT system events. AIOps can be invaluable here by effectively discovering and managing the higher-level IT system event patterns that are, in fact, business process patterns.

Will Cappelli is Field CTO at Moogsoft
Share this

The Latest

October 05, 2022

IT operations is a metrics-driven function and teams should keep score as a core practice. Services and sub-services break, alerts of varying quality come in, incidents are created, and services get fixed. Analytics can help IT teams improve these operations ...

October 04, 2022

Big Data makes it possible to bring data from all the monitoring and reporting tools together, both for more effective analysis and a simplified single-pane view for the user. IT teams gain a holistic picture of system performance. Doing this makes sense because the system's components interact, and issues in one area affect another ...

October 03, 2022

IT engineers and executives are responsible for system reliability and availability. The volume of data can make it hard to be proactive and fix issues quickly. With over a decade of experience in the field, I know the importance of IT operations analytics and how it can help identify incidents and enable agile responses ...

September 30, 2022

For businesses with vast and distributed computing infrastructures, one of the main objectives of IT and network operations is to locate the cause of a service condition that is having an impact. The more human resources are put into the task of gathering, processing, and finally visual monitoring the massive volumes of event and log data that serve as the main source of symptomatic indications for emerging crises, the closer the service is to the company's source of revenue ...

September 29, 2022

Our digital economy is intolerant of downtime. But consumers haven't just come to expect always-on digital apps and services. They also expect continuous innovation, new functionality and lightening fast response times. Organizations have taken note, investing heavily in teams and tools that supposedly increase uptime and free resources for innovation. But leaders have not realized this "throw money at the problem" approach to monitoring is burning through resources without much improvement in availability outcomes ...

September 28, 2022

Although 83% of businesses are concerned about a recession in 2023, B2B tech marketers can look forward to growth — 51% of organizations plan to increase IT budgets in 2023 vs. a narrow 6% that plan to reduce their spend, according to the 2023 State of IT report from Spiceworks Ziff Davis ...

September 27, 2022

Users have high expectations around applications — quick loading times, look and feel visually advanced, with feature-rich content, video streaming, and multimedia capabilities — all of these devour network bandwidth. With millions of users accessing applications and mobile apps from multiple devices, most companies today generate seemingly unmanageable volumes of data and traffic on their networks ...

September 26, 2022

In Italy, it is customary to treat wine as part of the meal ... Too often, testing is treated with the same reverence as the post-meal task of loading the dishwasher, when it should be treated like an elegant wine pairing ...

September 23, 2022

In order to properly sort through all monitoring noise and identify true problems, their causes, and to prioritize them for response by the IT team, they have created and built a revolutionary new system using a meta-cognitive model ...

September 22, 2022

As we shift further into a digital-first world, where having a reliable online experience becomes more essential, Site Reliability Engineers remain in-demand among organizations of all sizes ... This diverse set of skills and values can be difficult to interview for. In this blog, we'll get you started with some example questions and processes to find your ideal SRE ...