The New Normal for IT Ops Deepens Need for AI - Part 1
May 05, 2020

Will Cappelli
Moogsoft

Share this

The global pandemic has radically changed how enterprise IT services are consumed, both in the short and long term. Here's how AIOps can help IT Ops teams.

The current crisis has upended all aspects of our personal and work lives, and IT Ops pros aren't the exception. The abrupt shift to remote work has created unprecedented challenges for IT Ops teams, while increasing pressure on them to prevent outages and provide service assurance.

Specifically, new consumption patterns of enterprise IT services have put stress on systems, architectures and topologies at all stack layers. In response, IT Ops teams must rapidly implement structural and management changes to address both temporary and permanent shifts.

In this turmoil, AIOps has emerged as a lifeline. By streamlining and automating IT operations, AIOps helps IT leaders collaborate remotely and act quickly and precisely to maintain business-critical digital services — during the pandemic and beyond.

Let's look in more detail at these challenges and at how AIOps can help IT Ops teams cope and succeed.

AIOps: A Definition

An AIOps solution must have these five types of algorithms that fully automate and streamline five key dimensions of IT operations monitoring:

■ Data selection: Identifying and surfacing the most relevant information.

■ Pattern discovery: Correlating and finding relationships between events across your tool stack.

■ Inference: Identifying root causes and recurring issues.

■ Collaboration: Notifying appropriate operators, and facilitating collaboration.

■ Automation: Automating remediation

In a real world setting, an AIOps solution ingests heterogeneous data from many different sources. Using entropy algorithms, it removes noise and duplication, and selects only the truly relevant data. It then groups and correlates this relevant information using various criteria, like text, time and topology.

Next, it discovers patterns in the data, and infers which data items signify causes, and which signify events. It then communicates the result of that analysis to a collaborative environment, which will support automated responses to what has been discovered.

As such, an AIOps solution plays the role of organizing and integrating what an organization's domain-specific IT monitoring and management tools do, intelligently integrating the stack's functionalities. AIOps should act as the brain that brings together these tools, and becomes a coordinating, central layer.

Transitioning to the New Normal

As the workforce shifts to remote work, user behaviors will change and different elements of the IT infrastructure, both in-house and publicly sourced, will be stressed. This will result in new, quickly-evolving types of incidents and outages. With AIOps, IT Ops teams can detect and analyze genuinely novel anomalies which can cause incidents and outages rapidly and stealthily.

Cross-regional and intra-regional team collaboration among IT operations and NOC organizations will need to be reinforced virtually as the implicit supports derived from physical co-presence are removed. AIOps can enable and guide virtual collaborative observation, analysis and response efforts, helping IT Ops teams collaborate and communicate despite being physically dispersed.

Sharp and unpredictable levels of staff reduction due to illness and self-isolation will force IT operations and NOC organizations to "do more with less" on both the side of signal observation and the side of signal response. Here again AIOps can help IT Ops teams to respond by both dynamically filtering noisy alert streams, and integrating and automating platforms that support various aspects of incident and problem management.

Go to The New Normal for IT Ops Deepens Need for AI - Part 2

Will Cappelli is CTO, EMEA, at Moogsoft
Share this

The Latest

May 21, 2020

As cloud computing continues to grow, tech pros say they are increasingly prioritizing areas like hybrid infrastructure management, application performance management (APM), and security management to optimize delivery for the organizations they serve, according to ...

May 20, 2020

Businesses see digital experience as a growing priority and a key to their success, with execution requiring a more integrated approach across development, IT and business users, according to Digital Experiences: Where the Industry Stands ...

May 19, 2020

Fully 90% of those who use observability tooling say those tools are important to their team's software development success, including 39% who say observability tools are very important ...

May 18, 2020

As our production application systems continuously increase in complexity, the challenges of understanding, debugging, and improving them keep growing by orders of magnitude. The practice of Observability addresses both the social and the technological challenges of wrangling complexity and working toward achieving production excellence. New research shows how observable systems and practices are changing the APM landscape ...

May 14, 2020
Digital technologies have enveloped our lives like never before. Be it on the personal or professional front, we have become dependent on the accurate functioning of digital devices and the software running them. The performance of the software is critical in running the components and levers of the new digital ecosystem. And to ensure our digital ecosystem delivers the required outcomes, a robust performance testing strategy should be instituted ...
May 13, 2020

The enforced change to working from home (WFH) has had a massive impact on businesses, not just in the way they manage their employees and IT systems. As the COVID-19 pandemic progresses, enterprise IT teams are looking to answer key questions such as: Which applications have become more critical for working from home? ...

May 12, 2020

In ancient times — February 2020 — EMA research found that more than 50% of IT leaders surveyed were considering new ITSM platforms in the near future. The future arrived with a bang as IT organizations turbo-pivoted to deliver and support unprecedented levels and types of services to a global workplace suddenly working from home ...

May 11, 2020

The Internet of Things (IoT) is changing the world. From augmented reality advanced analytics to new consumer solutions, IoT and the cloud are together redefining both how we work and how we engage with our audiences. They are changing how we live, as well ...

May 07, 2020

Despite IT professionals' confidence in their ability to support today's much greater dependence on digital services, there is a rise in application performance errors reported by more than half of consumers, according to the Impact of COVID-19 on Digital Transformation survey from xMatters ...

May 06, 2020

The new normal includes not only periodic recurrences of Covid-19 outbreaks but also the periodic emergence of new global pandemics. This means putting in place at least three layers of digital business continuity practice ...