What Can AIOps Do For IT Ops? - Part 4
October 28, 2021
Share this

APMdigest asked the top minds in the industry what they think AIOps can do for IT Operations. Part 4 covers root cause analysis and automation.

Start with What Can AIOps Do For IT Ops? - Part 1

Start with What Can AIOps Do For IT Ops? - Part 2

Start with What Can AIOps Do For IT Ops? - Part 3

SINGLE PANE OF GLASS

AIOps provides a much needed real-time "single-pane-of-glass" view into complex IT infrastructures that encompass fragmented and distributed multi-vendor, multi-domain technologies including legacy, virtualization, hybrid cloud, containers, microservices, and others. Although AIOps is a seismic change for IT operations, it's not a radical application of analytics and machine learning. The potential of AIOps is enormous. Enterprises that have deployed AIOps solutions are experiencing transformational benefits in revenue growth, better customer retention, improved customer experience, lower costs, and enhanced performance. The time to move is now.
Maruti Sivakumar V
SVP, Head of Digital & Practices, Blue.cloud

ISOLATING THE ROOT CAUSE

AIOps helps build high-quality incidents that include all the necessary technical and business context, alongside AI/ML-identified probable root cause and root cause changes — and present it all within a single pane of glass.
Mohan Kompella, VP Product Marketing,
Adam Blau, Director of Product Marketing,
Anirban Chatterjee, Director of Product Marketing, BigPanda

AIOps is a buzzword 6 different types of products designed to create value for IT Operations professionals. Always pick specific use cases you wish to solve and then understand how machine learning and AI can apply to solve that issue or set of issues. Good examples of this are to help the user isolate the root cause down to a specific component, highlight outliers in graphs and other views, correlate likely related data types together. Generally, these technologies help augment the operator of the software versus being automation magic. Most often these are features in other Observability tools versus AIOps platforms. AIOps platforms are fantasy because the semantic meaning of data is not clear. The result is vendors write rules to analyze the data, making the resulted outcomes only work in specific situations which makes them useless when a major problem happens across a set of complex systems.
Jonah Kowall
CTO, Logz.io

AUTOMATED ROOT CAUSE ANALYSIS

Response automation is one of the most value-driving features of AIOps software tools. IT operators are able to conduct performance tests to establish a baseline for each metric or KPI and define acceptable thresholds for the ones they want to prioritize. When a KPI breach is detected, AIOps software can perform an automated root cause analysis to automatically determine why a problem occurred and implement a solution if one is available.
Abel Gonzalez
Director of Product Marketing, Sumo Logic

Machine learning and AI are not just critical — but foundational — components of a dynamic monitoring platform. Modern applications are constantly in flux, and microservices scale through ephemeral cloud and container infrastructure in response to demand. As these systems become more complex and dynamic, operational tasks consume an increasing share of engineering time. AIOps optimizes and automates IT operations so that engineers can get proactively alerted no matter the size of the workloads, and benefit from an augmented troubleshooting experience by cutting through noise to glean key insights. In some cases, AI can auto-discover the root cause of an issue, saving minutes or hours of stressful investigations. This is the core advantage of effective AIOps — less engineering time wasted on managing complex operations, and more time building new products for customers.
Renaud Boutet
VP of Product, Datadog

BETTER DECISION-MAKING

From a monitoring and observability perspective, a key benefit of AIOps has been the ability to use historical data to increase confidence in decisions that we previously thought were black-and-white. It's relatively simple to have a machine check if a service is up or down, but how do we find the trends that show that whilst the website is up, it's gradually been getting slower over the past few months? Modern tooling allows us to collect enough data and process it fast enough — often in real-time — for the machines to be able to make better-informed decisions, faster. Such decisions could only be made by lengthy human inspection previously. It's a great example of modern tooling working in the background to make sure everything is okay, so we don't have to.
Matt Saunders
Head of DevOps, Adaptavist

AIOps observability can play a critical role in terms of expected trends using the data from users, systems and processes and provide the data back to the decision-makers to make the investment call based on the pattern, trends, etc. With growing Cloud demand, it is imperative the enterprises start investing in AIOps before it is too late.
Vishnu Vasudevan
Head of Product Engineering and Management, Opsera

SYNCING WITH ITSM

Create automated, bi-directional syncing with your ITSM platform, on-call or other collaboration tools and reduce ticket/notification volumes by up to 95%
Mohan Kompella, VP Product Marketing,
Adam Blau, Director of Product Marketing,
Anirban Chatterjee, Director of Product Marketing, BigPanda

First generation AIOps solutions are a step in the right direction, to address the unending IT complexity, but needed more care and feed and only solved limited set of problems for ITOps teams. Looking ahead, new age AIOps platforms are poised to make AIOps faster, better and cheaper — by automating data preparations and integrations, by having native asset/topology intelligence and by using expanded AI/ML frameworks like neural networks, NLP, transformer models and graph databases to address a lot more use cases. This paves a path where everybody in the IT benefits — ITSM, Service Desk, IT Asset/Planning and more.
Tejo Prayaga
Product Management, CloudFabrix

UNDERSTANDING ALGORITHMS

The last several years have seen a dramatic increase in the use of AI across all types of companies and platforms. These complex solutions require more parts of an organization to be knowledgeable of AI, from data pipelines to the workflows that build, qualify and optimize the models. Having a specialized Ops function that understands this end-to-end is going to be critical for maximizing AI's effectiveness in a production environment. Over time, AIOps can build a deeper understanding of the algorithms, then use that knowledge to enhance the infrastructure with automated services around data cleaning, model tuning and scaling that will continue delivering key results for the business. This kind of specialty is beyond what a traditional IT Operations team can do with the breadth that they are normally expected to maintain.
David Luks
VP of Engineering, Smart Applications, Lucidworks

AUTOMATION

AIOps delivers significant value to businesses by automating many of the manual, tedious tasks that distract IT from working on higher level projects, especially when it comes to data prep.
David P. Mariani
CTO and Founder, AtScale

As the cadence of business continues to gain momentum and competition builds, organizations must not only innovate but also identify business problems and inefficiencies and utilize technology to overcome them. AIOps acts as the salve for many enterprise challenges by anchoring a triangulation of machine learning, decision automation and advanced analytics to automate repetitive tasks, freeing IT teams to work on new mission critical and challenging problems — resulting in faster completion of projects and improved business outcomes.
Alan Young
CPO, InRule

REMEDIAL OPTIMIZATION

IT Operations cannot keep up with the requirements of keeping cloud applications functional and running their best. IT Ops needs to utilize the power of AI to keep the many combinations of app parameters and metrics in an optimal state. Moreso, for AIOps to keep operational apps optimized it needs to be continuous (always on) and autonomous (no human intervention). This way AIOps can perform the remedial optimization work the IT Ops SREs would do, but much faster and with more accuracy.
Peter Nickolov
Co-Founder and VP of Engineering, Opsani

Go to What Can AIOps Do For IT Ops? - Part 5

Share this

The Latest

April 19, 2024

In MEAN TIME TO INSIGHT Episode 5, Shamus McGillicuddy, VP of Research, Network Infrastructure and Operations, at EMA discusses the network source of truth ...

April 18, 2024

A vast majority (89%) of organizations have rapidly expanded their technology in the past few years and three quarters (76%) say it's brought with it increased "chaos" that they have to manage, according to Situation Report 2024: Managing Technology Chaos from Software AG ...

April 17, 2024

In 2024 the number one challenge facing IT teams is a lack of skilled workers, and many are turning to automation as an answer, according to IT Trends: 2024 Industry Report ...

April 16, 2024

Organizations are continuing to embrace multicloud environments and cloud-native architectures to enable rapid transformation and deliver secure innovation. However, despite the speed, scale, and agility enabled by these modern cloud ecosystems, organizations are struggling to manage the explosion of data they create, according to The state of observability 2024: Overcoming complexity through AI-driven analytics and automation strategies, a report from Dynatrace ...

April 15, 2024

Organizations recognize the value of observability, but only 10% of them are actually practicing full observability of their applications and infrastructure. This is among the key findings from the recently completed Logz.io 2024 Observability Pulse Survey and Report ...

April 11, 2024

Businesses must adopt a comprehensive Internet Performance Monitoring (IPM) strategy, says Enterprise Management Associates (EMA), a leading IT analyst research firm. This strategy is crucial to bridge the significant observability gap within today's complex IT infrastructures. The recommendation is particularly timely, given that 99% of enterprises are expanding their use of the Internet as a primary connectivity conduit while facing challenges due to the inefficiency of multiple, disjointed monitoring tools, according to Modern Enterprises Must Boost Observability with Internet Performance Monitoring, a new report from EMA and Catchpoint ...

April 10, 2024

Choosing the right approach is critical with cloud monitoring in hybrid environments. Otherwise, you may drive up costs with features you don’t need and risk diminishing the visibility of your on-premises IT ...

April 09, 2024

Consumers ranked the marketing strategies and missteps that most significantly impact brand trust, which 73% say is their biggest motivator to share first-party data, according to The Rules of the Marketing Game, a 2023 report from Pantheon ...

April 08, 2024

Digital experience monitoring is the practice of monitoring and analyzing the complete digital user journey of your applications, websites, APIs, and other digital services. It involves tracking the performance of your web application from the perspective of the end user, providing detailed insights on user experience, app performance, and customer satisfaction ...

April 04, 2024
Modern organizations race to launch their high-quality cloud applications as soon as possible. On the other hand, time to market also plays an essential role in determining the application's success. However, without effective testing, it's hard to be confident in the final product ...