Monitoring as a Differentiator: Breaking Silos and Building Understanding
March 27, 2017

David Drai
Anodot

Share this

Monitoring a business means monitoring an entire business – not just IT or application performance. If businesses truly care about differentiating themselves from the competition, they must approach monitoring holistically. Separate, siloed monitoring systems are quickly becoming a thing of the past.

I see time and again cloud monitoring companies working with a myopic focus on the Infrastructure area – a critical mistake. They concentrate on system health but avoid business health like the plague. Although CPU, Disk, Memory and other infrastructure KPIs are essential to maintain a healthy system, their coverage is limited and lacking an equally crucial component that drives how well a company is operating – its business. Today there is simply no excuse for having incomplete monitoring capabilities, and it is more necessary than ever to get out of monitoring siloes.

Cloud Monitoring 1.0 and the Evolution of Metrics

Monitoring infrastructure provides some visibility to overall system health by keeping machines up and running – but it is not at all adequate to determine what is occurring on the business side of a company. Infrastructure monitoring is also far too basic to keep up with updates within applications – essentially putting blinders on a company's leadership.

As it stands, infrastructure monitoring tools usually run in conjunction with other internal tools to gain an angle on the business, or analysts rely on Business Intelligence solutions that may be connected to infrastructure monitoring through internal scripts. In most cases, these 1.0 level tools require a great deal of internal development and maintenance which are difficult to scale.

In the past few years, time series metrics have been the main driver of growth in cloud monitoring systems. This approach of normalizing almost all data per a single time series representation has enabled the provision of generic solutions for many cases and different customers. Because of its rudimentary ability, it is not surprising that open source solutions are becoming so widespread among the businesses which are beginning to understand the importance of monitoring. The ability to represent all metrics in the same manner using the same dashboards and time series function sets has significantly simplified this monitoring method providing good but not fully comprehensive information.

Today's Challenges of Monitoring Business

One of the main challenges of monitoring business KPIs is that static rules and alerts are too limiting. Particularly for metrics that change per trends or seasons, static alerts are difficult to maintain because of their inherent variability. Even in the simplest cases, it is very difficult to define thresholds for thousands of metrics because it requires the user to have working knowledge of their normal range. For e-commerce companies, the holiday season is always a peak time in sales and every metric is going to behave "abnormally." It is nearly impossible for large data-driven companies, which are monitoring so much, to start making changes to reset the threshold for every single metric – talk about a nightmare.

Another challenge of monitoring so many metrics is defining rules manually especially when each metric has a different normal range. Unfortunately, it is essential that this be done to achieve effective configuration. Amazon needs to know that "Elf on a Shelf" dolls are going to sell heavily in November and that gift certificates will be sold later in the month.

Cloud Monitoring 2.0: for IT, applications AND BUSINESS

The newest generation of monitoring centralizes all company activity into a single unified solution, rather than separate solutions for IT, application, and business. This is the holistic understanding that companies have been working towards for so long – the ability to understand every metric separately and together. It is one thing to see an infrastructure anomaly on its own, but to be able to contextualize it with the correlated impact on the business affords an entirely new way to problem-solve and measure the health of a company. Beyond addressing the immediate issues this type of top-down monitoring approach offers tremendous value.

Without a smart mechanism to monitor so many rules and alerts, companies are bound to compromise what they monitor, sacrificing all for a few selected metrics. Analysts are not fortune tellers – there is no way to define what the best metrics are to monitor. This creates an inevitable delay in detection of issues, which severely limits how proactive a company can be in the varied business scenarios it faces. It also limits the granularity of the organization's visibility – bringing us back to where we were with Cloud Monitoring 1.0.

Only recently the implementation of AI in BI is enabling companies to solve challenges in monitoring. By automating the ability to differentiate between what is normal and abnormal behavior (no matter the trend or time of year) businesses finally have a chance to review a comprehensive and automatic evaluation of anomalies. With the addition of AI to monitoring, companies can differentiate themselves by how quickly they respond to changing conditions; how quickly they find bugs and glitches, how rapidly they respond to customers in crisis, and how swiftly they leverage a business opportunity triggered by a celebrity's viral Instagram post.

While companies engage with their customers in more ways than ever before, finding ways to break out of monitoring silos is going to be the key that companies use to successfully scale and compete with industry giants.

David Drai is CEO and Co-Founder of Anodot.

Share this

The Latest

November 16, 2018

Everyone talks about automating the software development lifecycle (SDLC) but the first question should be: What should you automate? With this question in mind, DEVOPSdigest asked experts from across the IT industry for their opinions on what steps in the SDLC should be automated ...

November 15, 2018

We all know artificial intelligence (AI) is a hot topic — but beyond the buzzword, have you ever wondered how IT departments are actually adopting AI technologies to improve on their operations? ...

November 14, 2018

How can IT teams focus on the critical events that can impact their business instead of wading through false positives? The emerging discipline of AIOps is a much-needed panacea for detecting patterns, identifying anomalies, and making sense of alerts across hybrid infrastructure ...

November 09, 2018

In a recent webinar AIOps and IT Analytics at the Crossroads, I was asked several times about the borderline between AIOps and monitoring tools — most particularly application performance monitoring (APM) capabilities. The general direction of the questions was — how are they different? Do you need AIOps if you have APM already? Why should I invest in both? ...

November 08, 2018

There's no place like the web and smartphones for the holidays. With the biggest shopping season of the year quickly approaching, retailers are gearing up to experience the most traffic their online platforms (web, mobile, IoT) have ever seen. To avoid missing out on millions this holiday season, below are the top five ways developers can keep their apps and websites up and running without a hitch ...

November 07, 2018

Usage data is multifaceted, with many diverse benefits. Harvesting usage-driven insights effectively requires both good foundational technology and a nimbleness of mind to unify insights across IT's many silos of domains and disciplines. Because of this, leveraging usage-driven insights can in itself become a catalyst for helping IT as a whole transform toward improved efficiencies and enhanced levels of business alignment ...

November 06, 2018

The requirements to maintain the complete availability and superior performance of your mission-critical workloads is a dynamic process that has never been more challenging. Here are five ways IT teams can measure and guarantee performance-based SLAs in order to increase the value of the infrastructure to the business, and ensure optimal digital performance levels ...

November 05, 2018

APMdigest asked experts from across the IT industry for their opinions on what IT departments should be monitoring to ensure digital performance. Part 5, the final installment, offers some recommendations you may not have thought about ...

November 02, 2018

APMdigest asked experts from across the IT industry for their opinions on what IT departments should be monitoring to ensure digital performance. Part 4 covers the infrastructure, including the cloud and the network ...

November 01, 2018

APMdigest asked experts from across the IT industry for their opinions on what IT departments should be monitoring to ensure digital performance. Part 3 covers the development side ...