Monte Carlo Launches Incident IQ
July 14, 2021
Share this

Monte Carlo released Incident IQ, a new suite of capabilities that help data engineers better pinpoint, address, and resolve data downtime at scale through the Monte Carlo Data Observability Platform.

Incident IQ automatically generates rich insights about critical data issues through root cause analysis, giving teams unprecedented visibility into the end-to-end health and trust of their data beyond the scope of traditional data quality solutions.

On average, companies lose over $15 million per year on bad data, with data engineers spending upwards of 40 percent - or 120 hours per week - of their time tackling broken data pipelines. In the same way that New Relic, DataDog, and other Application Performance Management (APM) solutions ensure reliable software and keep application downtime at bay, Data Observability solves the costly problem of data downtime, in other words, periods of time when data is missing, inaccurate, or otherwise unreliable.

To help companies eliminate data downtime, Monte Carlo built Incident IQ, the first end-to-end solution that conducts root cause analysis for data issues at each stage of the pipeline, from ingestion in the data warehouse or lake to analytics in your business intelligence dashboards. Incident IQ automatically generates historical insights about your data to identify patterns in query logs, trigger investigative follow-on query results, and monitor upstream dependency changes to pin-point exactly what caused the issue to occur, reducing the amount of data incidents by 90 percent at each stage of the pipeline.

Developed after reviewing thousands of real data incidents from our customers, Incident IQ gives data engineers access to insights about their code, their data, and their operational environment that allows them to quickly and collaboratively get to the root cause of data problems -- all in a single UI.

With Incident IQ, everything related to the data issue is captured in an elegant timeline with easy commenting, documentation, and collaboration features to create rich post-mortems. This level of detail, common in software engineering and DevOps tooling, helps data teams learn from past incidents and determine where to allocate future investment. Additionally, Incident IQ makes it easy to create and share high-level incident reporting with CTOs and CDOs, fostering greater data trust and ownership across the company.

Core capabilities of Incident IQ include:

- Central UI that connects the dots between correlated causes of data incidents, and surfaces a historical collection of data incidents for quick comparison.

- Access to example queries that pull sample data, as well as rich query logs, historical incidents, and quick links to Monte Carlo’s Lineage and Catalog features, making it easy to identify, root cause, and fix data issues all from the same interface.

- Automatic insights based on the statistical correlation between table fields in anomalous records (for instance, Incident IQ can surface if an increase in order_id null values correlates with a specific order source).

- Automatic, end-to-end lineage that maps impacted downstream BI dashboards to the furthest upstream tables, helping teams narrow the focus of root cause investigations.

- Automatic runbooks and workflows to make the incident resolution and triaging process easy, fast, and collaborative between data engineers and analysts.

- Comprehensive query logs that reveal periodic vs. ad hoc queries, changes in query patterns, and more.

“As companies become more data driven, it’s fundamental that organizations not only understand the health of their data, but also have the data observability necessary to trust it from end to end,” said Lior Gavish, CTO, Monte Carlo. “As the data stack fragments to incorporate new tools, it’s becoming increasingly difficult to identify when data pipelines break and take action to fix them. With Incident IQ, data practitioners and leaders alike can holistically understand and respond to issues faster, before they become a serious problem for the business. We believe these features will help customers eliminate hundreds of hours of data downtime and thousands to millions of dollars in savings each month, as well as enable data platform teams to scale with rich post-mortems that track performance and facilitate greater learning.”

Monte Carlo is a Data Observability partner for the FinTech, e-commerce, media, B2B software, and retail industries, counting data teams at Fox, Vimeo, ThredUp, and PagerDuty among their customers.

In February 2021, the company announced their $25M Series B funding, led by Redpoint Ventures and GGV Capital, and was named one of the 2021 Enterprise Tech 30.

Share this

The Latest

April 25, 2024

The use of hybrid multicloud models is forecasted to double over the next one to three years as IT decision makers are facing new pressures to modernize IT infrastructures because of drivers like AI, security, and sustainability, according to the Enterprise Cloud Index (ECI) report from Nutanix ...

April 24, 2024

Over the last 20 years Digital Employee Experience has become a necessity for companies committed to digital transformation and improving IT experiences. In fact, by 2025, more than 50% of IT organizations will use digital employee experience to prioritize and measure digital initiative success ...

April 23, 2024

While most companies are now deploying cloud-based technologies, the 2024 Secure Cloud Networking Field Report from Aviatrix found that there is a silent struggle to maximize value from those investments. Many of the challenges organizations have faced over the past several years have evolved, but continue today ...

April 22, 2024

In our latest research, Cisco's The App Attention Index 2023: Beware the Application Generation, 62% of consumers report their expectations for digital experiences are far higher than they were two years ago, and 64% state they are less forgiving of poor digital services than they were just 12 months ago ...

April 19, 2024

In MEAN TIME TO INSIGHT Episode 5, Shamus McGillicuddy, VP of Research, Network Infrastructure and Operations, at EMA discusses the network source of truth ...

April 18, 2024

A vast majority (89%) of organizations have rapidly expanded their technology in the past few years and three quarters (76%) say it's brought with it increased "chaos" that they have to manage, according to Situation Report 2024: Managing Technology Chaos from Software AG ...

April 17, 2024

In 2024 the number one challenge facing IT teams is a lack of skilled workers, and many are turning to automation as an answer, according to IT Trends: 2024 Industry Report ...

April 16, 2024

Organizations are continuing to embrace multicloud environments and cloud-native architectures to enable rapid transformation and deliver secure innovation. However, despite the speed, scale, and agility enabled by these modern cloud ecosystems, organizations are struggling to manage the explosion of data they create, according to The state of observability 2024: Overcoming complexity through AI-driven analytics and automation strategies, a report from Dynatrace ...

April 15, 2024

Organizations recognize the value of observability, but only 10% of them are actually practicing full observability of their applications and infrastructure. This is among the key findings from the recently completed Logz.io 2024 Observability Pulse Survey and Report ...

April 11, 2024

Businesses must adopt a comprehensive Internet Performance Monitoring (IPM) strategy, says Enterprise Management Associates (EMA), a leading IT analyst research firm. This strategy is crucial to bridge the significant observability gap within today's complex IT infrastructures. The recommendation is particularly timely, given that 99% of enterprises are expanding their use of the Internet as a primary connectivity conduit while facing challenges due to the inefficiency of multiple, disjointed monitoring tools, according to Modern Enterprises Must Boost Observability with Internet Performance Monitoring, a new report from EMA and Catchpoint ...