Aviatrix® announced general availability of the new Aviatrix Network Insights API added to its management interface, Aviatrix CoPilot.
Aviatrix's approach grants control over the cloud data plane, offering a vantage point for monitoring network traffic and performance. With Aviatrix Network Insights API, businesses are now able to access and leverage that insight through their existing observability tools.
At the request of customers who also use application monitoring platforms for their cloud operations, incident management, and app development teams, Aviatrix now adds support for OpenMetrics standards. More than 40 observability vendors now support Prometheus and OpenMetrics. As a first step, Aviatrix is providing a Prometheus endpoint collector.
The traffic insights in Aviatrix CoPilot – previously available in a UI as well as through cloud service provider tools including AWS CloudWatch and Azure Log Analytics – need to be callable by New Relic and other observability tools. Cost constrained customers cannot keep adding staff, but instead seek to gain greater productivity and cost efficiency by enabling small, centralized teams to support hundreds of application workloads and sustain the right level of performance and availability. Aviatrix's launch of Aviatrix Network Insights API marks a significant advancement in cloud network application management, offering cloud administrators a powerful tool for seamlessly accessing network data to third-party analytics and visualization platforms.
Aviatrix Network Insights API provides fine-grained access to a broader set of advanced telemetry and network metrics not provided by cloud APIs. Designed for seamless integration with popular platforms like New Relic and Grafana, Network Insights API facilitates publishing the specific network characteristics a network operations team wants to see, providing deeper understanding of network performance and security for a workload at a given point in time. Network Insights API provides simplified access to over a dozen network performance characteristics, including packet drop rates, bandwidth limit exceptions, and packet-per-second limit exceptions. The new API also gives border gateway protocol (BGP) status updates on the core routing protocol run by Aviatrix on the cloud provider, one of the primary indicators for root cause diagnosing an application failure. Instead of providing a full view of all network resources in CoPilot, the network or cloud operations team responsible for a set of critical applications can more easily isolate insight on the routes, VPCs, and traffic of a given account associated to a given workload.
"We took input in from more than 50 top Aviatrix customers over the last year, spanning financial services institutions, health and life sciences leaders, retail organizations, and government agencies across a multitude of countries. The common request was to show critical network metrics tied to specific applications and to deliver that detailed network traffic flow in the application observability tools already in use," said Josh Cridlebaugh, Principal Product Manager at Aviatrix. "In addition to enrichening these valuable tools – because Aviatrix deploys in AWS, Azure, GCP, OCI, and backbone networks provided by companies like Equinix – the visibility into traffic and availability that Aviatrix delivers is both unique and indispensable to those companies who use two or more cloud providers. That's 85% of all organizations, according to Pluralsight. By offering real-time access to detailed network telemetry, Network Insights API enables our customers' cloud administrators to improve their network management and security strategies while simultaneously helping application development and DevOps teams diagnose network related application issues. This improved observability will yield greater operational excellence and present additional opportunities for cost optimization."
The Latest
Developers building AI applications are not just looking for fault patterns after deployment; they must detect issues quickly during development and have the ability to prevent issues after going live. Unfortunately, traditional observability tools can no longer meet the needs of AI-driven enterprise application development. AI-powered detection and auto-remediation tools designed to keep pace with rapid development are now emerging to proactively manage performance and prevent downtime ...
Every few years, the cybersecurity industry adopts a new buzzword. "Zero Trust" has endured longer than most — and for good reason. Its promise is simple: trust nothing by default, verify everything continuously. Yet many organizations still hesitate to implement Zero Trust Network Access (ZTNA). The problem isn't that ZTNA doesn't work. It's that it's often misunderstood ...
For many retail brands, peak season is the annual stress test of their digital infrastructure. It's also when often technical dashboards glow green, yet customer feedback, digital experience frustration, and conversion trends tell a different story entirely. Over the past several years, we've seen the same pattern across retail, financial services, travel, and media: internal application performance metrics fail to capture the true experience of users connecting over local broadband, mobile carriers, and congested networks using multiple devices across geographies ...
PostgreSQL promises greater flexibility, performance, and cost savings compared to proprietary alternatives. But successfully deploying it isn't always straightforward, and there are some hidden traps along the way that even seasoned IT leaders can stumble into. In this blog, I'll highlight five of the most common pitfalls with PostgreSQL deployment and offer guidance on how to avoid them, along with the best path forward ...
The rise of hybrid cloud environments, the explosion of IoT devices, the proliferation of remote work, and advanced cyber threats have created a monitoring challenge that traditional approaches simply cannot meet. IT teams find themselves drowning in a sea of data, struggling to identify critical threats amidst a deluge of alerts, and often reacting to incidents long after they've begun. This is where AI and ML are leveraged ...
Three practices, chaos testing, incident retrospectives, and AIOps-driven monitoring, are transforming platform teams from reactive responders into proactive builders of resilient, self-healing systems. The evolution is not just technical; it's cultural. The modern platform engineer isn't just maintaining infrastructure. They're product owners designing for reliability, observability, and continuous improvement ...
Getting applications into the hands of those who need them quickly and securely has long been the goal of a branch of IT often referred to as End User Computing (EUC). Over recent years, the way applications (and data) have been delivered to these "users" has changed noticeably. Organizations have many more choices available to them now, and there will be more to come ... But how did we get here? Where are we going? Is this all too complicated? ...
On November 18, a single database permission change inside Cloudflare set off a chain of failures that rippled across the Internet. Traffic stalled. Authentication broke. Workers KV returned waves of 5xx errors as systems fell in and out of sync. For nearly three hours, one of the most resilient networks on the planet struggled under the weight of a change no one expected to matter ... Cloudflare recovered quickly, but the deeper lesson reaches far beyond this incident ...
Chris Steffen and Ken Buckler from EMA discuss the Cloudflare outage and what availability means in the technology space ...
Every modern industry is confronting the same challenge: human reaction time is no longer fast enough for real-time decision environments. Across sectors, from financial services to manufacturing to cybersecurity and beyond, the stakes mirror those of autonomous vehicles — systems operating in complex, high-risk environments where milliseconds matter ...