
PagerDuty announced new generative AI (genAI) and automation features of PagerDuty Advance, which is embedded across the PagerDuty Operations Cloud platform, in collaboration with Amazon Web Services (AWS).
The latest collaboration combines the capabilities of PagerDuty Advance with Amazon Q Business, Amazon Bedrock and Amazon Bedrock Guardrails, empowering organizations to safely deploy genAI into their incident management processes while becoming more automated, connected and operationally resilient.
“AWS and PagerDuty share a commitment to operational excellence, and together we help organizations around the world reliably serve millions of enterprise customers at scale, with speed and resilience,” said Matt Garman, CEO of AWS. “We’re expanding our longstanding partnership to help businesses safely apply generative AI to make operations more automated, responsive, and efficient. This new collaboration helps our customers proactively prevent and resolve issues to maintain the highest reliability standards for their digital operations.”
“Today’s customers demand flawless digital experiences, making reliable technology essential for achieving ‘always-on’ operational resilience,” said Jennifer Tejada, Chairperson and CEO, PagerDuty. “PagerDuty and AWS are strengthening our 11-year partnership, which already supports nearly 6,000 joint customers, to further integrate generative AI into digital operations management. Together, we are driving transformative outcomes — helping organizations grow revenue, accelerate innovation, and reduce risk with confidence."
Through the power of AI and automation, the PagerDuty Operations Cloud detects and diagnoses disruptive events, mobilizes the right team members to respond, and streamlines infrastructure and workflows across digital operations. The launch of these new AI and automation tools will help joint customers of AWS and PagerDuty enhance their operational efficiency and redefine what it means to respond to incidents faster and smarter.
PagerDuty Advance integration with Amazon Bedrock: PagerDuty Advance assistant for Slack and Microsoft Teams delivers AI-powered incident context support through chat. These capabilities are powered by Amazon Bedrock, a fully managed service on AWS offering customers a broad set of capabilities to easily build, deploy and scale generative AI applications. PagerDuty Advance is embedded across the PagerDuty Operations Cloud platform, to improve situational awareness and accelerate triage during an incident. Using PagerDuty Advance with Bedrock, incident responders can more easily get answers to questions, such as “Are there any other affected services?”, “Has this incident occurred in the past?” or “What’s the next best step toward resolution?” — without having to ask others these questions. It is estimated that organizations could save hundreds of thousands per incident, factoring in the cost impact of incidents and staffing costs.
PagerDuty Advance integration with Amazon Bedrock Guardrails: While genAI and automation can help save time and create significant efficiencies, implementing safeguards can be critical to safer and more accurate query responses. The integration of PagerDuty Advance with Amazon Bedrock Guardrails helps to improve relevance and factual accuracy in genAI responses, which is essential to effective incident management. Additionally, the integration offers certain protections against hallucinations and helps block undesired topics and harmful content, such as malicious text inputs from generating responses.
PagerDuty Advance plugin integration with Amazon Q Business: PagerDuty is the first incident management platform to integrate with Amazon Q Business, one of the most capable genAI-powered assistants for leveraging companies’ internal data. PagerDuty Advance customers can now utilize a single user interface via Amazon Q Business plugins to query on data from multiple applications, including PagerDuty Advance. Using the new Amazon Q plugin, this centralized source of truth reduces the need to shuffle between third-party applications to retrieve critical incident data. In interviews with PagerDuty Advance early access adopters, customers indicated that they saved on average 30 minutes per incident. This could mean hundreds of thousands of dollars saved per incident, factoring in the cost impact of incidents and staffing costs.
PagerDuty Advance integration with Amazon Bedrock is now generally available in all regions in which PagerDuty operates.
PagerDuty Advance integration with Amazon Bedrock Guardrails is now generally available in all regions in which PagerDuty operates.
PagerDuty Advance plugin integration with Amazon Q Business is now generally available in the U.S. service region.
The Latest
Rapid AI adoption and the unique ways AI workloads operate is redefining the scope and structure of what these teams must deliver. This shift is forcing organizations to rethink how they manage scale, automation, and control, according to The State of SRE and Platform Engineering 2026, a new report from Dynatrace ...
AI is usually talked about as a software tool, but it also depends heavily on the network behind it. Whether a company is using AI for chatbots, automation, monitoring, analytics, or employee support, all of that information has to move across the network in a reliable and secure way. That means AI is not just an application decision. It is also an infrastructure decision. Before organizations rush into AI, they should ask a simple question: Is our network ready to support it? ...
Enterprise AI often lacks governed access to where business processes actually execute. Without that access, AI agents may be able to reason, but they cannot operate reliably across enterprise workflows. For AI agents to effectively carry out workflows, they will require integration-layer context and controls. Organizations can implement these prerequisites by providing AI with managed access to the middleware layer ...
Enterprise networks rarely behave the same way for very long. A routing adjustment in one region may unexpectedly alter application performance in another. A cloud migration may introduce hidden dependencies that go unnoticed until an outage occurs. All the while, the network is managed by several different teams, each of whom use different tool sets — and as a result, have different views of the network ... There’s usually an engineer who remembers why traffic fails over a certain way between sites, or which transparent firewall was added where. The problem is that human memory cannot scale alongside enterprise-scale networks ...
Ask an infrastructure team how confident they are in their ability to govern AI, and most will tell you they've got it handled. A recent survey of 406 IT decision-makers and platform engineering leaders found 86% expressing exactly that confidence. Ask the same group whether they have a formal written AI governance policy, and the number drops to 30%, according to Spacelift's Infrastructure Automation Report ...
In MEAN TIME TO INSIGHT Episode 27, Shamus McGillicuddy, EMA VP of Research, Network Infrastructure and Operations, and Parker Hathcock, EMA Research Director covering IT Service/Operations (ServiceOps), discuss observability unification in modern IT operations ...
Virtual Private Networks became a cornerstone of enterprise security at a time when corporate infrastructure looked very different from today ... For years, this model worked well. But the architecture behind VPNs assumed a centralized corporate environment—one where the network itself was the hub of activity. In a cloud — first world, that assumption no longer holds ...
Website outages get resolved just as fast in August as they do in November. I went looking for the opposite: the summer slowdown everyone assumes is there once the people who fix things are away. It isn't in the data we collected, covering 1.8 million confirmed outages across tens of thousands of websites ...
This year, many of the cloud infrastructure contracts signed in the early days of the AI boom will come up for renewal. As the year goes on, I anticipate we'll see a significant amount of cloud vendor swapouts and multi-cloud adoption, and the reason isn't just GPU depreciation. It's because they're tired of their current cloud providers ...
There's a moment the many observability teams have experienced days into bringing a new service into production: you realize that the vendor's claims of "intelligent" behavior included a large serving of hype. Their dashboards look nice until they don't, the failure modes are a black box, and no one on the team can confidently explain why the system did what it did at 2 am. Agentic AI is about to force every Ops team to relive that moment at web-scale until they start treating these systems as the dependencies they actually are ...