The PagerDuty Operations Cloud is essential infrastructure for all unplanned, time-sensitive, critical work. It automatically detects and diagnoses disruptive events mobilizes the right team members to respond and automate infrastructure and workflows across your digital operations. This means you can resolve unplanned, unstructured, time-sensitive, and high-impact issues quickly - with fewer escalations to your technical teams while minimizing the impact on your customers and maintaining brand trust.
High customer expectations and increasingly distributed systems mean disruptions to digital service can have catastrophic effects on sales, brand loyalty, and operating costs. The PagerDuty Operations Cloud deflects unnecessary work from teams and subject matter experts so they can focus on delivering business value. Urgent work is escalated to the right teams and routine work is made self-service. Teams can automate and accelerate issue resolutions with minimal human interruption -and improve system resilience and team capacity while reducing the strain of operational complexity and the unexpected.
With more than 700 integrations, APIs, and apps for customer service, the PagerDuty Operations Cloud empowers rapid responses in any environment. And thanks to more than 10 years of data ingestion, its machine learning-powered AIOps functionality can reduce alert noise by up to 98% and drive down MTTR with critical context for faster triage and effective automation.
PagerDuty integrates with various AWS services, including AWS CloudWatch, Amazon GuardDuty, AWS CloudTrail, AWS Personal Health Dashboard, Amazon EventBridge, AWS Security Hub, Amazon DevOps Guru, AWS Control Tower, AWS Outposts, and AWS S3 Storage Lens.
AIOps
PagerDuty AIOps helps teams reduce noise, triage efficiently to drive the right actions towards resolution, and remove manual, repetitive work from the incident response process. Noise reduction baked in with an ML model that learns and adapts based on user behavior means teams see fewer incidents overall. And automating toil from manual event processing results in greater efficiency, saving teams valuable time for innovating.
Process Automation
PagerDuty Runbook Automation is a managed cloud service that enables DevOps teams and SREs to create and delegate operational tasks in automated runbooks to other stakeholders such as developers, NOC personnel, and incident responders. Runbook Automation provides automated workflows and task automation focused on IT and developer process automation. Examples include service provisioning, CI/CD, configuration management, incident diagnosis and remediation, and more. With PagerDuty Runbook Automation, you can resolve requests in minutes, rather than days, optimize security and compliance, and give your engineers more time to spend on innovation rather than firefighting.
Incident Response
PagerDuty helps you save time and money by bringing together the right teams with the right information to resolve incidents faster. Replace manual processes with automation to streamline incident response, freeing up time and resources for more innovation. Orchestrate end-to-end incident response with a service ownership model that only brings in the teams you need. Over 21K organizations trust PagerDuty to help them adopt DevOps best practices and build more resilient operational practices to minimize costly downtime and protect the customer experience.
Custom Private Offer
We can create a custom offer tailored to your needs. Please contact us at aws-sales@pagerduty.com
Highlights
Incident Response - Manage incidents end-to-end
Process Automation - Automate and delegate business and IT processes
AIOps - Maximize IT capacity with fewer incidents and faster resolution
Get personalized pricing in minutes - New
If qualified, an express private offer gets you custom pricing and terms. Finalize your purchase in the AWS Marketplace console.
AWS Marketplace now accepts line of credit payments through the PNC Vendor Finance program. This program is available to select AWS customers in the US, excluding NV, NC, ND, TN, & VT.
Pricing is based on the duration and terms of your contract with the vendor, and additional usage. You pay upfront or in installments according to your contract terms with the vendor. This entitles you to a specified quantity of use for the contract duration. Usage-based pricing is in effect for overages or additional usage not covered in the contract. These charges are applied on top of the contract price. If you choose not to renew or replace your contract before the contract end date, access to your entitlements will expire.
Additional AWS infrastructure costs may apply. Use the AWS Pricing Calculator to estimate your infrastructure costs.
Pricing centers on per-user incident response plans. You pick Professional or Business based on the depth of workflow and admin features your team needs. CustomerServProfessional and CustomerService Business apply the same tiering to support-to-engineering coordination. Several dimensions are usage- or add-on-based: AIOps bills by annual events, with a separate charge for events over your contracted amount. Runbook Automation, Automation Actions, Live Call Routing, and Runbook Auto Job Runner add capabilities on top of a plan. Stakeholder Users sell in bundles of 50, and Status Pages sell in 1,000-user packs.
Top-of-mind questions for buyers
How is AIOps counted, and what happens if I go over my contracted event volume?
AIOps is licensed per accepted event. An accepted event is any valid event sent to and processed by PagerDuty with a successful 2xx response. Your contract covers up to 96,000 annual events. If you exceed that amount, the Additional events over contracted value dimension charges for the overflow.
What counts as a billable user on the Professional and Business plans?
Every person added to your PagerDuty account is a paid user. This includes anyone who receives notifications or appears in an on-call schedule. Read-only business stakeholders are handled separately through Stakeholder Users, which sell in bundles of 50.
Can I buy AIOps, Runbook Automation, or the add-ons on their own?
No. You must first purchase at least one user on a Professional or Business incident response plan before you can add AIOps. Add-ons such as Automation Actions, Live Call Routing, and Runbook Auto Job Runner layer on top of a base plan rather than standing alone.
www.pagerduty.com
Helpful?
Vendor refund policy
All fees are non-cancellable and non-refundable except as required by law.
Request a private offer to receive a custom quote.
How can we make this page better?
Tell us how we can improve this page, or report an issue with this product.
Give us feedbackReport a problem with this product or seller
Legal
Vendor terms and conditions
Upon subscribing to this product, you must acknowledge and agree to the terms and conditions outlined in the vendor's End User License Agreement (EULA).
Content disclaimer
Vendors are responsible for their product descriptions and other product content. AWS does not warrant that vendors' product descriptions or other product content are accurate, complete, reliable, current, or error-free.
SaaS delivers cloud-based software applications directly to customers over the internet. You can access these applications through a subscription model. You will pay recurring monthly usage fees through your AWS bill, while AWS handles deployment and infrastructure management, ensuring scalability, reliability, and seamless integration with other AWS services.
Our team provides multiple resources for customers to find answers to questions and get help with our product. Users may browse our integration guides (pagerduty.com/integrations) to integrate with partner tools, our knowledge base (support.pagerduty.com) to learn more about using PagerDuty, and our developer docs (developer.pagerduty.com) to use our APIs. Additionally, anyone can interact with other PagerDuty users and PagerDuty employees via the PagerDuty Community (community.pagerduty.com). Our Support team is available during regular business hours around the globe, Monday through Friday, and can be contacted at: Email: support@pagerduty.com or via a ticket submitted at tickets.pagerduty.com
AWS infrastructure support
AWS Support is a one-on-one, fast-response support channel that is staffed 24x7x365 with experienced and technical support engineers. The service helps customers of all sizes and technical abilities to successfully utilize the products and features provided by Amazon Web Services.
Machine learning-powered functionality that reduces alert noise by up to 98% through adaptive models based on user behavior patterns
Incident Response Orchestration
End-to-end incident management with service ownership model that mobilizes appropriate team members and automates infrastructure workflows
Runbook Automation
Managed cloud service enabling creation and delegation of automated operational tasks including service provisioning, CI/CD, configuration management, and incident remediation
Multi-Platform Integration
Over 700 integrations, APIs, and applications supporting integration with AWS services including CloudWatch, GuardDuty, CloudTrail, EventBridge, Security Hub, and DevOps Guru
Event Detection and Diagnosis
Automatic detection and diagnosis of disruptive events with machine learning-powered AIOps functionality leveraging over 10 years of data ingestion
Unified Observability Platform
Comprehensive visibility across applications, infrastructure, logs, databases, networks, and digital experiences through a single-pane-of-glass interface
AIOps and Machine Learning
AIOps enhanced with machine learning capabilities to simplify management of distributed environments and automatically prioritize alerts to reduce alert fatigue
Automated Instrumentation and Dependency Mapping
Automated instrumentation with dependency mapping and service relationship views to identify multi-level relationships across services
Open Source and Container Support
Support for open-source frameworks, container technologies, and third-party integrations for cloud-native environments
Rapid Deployment and Integration
Quick installation with automated setup and easy integration with SolarWinds Hybrid Cloud Observability for reduced time to value
AI-Powered Monitoring and Anomaly Detection
Cognitive AI engine automates issue classification, pattern detection, and anomaly identification across managed applications
Multi-Source Data Integration
Combines CloudWatch, CloudTrail logging, and real-time metrics collection for comprehensive application monitoring
Automated Incident Response and Remediation
Provides root cause analysis, corrective action recommendations, and automated response actions integrated with ServiceNow, email, and Slack
Multi-Region Failover and Business Continuity
Supports automatic failover, failback coordination, and state monitoring across multiple AWS regions with automated recovery processors
Application Topology and Resource Management
Delivers topology views, task lists, notifications, and automation runbooks for resource management and operational reporting
Alert noise has been cut and on-call teams respond faster to critical incidents
Reviewed on Aug 10, 2026
Review from a verified AWS customer
What is our primary use case?
Our main use case for using PagerDuty Operations Cloud is because our L1 team was under constant stress from alert storms and missed escalations that were threatening our service reliability. We chose it to streamline the escalation matrices, and it has been excellent for us. It has streamlined our escalation metrics and drastically reduced erroneous notifications, meaning my engineering teams are now only waking up in the middle of the night for actual revenue-impacting emergencies. It has brought stability to our operations and allowed us to respond accurately and swiftly.
The main use case is to ensure that our services run uninterrupted while optimizing our operational cost. We are using it as the central hub for IT alerting, incident management, and automated on-call scheduling across the organization.
How has it helped my organization?
PagerDuty Operations Cloud has definitely been a game changer for the organization. The return on investment has been significant and highly visible. It has helped us reduce alert noise and surface resolution insight. We have saved an average of 30 minutes per incident, which totals a time savings of at least 70% to 80% per year. On a broader operational scale, the built-in ticketing management, the macros, and the automation have saved us the equivalent of at least three to four full-time employees, which has minimized our cost.
Most importantly, it has minimized our downtime by catching signals earlier and translated that to productive revenue, which has avoided an estimated loss worth one million dollars for our company.
What is most valuable?
PagerDuty Operations Cloud has given us very clear AI-driven alert grouping. The autonomous AI agent calculated alerts and suppressed the noise. We used to get temporary traffic spikes or minor upgrade dips, and it has ensured that wherever there is a need for a security alert, it provides the security alert in real time, 24/7.
PagerDuty Operations Cloud includes automated escalation, which helps us set up a call and join the right person via SMS, mobile push, or direct phone call. Another valuable feature is historical data insight, which gives clear insight during an incident and surfaces the historical data on forward-looking resolution from similar past events. This helps the team resolve issues faster.
These features have given us real-time insight and a historical and present-time alert system via SMS, mobile push, or direct phone call. Even if the team has missed an alert at the regular time, it provides historical data and insight to ensure we understand why an incident has been raised or why a security breach has occurred.
What needs improvement?
PagerDuty Operations Cloud is working as expected, but if I were to suggest one improvement, I believe that the AI could be better. The accuracy can be improved because during consistent critical events, it occasionally groups unrelated alerts or misses a correlation, which forces our on-call team to fall back to manual triage.
For how long have I used the solution?
We have been using PagerDuty Operations Cloud in our organization for the last three and a half years.
What do I think about the stability of the solution?
We experienced 100% stability. We did not face any challenges.
What do I think about the scalability of the solution?
It is definitely a scalable product. We can easily scale up and down as per our needs. It has great scale.
How are customer service and support?
Customer support is a 10 out of 10. We had an interaction with them for one of the incidents, and they resolved it within 30 minutes.
What's my experience with pricing, setup cost, and licensing?
Our experience with the pricing, setup cost, and licensing of PagerDuty Operations Cloud has been excellent. We have chosen the pay-as-you-go licensing through AWS Marketplace, which has been a great option for us.
Which other solutions did I evaluate?
Before choosing PagerDuty Operations Cloud, we definitely evaluated other options and alternatives to ensure we were selecting the right solution. Some solutions we considered included Opsgenie by Atlassian, Better Stack, xMatters, and Splunk On-Call.
What other advice do I have?
The alert reduction feature of PagerDuty Operations Cloud has given us at least 30 minutes per incident and has reduced the overall tasks and work that was being done. We have seen a reduction in time and a reduction in the incidents, and we have saved a lot of money as well.
The team has leveraged PagerDuty Operations Cloud to optimize the incident workflows in a positive way. It has given us clear insight into all incidents. We can easily track them in the history section. The impact is significant because it has optimized our workflows, optimized team capacity, and optimized the quality of incidents. The major impact is on time. The team has been allocated tasks on time, and they are able to deliver them on time.
I give PagerDuty Operations Cloud a rating of nine out of ten. Since it has provided us exceptional results, I have given it a nine, and one point is deducted for the improvement I have suggested. Anyone who is looking for a dedicated platform for alerting and automation-related incidents should definitely go with PagerDuty Operations Cloud. They are the best in this category.
Which deployment model are you using for this solution?
Public Cloud
If public cloud, private cloud, or hybrid cloud, which cloud provider do you use?
Amazon Web Services (AWS)
Anonymous
Reliable On-Call Alerting with AIOps Benefits
Reviewed on Aug 03, 2026
Review provided by G2
What do you like best about the product?
I like using PagerDuty for its reliability, as it ensures that we don't miss major critical issues. I use it mainly for on-call schedules and catching alerts when things go wrong in the infrastructure. I find the AIOps and noise reduction features valuable. The setup was pretty simple, which I appreciated.
What do you dislike about the product?
I find PagerDuty can get expensive over time.
What problems is the product solving and how is that benefiting you?
I use PagerDuty for on-call scheduling and catching alerts for critical issues, like server failures. It helps us avoid missing major critical issues.
Marketing and Advertising
Easy-to-Use Urgent Agent Pings with a Great Mobile App
Reviewed on Aug 03, 2026
Review provided by G2
What do you like best about the product?
Useful way to ping agents urgently for cases based on severity. UI in browser is easy to use.
Integrated mobile application is great for agents who are away from the office to keep up-to-date.
What do you dislike about the product?
Inability to modify the notification settings can be irritating.
Browser calendar can become clunky when a large number of agents are on the same roster.
Integration with Slack can break sometimes and leads to distrust of notifications.
What problems is the product solving and how is that benefiting you?
PagerDuty allows my managers to concisely organise agents into schedules that fit around their daily lives.
It allows me to know when I am required access to my work computer, or when I'm free to leave it at home.
Sofia V.
Powerful Incident Response, But Advanced Setup Takes Time
Reviewed on Jul 28, 2026
Review provided by G2
What do you like best about the product?
What I like best about PagerDuty is how it streamlines incident response and ensures the right people are notified immediately when critical issues occur. The alerting and escalation policies are highly customizable, making it easy to manage on-call rotations without worrying that important incidents will be missed.
What do you dislike about the product?
While PagerDuty is a powerful platform, some of its more advanced features come with a learning curve, especially when configuring complex escalation policies, event orchestration, and automation workflows. Initial setup can take time for larger organizations with multiple teams and integrations.
What problems is the product solving and how is that benefiting you?
PagerDuty helps us solve the challenge of responding quickly to critical incidents by ensuring alerts reach the right people at the right time. Instead of relying on manual notifications or monitoring dashboards, incidents are automatically routed through predefined escalation policies, reducing the risk of missed alerts and delayed responses.
reviewer2879727
Unified incident response has reduced alert noise and improves on-call focus and coordination
Reviewed on Jul 26, 2026
Review from a verified AWS customer
What is our primary use case?
I have used PagerDuty Operations Cloud for two years, so I have over two years of experience with it.
My main duties in PagerDuty Operations Cloud are incident routing, escalation policies, on-call management, incident response coordination, and integrations.
For incident routing, I connected AWS CloudWatch alarms to PagerDuty using the Event API. When our EC2 CPU crosses 85%, CloudWatch pushes the alert to PagerDuty, which then routes it to the backend services team. I also used event rules to auto-tag alerts and suppress duplicates. Another use case is on-call management. I used PagerDuty's mobile app to acknowledge alerts during off hours. For example, on weekends, such as Sunday morning, when there is an incident I acknowledged a high-priority alert from my phone and joined the Slack room directly from the PagerDuty incident link. This is helpful for those who are on call during off shifts. Regarding integrations, since we use many integrations on PagerDuty, we integrated PagerDuty with Slack so incidents automatically created a dedicated channel. I also connected it to Jira so incidents could generate tickets for follow-up actions, and also to ServiceNow as well.
What is most valuable?
PagerDuty Operations Cloud excels at routing alerts, managing on-call schedules, coordinating incidents, and automating remediation. These features make it helpful for the on-call team and the people who are working with PagerDuty Operations Cloud.
The feature I use the most is incident routing. We use incident routing more than anything else because it is the main foundation of PagerDuty Operations Cloud. If routing is incorrect, nothing else is going to work, no escalations, no on-call, and no automation at all. I use it the most because it is the first step in every incident; it ensures that alerts from different cloud platforms or any other third-party providers such as AWS, Azure, and other platforms including Datadog reach the right service. It prevents misrouted alerts and missed incidents and it directly impacts the MTTA, because responders get the alert instantly. Every other PagerDuty Operations Cloud feature depends on routing being accurate. I use incident routing daily in this way. This is the main foundation of PagerDuty Operations Cloud, to work with it and to make it a very easy platform for us to work with.
PagerDuty Operations Cloud excels at routing alerts, managing on-call schedules, coordinating incidents, and automating remediation. It ensures critical alerts reach the right teams instantly, reducing the MTTA and MTTR. It strengthened the on-call process with dependable escalations and real-time incident coordination, preventing missed alerts and speeding up resolution tasks. Automation and noise reduction features help cut down repetitive manual work and alert fatigue, allowing engineers to focus on real issues.
PagerDuty Operations Cloud significantly reduced alert noise and improved MTTA. Our acknowledgement time dropped almost by half, and MTTR improved by around 15–20% due to improved automation and routing.
The alert noise dropped by 20–25% after tuning event rules. After implementing PagerDuty Operations Cloud, we saw a 20–25% reduction in alert noise and a noticeable improvement in MTTA. Our acknowledgement time dropped by almost half and the MTTR improved by around 15–20%. MTTR improved by 15–20% due to major incident automation and better routing, which also reduced unnecessary on-call interruptions. It made a good impact on our current schedules and in my current organization.
PagerDuty Operations Cloud's alert reduction features helped lower operational cost by cutting unnecessary alerts and reducing on-call interruptions by grouping and suppressing noisy alerts. The team spent less time triaging low-value incidents, which directly reduced overtime and burnout. The biggest impact for us was fewer false alarms, meaning engineers could focus on real issues instead of constant alert fatigue.
What needs improvement?
PagerDuty Operations Cloud could improve its noise reduction by making deduplication and suppression more automated instead of manually tuned, and the service dependency graph could be more intuitive with cleaner visuals and easier-to-understand root cause tracing during major incidents. Automation could go further with smarter runbook triggers and AI-driven suggestions to help find root causes, which could save a lot of time for engineers who are struggling to understand what is actually happening. These AI capabilities could lower the time by maybe 50–60%. Analytics and reporting could be more flexible, allowing custom dashboards and filters and team-level MTTA and MTTR breakdowns, so that it is segregated based on teams and it is much easier to have custom dashboards for the teams to understand more.
Alert storms were a recurring frustration for the on-call team, and escalation overrides and service dependency graphs can get a bit confusing in larger environments. More customizable analytics and smarter automation would make the platform even easier, more flexible and more powerful for the team to understand.
PagerDuty Operations Cloud could improve its analytics flexibility with customized dashboards and AI capabilities to be more trained and more reliable. Service dependency mapping can also feel cluttered in big environments, making it harder to trace upstream and downstream impacts during bigger incidents.
We implemented PagerDuty Operations Cloud's AI to help with alert grouping and early incident insights, but accuracy was not consistent enough to rely on during critical events. It occasionally grouped unrelated alerts or missed correlations, which limited the operational efficiency gains we expected. The on-call team still depended heavily on manual triage because AI suggestions were not always aligned with the real root cause. Overall, AI added some value but it has not yet reached the reliability needed to significantly improve the incident response efficiency. It needs more training or more work.
What do I think about the stability of the solution?
PagerDuty Operations Cloud is stable in day-to-day use with no major outages or reliability issues affecting our on-call workflows. Alert delivery, incident routing and escalation chains have been consistently working without delays or missed notifications. Its cloud-native architecture also means updates and patches roll out smoothly without impacting uptime. Overall, the stability has been one of the strongest features for our team.
What do I think about the scalability of the solution?
PagerDuty Operations Cloud scales well for growing teams. Adding services, integrations and new on-call groups does not impact performance. Its native cloud architecture handles higher alert volumes smoothly, especially when paired with strong incident routing. We have seen it manage increased workloads without slowing down or causing notification delays. On the whole, scalability has been reliable even as our environment and service footprint has expanded.
How are customer service and support?
The customer support has been responsive and generally helpful when we have raised any issue. They provided clear troubleshooting steps and follow-ups. Resolution speed can vary depending on the severity of the issue, but support has been good and reliable for day-to-day operational needs.
Which solution did I use previously and why did I switch?
Before PagerDuty Operations Cloud, we used basic native alerting tools such as AWS CloudWatch since AWS is our primary cloud platform. We switched because those tools lacked valuable incident routing and consistent on-call escalation workflows. PagerDuty Operations Cloud offered stronger coordination, better noise reduction, and more mature incident response features. The move was mainly driven by the need for a unified and dependable on-call and incident management platform.
What was our ROI?
We have actually seen a moderate return on investment mainly through reduced alert noise and more efficient incident routing. Alert storms dropped by roughly 20–25%, saving the team several hours per week that used to be spent on low-value triage. We estimate a small cost reduction from fewer overtime hours and less on-call fatigue, though not enough to reduce headcount. Overall, the return on investment is positive but incremental — more time saved than money saved.
What's my experience with pricing, setup cost, and licensing?
PagerDuty Operations Cloud's pricing felt reasonable but it definitely is not the cheapest option. The value comes more from reliability than cost savings. Setup costs were minimal since it is a SaaS platform and onboarding did not require any infrastructure or hidden implementation costs or fees. Licensing is straightforward but scaling seats for larger teams can get expensive, especially when adding advanced features. Overall the experience was smooth but the price point could be more flexible for growing teams. Our team is small now but in the future it might become bigger.
Which other solutions did I evaluate?
Before adopting PagerDuty Operations Cloud, we evaluated options such as Opsgenie and Splunk On-Call for incident management. We also looked at other native cloud tools such as AWS CloudWatch, but they lacked the escalation workflows. Opsgenie had good features but did not match PagerDuty Operations Cloud's reliability and integrations for bigger teams. Overall, PagerDuty Operations Cloud offered stronger coordination and more consistent on-call performance which made us choose it.
What other advice do I have?
Escalation policies are a very helpful feature for managing the team, such as multi-step escalation logic which is extremely dependable and prevents missed alerts, and also the automation runbook. I rely on incident routing daily because it ensures the alerts reach the right team. The other features, such as integrations, are really good to integrate PagerDuty Operations Cloud with different platforms, such as Jira and ServiceNow platforms to automatically open an incident. These are really helpful top features.
I would give PagerDuty Operations Cloud eight out of ten, considering all the features I have been using and also the improvements I have suggested. I chose eight out of ten because PagerDuty Operations Cloud is generally strong in the areas that matter the most, especially incident routing and escalation management which are extremely reliable and directly reduce MTTA and MTTR. It stands out for dependable on-call orchestration and smooth incident response coordination, especially with Slack and Jira integrations. But it does not reach a ten because noise reduction still needs more automation, analytics lack deeper customization such as custom dashboards to filter among the teams, and the service dependency graph gets cluttered in large environments. With smarter automation and improved visualization of the analytics, it could easily move closer to a perfect ten.
Which deployment model are you using for this solution?
Public Cloud
If public cloud, private cloud, or hybrid cloud, which cloud provider do you use?