PagerDuty Operations Cloud logo

    PagerDuty Operations Cloud

    The PagerDuty Operations Cloud is essential infrastructure for all unplanned, time-sensitive, critical work. It automatically detects and diagnoses disruptive events mobilizes the right team members to respond and automate infrastructure and workflows across your digital operations. This means you can resolve unplanned, unstructured, time-sensitive, and high-impact issues quickly with fewer escalations to your technical teams while minimizing the impact on your customers and maintaining brand trust.

    Ratings and reviews

    4.3
    68 ratings
    2 star
    1 star
    57%
    40%
    3%
    0%
    0%
    40 AWS reviews
    |
    28 external reviews
    External reviews are from PeerSpot .

    Filters

    Review type

    AWS Marketplace reviews
    External reviews
    Reviews (68)
    Yuvashree Yuvashree

    Intelligent incident automation has improved on-call routing and reduced downtime risks

    Reviewed on Sep 21, 2026
    Review provided by PeerSpot

    What is our primary use case?

    My main use case for PagerDuty Operations Cloud is integrating with ITOM, incident management, and also using PagerDuty for automation.

    A quick specific example of how I use PagerDuty Operations Cloud for incident management is that I commonly use it for critical server down alerts for certain application performance issues, database capacity alerts, network device failures, etc.

    In addition to that, for a network device failure when a core switch or a router goes offline, SolarWinds detects the node status is down alert. So whenever an alert is detected, PagerDuty Operations Cloud opens a P1 incident and relevant teams like the NOC team, network team, and incident management team get notified. They automatically get paged through PagerDuty Operations Cloud, and if the incident is not acknowledged, it gets escalated to the managers. This is the automation use case which I created for the network device failure alerts.

    The feature I find myself using the most is the best use case from Datadog to PagerDuty Operations Cloud to the Linux team, which I have created. Datadog detects CPU memory or disk or service failure alerts for the Linux server, and PagerDuty Operations Cloud automatically creates the incident and notifies the Linux on-call engineer.

    What is most valuable?

    According to PagerDuty Operations Cloud, the best features offered are intelligent incident management, AI Ops and alert noise reduction, automation and runbook execution, cloud and hybrid infrastructure monitoring, ChatOps integration, AI-powered operations, service ownerships, and business visibility.

    PagerDuty Operations Cloud has positively impacted my organization by reducing mean time to resolution, automatically routing incidents to the correct on-call engineers, reducing alert fatigue, whereas PagerDuty Operations Cloud AI Ops correlates duplicate and related alerts. It also provides 24/7 operational coverage, faster incident management, increased automation, better visibility and accountability, and improved knowledge management.

    What needs improvement?

    PagerDuty Operations Cloud can be improved by eliminating alert storms, automating common incident resolutions, improving major incident management, using AI for faster troubleshooting, and improving operational metrics.

    I could add a feature to automate common incident resolution where engineers perform repetitive actions like restarting servers, clearing disk space, collecting logs, or running diagnostics.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for almost six plus years.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is stable.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is highly scalable and is desired for organizations ranging from small operation teams to larger enterprises managing thousands of services, responders, and alerts.

    How are customer service and support?

    PagerDuty's customer support is generally considered strong from enterprise customers, particularly those running mission-critical operations and requiring 24/7 incident management.

    Which solution did I use previously and why did I switch?

    Before adopting PagerDuty Operations Cloud, I primarily relied on monitoring platforms such as Datadog and SolarWinds and native cloud monitoring solutions for alert generation.

    What was our ROI?

    Organizations commonly see ROI from PagerDuty Operations Cloud in both time-saving and downtime reduction, especially when integrated with monitoring platforms and automation workflows.

    What's my experience with pricing, setup cost, and licensing?

    From my experience evaluating and implementing PagerDuty Operations Cloud for incident management and on-call operation, PagerDuty Operations Cloud offers tiered plans that cost varies based on features such as incident management, AI Ops, event orchestration, automation, and AI capabilities.

    Which other solutions did I evaluate?

    During the evaluation phase, I considered other incident management platforms such as ServiceNow, Opsgenie, etc., before selecting PagerDuty Operations Cloud.

    What other advice do I have?

    PagerDuty Operations Cloud helps my team focus on core tasks, which means it helps engineers spend less time on operations overhead and more time on engineering work.

    My team has leveraged automation within PagerDuty Operations Cloud by integrating it with monitoring platforms such as Datadog, SolarWinds, and cloud monitoring tools to automate the end-to-end incident management process. Alerts are automatically correlated, routed through the appropriate on-call engineers, escalated when required, and tracked through resolution. The team also leverages automated runbooks and AI-assisted triage to reduce manual intervention during incidents.

    PagerDuty Operations Cloud's alert reduction feature has significantly reduced the risk of costly outages in my organization by ensuring critical issues are identified, routed, and escalated to the appropriate responders in real-time. Automation, notifications, intelligent alert grouping, and escalation policies help prevent incidents from being overlooked, reducing downtime and business impact.

    I use workflows to standardize the incident response process by automatically engaging the correct support teams, creating collaboration channels, notifying stakeholders, and tracking incidents through resolution. I have also utilized event orchestration to filter, enrich, suppress, and route alerts from monitoring tools such as Datadog and SolarWinds.

    Organizations get the highest value from PagerDuty Operations Cloud when they use it for full operation platforms combining incident management, event orchestration, AI Ops, automation, governance, and service ownership.

    Regarding governance and security, PagerDuty Operations Cloud provides strong controls that help the cloud operations team maintain compliance, accountability, and operational resilience.

    For cloud operations, AI is valuable only if it is accurate, reliable, and actionable. PagerDuty Operations Cloud improves this through an AI-first operation platform, which is built on large-scale operational data, event correlation, and incident response workflows.

    I rate this product a 10 out of 10.

    MayankGarg

    Centralized incident routing has streamlined on-call response and reduced downtime

    Reviewed on Sep 18, 2026
    Review provided by PeerSpot

    What is our primary use case?

    I primarily use PagerDuty Operations Cloud for managing incidents and monitoring all critical services on our application. I rely on it most for real-time alerts, automated incident routing, on-call coordination, and quickly bringing the right team members into an issue when something goes wrong with our production applications.

    A recent issue occurred when one of our production services started showing a spike in errors. PagerDuty Operations Cloud triggered an alert and automatically routed the incident to our on-call engineer. We were able to acknowledge and quickly coordinate the response while keeping track of updates in one place, making it easy to identify issues and avoid delays in determining who should handle the problem and how to reduce the time to restore service to normal so we could maintain the same user experience.

    PagerDuty Operations Cloud fits well into my daily workflow because it provides one place to manage all alerts, incidents, and on-call responsibilities. I especially appreciate having clear ownership and escalation when an issue needs immediate attention. It reduces manual coordination and makes incident response more organized and structured.

    What is most valuable?

    Several features stand out for me, including automated alerting, intelligent incident routing, on-call scheduling, and escalation policies. I find it especially useful that alerts reach the right person automatically instead of relying on manual communication. The incident timeline and collaboration features make it easier to understand what happened and keep everyone aligned during an outage.

    Escalation policies make the biggest difference for my team because they ensure that if the first person does not respond, the incident automatically moves to the next person instead of getting missed. This provides us more confidence during critical issues, especially outside normal working hours.

    PagerDuty Operations Cloud has made our incident response more organized, well-structured, and consistent. Instead of relying on manual communications, alerts are automatically routed to the appropriate team and escalation ensures issues do not get overlooked. It has also improved collaboration during an incident because everyone can see the status and ownership clearly. Overall, my team spends less time coordinating and more time actually resolving problems.

    What needs improvement?

    One area I would like to see improved is making complex alerts configuration and routing rules easier to set up and manage.

    PagerDuty Operations Cloud's flexibility is useful, but some advanced configuration can take time to understand. Simplifying the setup of routing rules and escalation policies with clear guidance for complex workflows would make the platform easier for new team members to adopt.

    The platform feels a little complex when setting up advanced routing rules, escalation policies, and integrations. For new users, it takes time to understand how everything fits together. I would like to see simpler configuration workflows and clear guidance, particularly for new users or new team members. Once everything is configured, the day-to-day incident management experience is fairly smooth.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for two years.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is very stable.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is very efficient, robust, and scalable.

    How are customer service and support?

    Customer support is wonderful and always helpful.

    Which solution did I use previously and why did I switch?

    We are using this for monitoring all our production applications.

    What was our ROI?

    We see a good return on investment. It saves our time as well as money.

    What's my experience with pricing, setup cost, and licensing?

    Pricing is good and not too cheap.

    What other advice do I have?

    I would always recommend PagerDuty Operations Cloud to any organization because it monitors all production applications, enhances user experience, and reduces application downtime.

    I found the AI output to be generally accurate and reliable for common incident management tasks, especially when summarizing incidents and highlighting relevant information. It gives us a useful starting point and can save time when reviewing an issue. I still treat AI-generated recommendations as a support tool and verify important details before taking action, particularly during critical incidents.

    I give this product a rating of 9 out of 10.

    Kishan Chand

    Incident automation has reduced alert noise and now improves response times for critical issues

    Reviewed on Sep 17, 2026
    Review from a verified AWS customer

    What is our primary use case?

    I use PagerDuty Operations Cloud as my primary incident management system and automated escalation platform because it integrates directly with my monitoring tools like CloudWatch and Datahog to reduce aggregate alerts and alert noise through duplication. It centralizes real-time alert triage and automatic incident response across my infrastructure and microservices. It automatically logs issues and transfers them to my IT team to resolve as soon as possible. Most significantly, it is lowering my mean time to acknowledge (MTTA) and mean time to resolve (MTTR).

    In a recent situation, I had a production service where response time suddenly increased. PagerDuty Operations Cloud detected the issue and triggered an alert to the right engineer over a call. I saw the incident details and escalation workflow to quickly identify the affected service and bring the right engineer team to troubleshoot the issue so that user experience would not be hampered. The issue was resolved before it had a major customer impact. I found it particularly useful that PagerDuty reduced the time between detecting the problem, notifying the right person, and getting it resolved as soon as possible.

    What is most valuable?

    I see many valuable features, but the most useful are incident alerting, incident escalation, on-call scheduling, integration with my monitoring tools, and better collaboration. There is a centralized dashboard where all incidents are logged, where I can track the responses and see who is handling the issue. The mobile application is also very useful because it automatically alerts me when I am away from my desk so that I can fix and resolve issues from anywhere.

    The mobile application is particularly helpful when I am away from my laptop. For example, when a critical incident occurred outside normal working hours, PagerDuty notified me on my phone and I acknowledged the alert and checked the incident details right away. I was able to share the details with my IT team without waiting for someone to return to their desk. This ability to stay updated and respond from the mobile app is especially useful for on-call situations.

    PagerDuty Operations Cloud is creating a positive impact on my enterprise by making incident response more structured and organized. Alerts are automatically routed to the right engineer and escalation policies ensure that important incidents are not missed. It has overall improved communication during incidents because everyone can see the status and take ownership. It has helped me reduce delays in responding to operational issues and has improved coordination across all teams.

    What needs improvement?

    I feel the user interface is not optimal for new users and needs some optimization because some configuration takes time to understand, especially when adding rules and workflows. The platform is otherwise good and the experience is wonderful and easier for my team.

    The setup could be simpler, especially for new users who want to add integrations. Better guidance during the initial configuration would make it easier to get started. Some improvements could be made to the administrator features and to make the platform easier to manage as the team grows or increases.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for the last two years.

    What do I think about the stability of the solution?

    I have experienced stability issues.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is very scalable.

    How are customer service and support?

    Customer support is excellent.

    How was the initial setup?

    Pricing is reasonable and good for my organization. It is very easy to set up.

    What was our ROI?

    I see a good ROI of approximately 50 to 60 percent because it makes my application more robust and reduces the need for my IT team to do manual work and troubleshoot issues.

    What other advice do I have?

    AI capabilities are always useful in terms of governance and security, and they give me clear control around all data access, permissions, and auditability. AI gives me incident information quickly. In my experience, I am using AI for recommendations on troubleshooting issues. The AI feature helps me handle issues much quicker and faster.

    AI is generally useful for getting better incident response and identifying the next steps I can take to resolve issues. Recommendations are always helpful when I am spending time troubleshooting issues and monitoring incident details. The value is mainly in reducing the time spent analyzing the issue while keeping human review as a part of the process.

    Alert reduction features have helped me cut down duplicate or low priority notifications so that my team can focus on alerts that actually need attention and have high priorities. It reduces unnecessary escalation so that I can spend less time investigating issues and respond to genuine incidents sooner.

    PagerDuty Operations Cloud helps me reduce the time spent on incident triage and follow-ups. Alerts are easier to prioritize so that my team can identify the right responder faster. I have seen less time spent manually coordinating incidents and keeping track of routine alerts. It also gives me more flexibility so that I can customize reporting and create my own dashboards and workflows to easily detect and resolve issues as soon as possible.

    I have heavily utilized event orchestration and automatic incident workflows to streamline my process. By setting up rules that automatically suppress alerts for non-critical issues, I can resolve genuine high-severity issues more effectively. I integrated automated diagnostic runbooks that execute immediately upon incident creation, gathering essential logs and service health metrics before an engineer even logs in. The automation eliminated repetitive manual overhead during outages and removed significant alert fatigue across my team, allowing them to redirect their focus toward planning project work and system resilience. This enables my engineers to work simultaneously to resolve issues more quickly and efficiently.

    I use PagerDuty SRE the most because it helps automate incident response, identify issues quickly, and reduce the time needed to resolve incidents. It makes on-call work more efficient and helps my team focus on high-priority tasks.

    I always suggest using PagerDuty Operations Cloud to monitor all production applications. I rate this solution a 10 out of 10.

    Harjoth Sudan

    Structured alerting has reduced recurring incidents and enables faster resolution for our team

    Reviewed on Sep 04, 2026
    Review provided by PeerSpot

    What is our primary use case?

    My main use case for PagerDuty Operations Cloud is to manage alerts and to monitor AI-powered workflows. For example, whenever a service goes down, it needs attention, and PagerDuty Operations Cloud helps in keeping track of all these things.

    PagerDuty Operations Cloud helps me when a service goes down by notifying everyone. Prior to this, we were using Slack, but we were relying on messages, and this platform gives more clear visibility to the team rather than a message.

    PagerDuty Operations Cloud identifies the right person and notifies that person, so it is way faster and efficient.

    What is most valuable?

    The best features PagerDuty Operations Cloud offers for us include incident management and automated alerting, which are helping us. These are the only reasons we started and onboarded this.

    Automated alerting and incident management features make my team's life easier on a day-to-day basis. We have set the hierarchy, and if the first person does not respond, it moves to the second person, so that way we do not miss any escalations or issues.

    PagerDuty Operations Cloud has positively impacted my organization by making us more structured. All the production-related issues which we were facing before are sorted.

    Since using PagerDuty Operations Cloud, I have noticed specific outcomes such as faster incident resolution. Recurring issues were monitored properly, and recurring incidents were corrected and have stopped.

    What needs improvement?

    PagerDuty Operations Cloud can be improved. The product itself is a bit complicated, and the initial configuration takes time, making it very difficult for new members to understand.

    The product is strong, but making the setup simpler and easy to operate is an area that can be improved.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for the last year.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is stable.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is highly scalable.

    How are customer service and support?

    Customer support is good. I would rate customer support nine out of ten.

    Which solution did I use previously and why did I switch?

    We have not used any different solution previously.

    What was our ROI?

    I have seen a return on investment. The ROI is great for us as repetitive incidents have decreased, the time in solving those issues has also decreased, and the team needs more projects and fewer team members.

    What's my experience with pricing, setup cost, and licensing?

    My experience with pricing, setup cost, and licensing revealed that licensing is fine. However, the pricing was a bit high, and the setup, as I mentioned, is a bit complicated, costing us four weeks and roughly one hundred dollars.

    Which other solutions did I evaluate?

    Before choosing PagerDuty Operations Cloud, one of the team heads suggested this to us, and we implemented it because they were using this in their previous organization.

    What other advice do I have?

    Regarding PagerDuty Operations Cloud's AI capabilities, I think its governance and security are trying to cope up. For now, this has not been a challenge for us, but we keep an eye always.

    Regarding PagerDuty Operations Cloud's AI capabilities, I find its accuracy and reliability of output to be very accurate and quite reliable.

    My advice to others looking into using PagerDuty Operations Cloud is that it helps a lot with production issues and solving them on time. I suggest they go ahead and set this up properly, as initially it takes time but will provide ROI.

    PagerDuty Operations Cloud is one of the reliable solutions that help in solving production issues. I have rated this review eight out of ten.

    Aaron Held

    Automated alerting has streamlined incident response and improved team coordination

    Reviewed on Aug 26, 2026
    Review provided by PeerSpot

    What is our primary use case?

    The usual use cases for PagerDuty Operations Cloud include incident management. I ran the infrastructure teams and SRE, so the primary use case was incident management and spinning up RCAs, tracking remediation. We also used it for identifying key stakeholders of projects. We relied on the escalation chains and the groups. Even if it wasn't an incident, if somebody had a question or they needed to get a hold of the CMS team, they would go into PagerDuty. We had a lot of integrations into Teams chat.

    The heart of it was obviously kicking off pages from alerts within PagerDuty Operations Cloud. It was fully automated so that when an anomaly was detected, it automatically created an incident in PagerDuty. That automation was just table stakes. We also had some good automations to follow up and close out Jira tickets that would integrate with the PagerDuty incident. The primary one was simply calling the people when something happens.

    What is most valuable?

    My favorite feature about PagerDuty Operations Cloud is that it just worked. It was one of the tools we didn't have to worry or think about. It was very well polished. It did what they said they were going to do, and it didn't require a lot of customization. I think it facilitated us organizing the teams, groups, and escalation chains really well. However, there is no feature that stands out significantly.

    PagerDuty Operations Cloud absolutely requires maintenance on my end. Escalation chains and people get stale very quickly. I wish it would integrate into org charts or an HR system so that when people get terminated or shift roles, PagerDuty knows about the changes. We had one escalation chain where it went up to an executive that wasn't in the company anymore.

    What needs improvement?

    The users would complain that they always had a hard time maintaining their schedules, their escalations, and installing the app. I got a lot of complaints from the people that responded to pages rather than from the people who administered PagerDuty.

    The complaints I remember were about how users wanted to be notified. It was very common for somebody to misconfigure their own communication chain and not test it, so when an alert went off, their excuse was they didn't get the page or they didn't get the SMS. I got a lot of feedback from users that they struggled to do their shift when they needed someone to cover their shift. Often that wouldn't get done. Anything that can be done to make that as easy and simple as possible would be helpful.

    For how long have I used the solution?

    I have probably used PagerDuty Operations Cloud for about five years total.

    What do I think about the stability of the solution?

    The stability of PagerDuty Operations Cloud has been great. I have not experienced any stability issues.

    What do I think about the scalability of the solution?

    I have no problems with scalability. I have used it for like 500 engineers and experienced no issues.

    How are customer service and support?

    I did not have any partnerships with them. I was just a customer and user.

    Which solution did I use previously and why did I switch?

    I have used some alternatives. I used Grafana incident management a lot and we used something called RingCentral.

    How was the initial setup?

    The initial deployment for a new client when starting with PagerDuty Operations Cloud was moderate. I would say it was more difficult because the biggest problem we had was just identifying who the responders were. It does take a bit of setup. We also did the Teams integration, which requires security and permission. The setup is always hard, though this is not necessarily a criticism on PagerDuty.

    What's my experience with pricing, setup cost, and licensing?

    I am not current on what the current pricing is for PagerDuty Operations Cloud, but I assume it is per seat and per responder.

    Which other solutions did I evaluate?

    If I were to pick one, PagerDuty Operations Cloud is the one I would pick if there were no budget concerns. If I was budget-constrained, I would pick Grafana.

    What other advice do I have?

    PagerDuty Operations Cloud did influence the volume and nature of alerts. We had more of them. I have implemented this three major times in my career, and the number of incidents immediately shot up and then came down. What happens is you identify the incidents more aggressively and then you actually fix the problems. I think that is a good thing.

    I have not used Incident Workflows or Event Orchestration to automate toil from the incident management process. We played with it, but I did not use them in practice.

    I did not use the AI agents at all. I would rate this review a 9.

    Aryan Dwivedi

    Real-time alerts have protected client sites and now keep our team focused on critical issues

    Reviewed on Aug 25, 2026
    Review provided by PeerSpot

    What is our primary use case?

    PagerDuty Operations Cloud is used for real-time alerts when something goes wrong with our system so the team can fix it right away, and for setting up on-call schedules to ensure the right person is always ready to answer the problem at any time of the day. We also use it for scaling up discovery calls with our clients and prospects.

    Last month, one of our main web servers stopped working in the middle of the day, and the system sent an instant notification to the team, who automatically used our on-call schedule to contact the right engineer on duty, saving us precious time instead of guessing who to call.

    We have used PagerDuty Operations Cloud to control the external noise from the alerts.

    What is most valuable?

    The real-time notification is top-notch, and critical warnings reach the right person right away without delay. Smart tools that reduce extra alert noise are great for keeping the team focused on real-time problems. The large number of connections to over 700 different software tools allows us to bring all work together in one workflow.

    PagerDuty Operations Cloud has helped fix critical problems much faster, including the ads we are running that do not crash, the websites we run for our clients, and our CRM website.

    I cannot say that PagerDuty Operations Cloud has reduced downtime or improved our team's response time because we have been using it for around two months only. However, we have seen reduced downtime and quick resolution of issues because everyone is notified when something is damaged, whereas earlier we would only find out about issues when the client called us. This also gave us a demerit because the client thought we could not do the work properly, resulting in a significant reduction in our monthly retainer. Now we are stable with our monthly retainer.

    The security governance of PagerDuty Operations Cloud is pretty good.

    Regarding its accuracy and reliability of output, I would say it is pretty much accurate.

    PagerDuty's generative AI is pretty good, although sometimes the decision can be a bit generic and not very reliable, but they are improving it.

    What needs improvement?

    The learning curve for PagerDuty Operations Cloud is a bit steep, with its complicated rules and schedules, and we have to dedicate a proper team to use that tool.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for two months.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is pretty much stable and I have not seen it crashing.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is pretty much scalable.

    How are customer service and support?

    I have not used customer support until now, but I believe they will be great.

    Which solution did I use previously and why did I switch?

    We were not using any different solution; this is the first time we are going for an on-cloud platform.

    What about the implementation team?

    I cannot share how my team has leveraged automation within PagerDuty Operations Cloud because it is a bit confidential.

    What was our ROI?

    It helps us save manual hours of work every week.

    For me as an employee, I would say the time has been reduced and the downtime has been reduced.

    Which other solutions did I evaluate?

    I was not on the team that evaluated options before choosing PagerDuty Operations Cloud, so I cannot tell you about this.

    What other advice do I have?

    We are still in the process of implementing AI and automation through PagerDuty for incident response.

    I cannot answer that question right now because we are still testing PagerDuty's autonomous AI agents.

    The alert reduction feature has been reporting to us issues such as a client's website being down or the CRM website being down, which I would say has saved us.

    I would say that if your team can purchase PagerDuty Operations Cloud, you can definitely use it, although it is a bit costly for a startup, as was pointed out by my team too, but we have saved a lot of time.

    I gave this review a rating of 9 out of 10.

    Raj kuruhuri

    Alert noise has been cut and on-call teams respond faster to critical incidents

    Reviewed on Aug 10, 2026
    Review from a verified AWS customer

    What is our primary use case?

    Our main use case for using PagerDuty Operations Cloud is because our L1 team was under constant stress from alert storms and missed escalations that were threatening our service reliability. We chose it to streamline the escalation matrices, and it has been excellent for us. It has streamlined our escalation metrics and drastically reduced erroneous notifications, meaning my engineering teams are now only waking up in the middle of the night for actual revenue-impacting emergencies. It has brought stability to our operations and allowed us to respond accurately and swiftly.

    The main use case is to ensure that our services run uninterrupted while optimizing our operational cost. We are using it as the central hub for IT alerting, incident management, and automated on-call scheduling across the organization.

    How has it helped my organization?

    PagerDuty Operations Cloud has definitely been a game changer for the organization. The return on investment has been significant and highly visible. It has helped us reduce alert noise and surface resolution insight. We have saved an average of 30 minutes per incident, which totals a time savings of at least 70% to 80% per year. On a broader operational scale, the built-in ticketing management, the macros, and the automation have saved us the equivalent of at least three to four full-time employees, which has minimized our cost.

    Most importantly, it has minimized our downtime by catching signals earlier and translated that to productive revenue, which has avoided an estimated loss worth one million dollars for our company.

    What is most valuable?

    PagerDuty Operations Cloud has given us very clear AI-driven alert grouping. The autonomous AI agent calculated alerts and suppressed the noise. We used to get temporary traffic spikes or minor upgrade dips, and it has ensured that wherever there is a need for a security alert, it provides the security alert in real time, 24/7.

    PagerDuty Operations Cloud includes automated escalation, which helps us set up a call and join the right person via SMS, mobile push, or direct phone call. Another valuable feature is historical data insight, which gives clear insight during an incident and surfaces the historical data on forward-looking resolution from similar past events. This helps the team resolve issues faster.

    These features have given us real-time insight and a historical and present-time alert system via SMS, mobile push, or direct phone call. Even if the team has missed an alert at the regular time, it provides historical data and insight to ensure we understand why an incident has been raised or why a security breach has occurred.

    What needs improvement?

    PagerDuty Operations Cloud is working as expected, but if I were to suggest one improvement, I believe that the AI could be better. The accuracy can be improved because during consistent critical events, it occasionally groups unrelated alerts or misses a correlation, which forces our on-call team to fall back to manual triage.

    For how long have I used the solution?

    We have been using PagerDuty Operations Cloud in our organization for the last three and a half years.

    What do I think about the stability of the solution?

    We experienced 100% stability. We did not face any challenges.

    What do I think about the scalability of the solution?

    It is definitely a scalable product. We can easily scale up and down as per our needs. It has great scale.

    How are customer service and support?

    Customer support is a 10 out of 10. We had an interaction with them for one of the incidents, and they resolved it within 30 minutes.

    What's my experience with pricing, setup cost, and licensing?

    Our experience with the pricing, setup cost, and licensing of PagerDuty Operations Cloud has been excellent. We have chosen the pay-as-you-go licensing through AWS Marketplace, which has been a great option for us.

    Which other solutions did I evaluate?

    Before choosing PagerDuty Operations Cloud, we definitely evaluated other options and alternatives to ensure we were selecting the right solution. Some solutions we considered included Opsgenie by Atlassian, Better Stack, xMatters, and Splunk On-Call.

    What other advice do I have?

    The alert reduction feature of PagerDuty Operations Cloud has given us at least 30 minutes per incident and has reduced the overall tasks and work that was being done. We have seen a reduction in time and a reduction in the incidents, and we have saved a lot of money as well.

    The team has leveraged PagerDuty Operations Cloud to optimize the incident workflows in a positive way. It has given us clear insight into all incidents. We can easily track them in the history section. The impact is significant because it has optimized our workflows, optimized team capacity, and optimized the quality of incidents. The major impact is on time. The team has been allocated tasks on time, and they are able to deliver them on time.

    I give PagerDuty Operations Cloud a rating of nine out of ten. Since it has provided us exceptional results, I have given it a nine, and one point is deducted for the improvement I have suggested. Anyone who is looking for a dedicated platform for alerting and automation-related incidents should definitely go with PagerDuty Operations Cloud. They are the best in this category.

    Which deployment model are you using for this solution?

    Public Cloud

    If public cloud, private cloud, or hybrid cloud, which cloud provider do you use?

    Amazon Web Services (AWS)
    reviewer2879727

    Unified incident response has reduced alert noise and improves on-call focus and coordination

    Reviewed on Jul 26, 2026
    Review from a verified AWS customer

    What is our primary use case?

    I have used PagerDuty Operations Cloud for two years, so I have over two years of experience with it.

    My main duties in PagerDuty Operations Cloud are incident routing, escalation policies, on-call management, incident response coordination, and integrations.

    For incident routing, I connected AWS CloudWatch alarms to PagerDuty using the Event API. When our EC2 CPU crosses 85%, CloudWatch pushes the alert to PagerDuty, which then routes it to the backend services team. I also used event rules to auto-tag alerts and suppress duplicates. Another use case is on-call management. I used PagerDuty's mobile app to acknowledge alerts during off hours. For example, on weekends, such as Sunday morning, when there is an incident I acknowledged a high-priority alert from my phone and joined the Slack room directly from the PagerDuty incident link. This is helpful for those who are on call during off shifts. Regarding integrations, since we use many integrations on PagerDuty, we integrated PagerDuty with Slack so incidents automatically created a dedicated channel. I also connected it to Jira so incidents could generate tickets for follow-up actions, and also to ServiceNow as well.

    What is most valuable?

    PagerDuty Operations Cloud excels at routing alerts, managing on-call schedules, coordinating incidents, and automating remediation. These features make it helpful for the on-call team and the people who are working with PagerDuty Operations Cloud.

    The feature I use the most is incident routing. We use incident routing more than anything else because it is the main foundation of PagerDuty Operations Cloud. If routing is incorrect, nothing else is going to work, no escalations, no on-call, and no automation at all. I use it the most because it is the first step in every incident; it ensures that alerts from different cloud platforms or any other third-party providers such as AWS, Azure, and other platforms including Datadog reach the right service. It prevents misrouted alerts and missed incidents and it directly impacts the MTTA, because responders get the alert instantly. Every other PagerDuty Operations Cloud feature depends on routing being accurate. I use incident routing daily in this way. This is the main foundation of PagerDuty Operations Cloud, to work with it and to make it a very easy platform for us to work with.

    PagerDuty Operations Cloud excels at routing alerts, managing on-call schedules, coordinating incidents, and automating remediation. It ensures critical alerts reach the right teams instantly, reducing the MTTA and MTTR. It strengthened the on-call process with dependable escalations and real-time incident coordination, preventing missed alerts and speeding up resolution tasks. Automation and noise reduction features help cut down repetitive manual work and alert fatigue, allowing engineers to focus on real issues.

    PagerDuty Operations Cloud significantly reduced alert noise and improved MTTA. Our acknowledgement time dropped almost by half, and MTTR improved by around 15–20% due to improved automation and routing.

    The alert noise dropped by 20–25% after tuning event rules. After implementing PagerDuty Operations Cloud, we saw a 20–25% reduction in alert noise and a noticeable improvement in MTTA. Our acknowledgement time dropped by almost half and the MTTR improved by around 15–20%. MTTR improved by 15–20% due to major incident automation and better routing, which also reduced unnecessary on-call interruptions. It made a good impact on our current schedules and in my current organization.

    PagerDuty Operations Cloud's alert reduction features helped lower operational cost by cutting unnecessary alerts and reducing on-call interruptions by grouping and suppressing noisy alerts. The team spent less time triaging low-value incidents, which directly reduced overtime and burnout. The biggest impact for us was fewer false alarms, meaning engineers could focus on real issues instead of constant alert fatigue.

    What needs improvement?

    PagerDuty Operations Cloud could improve its noise reduction by making deduplication and suppression more automated instead of manually tuned, and the service dependency graph could be more intuitive with cleaner visuals and easier-to-understand root cause tracing during major incidents. Automation could go further with smarter runbook triggers and AI-driven suggestions to help find root causes, which could save a lot of time for engineers who are struggling to understand what is actually happening. These AI capabilities could lower the time by maybe 50–60%. Analytics and reporting could be more flexible, allowing custom dashboards and filters and team-level MTTA and MTTR breakdowns, so that it is segregated based on teams and it is much easier to have custom dashboards for the teams to understand more.

    Alert storms were a recurring frustration for the on-call team, and escalation overrides and service dependency graphs can get a bit confusing in larger environments. More customizable analytics and smarter automation would make the platform even easier, more flexible and more powerful for the team to understand.

    PagerDuty Operations Cloud could improve its analytics flexibility with customized dashboards and AI capabilities to be more trained and more reliable. Service dependency mapping can also feel cluttered in big environments, making it harder to trace upstream and downstream impacts during bigger incidents.

    We implemented PagerDuty Operations Cloud's AI to help with alert grouping and early incident insights, but accuracy was not consistent enough to rely on during critical events. It occasionally grouped unrelated alerts or missed correlations, which limited the operational efficiency gains we expected. The on-call team still depended heavily on manual triage because AI suggestions were not always aligned with the real root cause. Overall, AI added some value but it has not yet reached the reliability needed to significantly improve the incident response efficiency. It needs more training or more work.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is stable in day-to-day use with no major outages or reliability issues affecting our on-call workflows. Alert delivery, incident routing and escalation chains have been consistently working without delays or missed notifications. Its cloud-native architecture also means updates and patches roll out smoothly without impacting uptime. Overall, the stability has been one of the strongest features for our team.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud scales well for growing teams. Adding services, integrations and new on-call groups does not impact performance. Its native cloud architecture handles higher alert volumes smoothly, especially when paired with strong incident routing. We have seen it manage increased workloads without slowing down or causing notification delays. On the whole, scalability has been reliable even as our environment and service footprint has expanded.

    How are customer service and support?

    The customer support has been responsive and generally helpful when we have raised any issue. They provided clear troubleshooting steps and follow-ups. Resolution speed can vary depending on the severity of the issue, but support has been good and reliable for day-to-day operational needs.

    Which solution did I use previously and why did I switch?

    Before PagerDuty Operations Cloud, we used basic native alerting tools such as AWS CloudWatch since AWS is our primary cloud platform. We switched because those tools lacked valuable incident routing and consistent on-call escalation workflows. PagerDuty Operations Cloud offered stronger coordination, better noise reduction, and more mature incident response features. The move was mainly driven by the need for a unified and dependable on-call and incident management platform.

    What was our ROI?

    We have actually seen a moderate return on investment mainly through reduced alert noise and more efficient incident routing. Alert storms dropped by roughly 20–25%, saving the team several hours per week that used to be spent on low-value triage. We estimate a small cost reduction from fewer overtime hours and less on-call fatigue, though not enough to reduce headcount. Overall, the return on investment is positive but incremental — more time saved than money saved.

    What's my experience with pricing, setup cost, and licensing?

    PagerDuty Operations Cloud's pricing felt reasonable but it definitely is not the cheapest option. The value comes more from reliability than cost savings. Setup costs were minimal since it is a SaaS platform and onboarding did not require any infrastructure or hidden implementation costs or fees. Licensing is straightforward but scaling seats for larger teams can get expensive, especially when adding advanced features. Overall the experience was smooth but the price point could be more flexible for growing teams. Our team is small now but in the future it might become bigger.

    Which other solutions did I evaluate?

    Before adopting PagerDuty Operations Cloud, we evaluated options such as Opsgenie and Splunk On-Call for incident management. We also looked at other native cloud tools such as AWS CloudWatch, but they lacked the escalation workflows. Opsgenie had good features but did not match PagerDuty Operations Cloud's reliability and integrations for bigger teams. Overall, PagerDuty Operations Cloud offered stronger coordination and more consistent on-call performance which made us choose it.

    What other advice do I have?

    Escalation policies are a very helpful feature for managing the team, such as multi-step escalation logic which is extremely dependable and prevents missed alerts, and also the automation runbook. I rely on incident routing daily because it ensures the alerts reach the right team. The other features, such as integrations, are really good to integrate PagerDuty Operations Cloud with different platforms, such as Jira and ServiceNow platforms to automatically open an incident. These are really helpful top features.

    I would give PagerDuty Operations Cloud eight out of ten, considering all the features I have been using and also the improvements I have suggested. I chose eight out of ten because PagerDuty Operations Cloud is generally strong in the areas that matter the most, especially incident routing and escalation management which are extremely reliable and directly reduce MTTA and MTTR. It stands out for dependable on-call orchestration and smooth incident response coordination, especially with Slack and Jira integrations. But it does not reach a ten because noise reduction still needs more automation, analytics lack deeper customization such as custom dashboards to filter among the teams, and the service dependency graph gets cluttered in large environments. With smarter automation and improved visualization of the analytics, it could easily move closer to a perfect ten.

    Which deployment model are you using for this solution?

    Public Cloud

    If public cloud, private cloud, or hybrid cloud, which cloud provider do you use?

    Amazon Web Services (AWS)
    Mukul Bhati

    Automation has reduced manual support effort and has improved incident response accuracy

    Reviewed on Jul 16, 2026
    Review from a verified AWS customer

    What is our primary use case?

    My main use case for PagerDuty Operations Cloud is for doing the automation part and minimizing the efforts. It decreases the count of the L1 support team because most of the tasks have been handled by the operation of PagerDuty Operations Cloud and the AI.

    What is most valuable?

    PagerDuty Operations Cloud offers several valuable features that have significantly benefited our operations. The solution keeps the person available and only rings the alert while there is any outage or any issue generated in the live environment. Additionally, it has an automated section where it handles similar issues if they happen again and is able to troubleshoot them by itself by running those jobs or actions that have been put in PagerDuty Operations Cloud.

    PagerDuty Operations Cloud positively impacts our organization by helping us increase the overall SLA delivery. Earlier, we were having a delivery count of 87% or 88%, but after setting up PagerDuty Operations Cloud, we were able to achieve the SLA up to 99.5% post-installation. The improvement in SLA and SLI is mostly because of the faster incident response and fixing the issues that are most generic and reoccurring, which has been minimized and helped us increase the SLA and SLI.

    What needs improvement?

    PagerDuty Operations Cloud is already at its best, but AI can be integrated just to set it up and continue learning and provide a list of fixes that can be applied as suggestions, as that might help improve PagerDuty Operations Cloud operations.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for around three years or more.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is stable and much more stable than previous solutions.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud's scalability is much more stable than expected. We can scale it up whatever the requirements are, so scalability is good.

    How are customer service and support?

    Regarding customer support, we never found any reason to get support from the customer. We were already supported while setting up the infrastructure and it was good. I would rate the customer support on a scale of one to ten as ten out of ten.

    Which solution did I use previously and why did I switch?

    Previously, we were using Teams for generating the alerts and Slack, but we prefer PagerDuty Operations Cloud, which offers many more options and setups that can be used.

    How was the initial setup?

    To deploy PagerDuty Operations Cloud in our organization, we have used public and private cloud. Earlier it was on-premises but we switched to private cloud.

    What about the implementation team?

    We have implemented the AI and automation through PagerDuty Operations Cloud for incident response. As mentioned earlier, it increased the SLA from 87% to 99% and our operations have been improved, and there is a very low count of clear or false alerts.

    What was our ROI?

    We have seen a return on investment. We have cost-cutting on the employees as we have decreased the headcount since there is less man labor required while PagerDuty Operations Cloud is able to handle all the alerts first as an L1. There is a lot of money saved compared to the pricing or the license that we have spent.

    What's my experience with pricing, setup cost, and licensing?

    My experience with pricing, setup cost, and licensing for PagerDuty Operations Cloud is that pricing and license are good to go. Everything is perfect as per the licensing and pricing setup cost. There is no more suggestion I can provide based on that. We completely achieved whatever we are spending.

    Which other solutions did I evaluate?

    Before choosing PagerDuty Operations Cloud, we evaluated other options and chose Teams.

    What other advice do I have?

    We are using PagerDuty Operations Cloud to our maximum capacity and we appreciate the support that it has for continuing to learn from the issues that we have and to predict the outages or any alerts or whatever escalation is required. It can fulfill that.

    PagerDuty Operations Cloud's embedded AI has significantly influenced revenue protection in terms of reducing alert fatigue and incident costs. We have improved a lot, and we have improved 25 to 27% of our overall incident management.

    The alert reduction feature of PagerDuty Operations Cloud has a significant impact on preventing costly incidents in our organization. Alert reduction has been decreased as a result of removing false alerts. Earlier, we were getting 30 to 35 alerts per day. Now, there are only four or five general alerts that are genuine. That is a significant amount of alert reduction.

    For others looking into using PagerDuty Operations Cloud, I would completely suggest that PagerDuty Operations Cloud is a good option to install or set up in infrastructure. It handles the infrastructure very well. I would rate this review nine out of ten overall.

    Abhishek Jadli

    Reliable incident paging has kept outages under control but separate client notes are still missing

    Reviewed on Jun 23, 2026
    Review from a verified AWS customer

    What is our primary use case?

    My usual use cases with PagerDuty Operations Cloud involve handling incidents through a full flow. When there is an outage, an incident is created that can be either severity one or severity two. As the person on call that day, I receive a page from PagerDuty on the app and three calls on my cell phone. When I pick up the call, PagerDuty IVR asks me to acknowledge the incident. Once I acknowledge the incident from the call, I go to PagerDuty through the website, which is much easier to navigate than the mobile app. I then page other teams responsible for the incident, as well as the stakeholders and product owners.

    PagerDuty is integrated with Microsoft Teams, so I open a Teams bridge call to resolve the issue and update all incident details in PagerDuty notes, which automatically integrates with ServiceNow incident and sends messages to stakeholders' phone numbers.

    One time while hanging out in the mountains, there was no internet signal but there was cell reception. An incident happened while I was on call that day. Normally, without internet, I would not be able to know about it, but because of PagerDuty, I was paged three times on my cell phone as well as through text message. I managed to call another colleague from my work and told him to take care of the incident. This helped me avoid breaching the SLAs on incident acknowledgment and allowed me to access remote incidents without relying solely on the internet.

    What is most valuable?

    The features of PagerDuty Operations Cloud that I find most valuable involve being automatically paged whenever an incident is triggered. PagerDuty has group names embedded into it, and when we set up PagerDuty in our organization, we embedded the group name, allowing me to page other respondents without having to go separately into Microsoft Teams, add everybody's name, and then ping and call them. I can do this directly from PagerDuty itself.

    The notes update feature allows me to put the details of the incident in the notes and click post, and it is integrated everywhere. Everything is centralized.

    PagerDuty Operations Cloud has improved my team's ability to focus on core tasks rather than routine issues primarily due to its availability and reliability. We do not worry about whether PagerDuty will call us when an incident triggers, allowing us to focus on that incident. The notes update process ensures that everyone gets informed with just one update.

    What needs improvement?

    I think PagerDuty Operations Cloud could be improved by having two fields for incident updates. In my work, I handle incidents that have two fields: work notes visible only to developers working on the incident and additional comments visible to the client. When I update through PagerDuty, everything gets updated into the additional comments. There should be two fields, perhaps based on how it is integrated with ServiceNow.

    I have not used PagerDuty's autonomous AI agents or generative AI, so I am not sure whether it is integrated.

    I can evaluate the effectiveness of PagerDuty Operations Cloud in providing insights for decision-making, and I rate it around seven out of ten. PagerDuty Operations Cloud is really helpful. With generative AI integrated and a chatbot, I think it would bump the rating up to nine, but I have not used it yet, so I cannot say for certain.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for around one and a half years.

    What do I think about the stability of the solution?

    Regarding the reliability and stability of this product, I find it really stable. Based on my personal experience in the mountains, I can vouch for it. My organization measures the acknowledgment rate by checking how many times a call rings and how quickly the responder acknowledges the incident. The acknowledgment rate is really good, around ninety to ninety-five percent.

    What do I think about the scalability of the solution?

    I think the scalability of PagerDuty Operations Cloud is superb. It can handle many requests according to demand. As other members and teams are added to our organization, it has not impacted the latency, call failure, or anything else. The performance is really good.

    How are customer service and support?

    I do not often communicate with the technical support of PagerDuty. The manuals are created by our team, so we use those.

    Which solution did I use previously and why did I switch?

    I did not use a different solution for the same use case before PagerDuty Operations Cloud. PagerDuty Operations Cloud was the first solution I used.

    How was the initial setup?

    I did not participate in the initial setup of PagerDuty Operations Cloud. When I joined the organization after one and a half years of use, it was already set up for me by the IT team.

    What about the implementation team?

    I have not personally implemented automation through PagerDuty for incident response, but I think some of my colleagues may have done so.

    What was our ROI?

    I am not aware of the impact of PagerDuty's alert reduction feature on preventing costly incidents in our organization, as that metric is not shared with me.

    What's my experience with pricing, setup cost, and licensing?

    I am not sure about the influence of PagerDuty Operations Cloud on revenue protection in terms of reducing alert fatigue and incident costs, as these are organizational-level decisions, and employees are not involved in those discussions.

    Which other solutions did I evaluate?

    Before PagerDuty Operations Cloud was chosen, I did not evaluate other options, as the decision was made by the board members.

    What other advice do I have?

    My review rating for PagerDuty Operations Cloud is seven point five out of ten.

    Which deployment model are you using for this solution?

    Public Cloud

    If public cloud, private cloud, or hybrid cloud, which cloud provider do you use?

    Amazon Web Services (AWS)