PagerDuty Operations Cloud logo

    PagerDuty Operations Cloud

    Sold by
    The PagerDuty Operations Cloud is essential infrastructure for all unplanned, time-sensitive, critical work. It automatically detects and diagnoses disruptive events mobilizes the right team members to respond and automate infrastructure and workflows across your digital operations. This means you can resolve unplanned, unstructured, time-sensitive, and high-impact issues quickly - with fewer escalations to your technical teams while minimizing the impact on your customers and maintaining brand trust.

    Ratings and reviews

    4.5
    1038 ratings
    42 AWS reviews
    |
    996 external reviews
    External reviews are from G2  and PeerSpot .

    Filters

    Review type

    AWS Marketplace reviews
    External reviews
    Reviews (1038)
    Yuvashree Yuvashree

    Intelligent incident automation has improved on-call routing and reduced downtime risks

    Reviewed on Sep 21, 2026
    Review provided by PeerSpot

    What is our primary use case?

    My main use case for PagerDuty Operations Cloud is integrating with ITOM, incident management, and also using PagerDuty for automation.

    A quick specific example of how I use PagerDuty Operations Cloud for incident management is that I commonly use it for critical server down alerts for certain application performance issues, database capacity alerts, network device failures, etc.

    In addition to that, for a network device failure when a core switch or a router goes offline, SolarWinds detects the node status is down alert. So whenever an alert is detected, PagerDuty Operations Cloud opens a P1 incident and relevant teams like the NOC team, network team, and incident management team get notified. They automatically get paged through PagerDuty Operations Cloud, and if the incident is not acknowledged, it gets escalated to the managers. This is the automation use case which I created for the network device failure alerts.

    The feature I find myself using the most is the best use case from Datadog to PagerDuty Operations Cloud to the Linux team, which I have created. Datadog detects CPU memory or disk or service failure alerts for the Linux server, and PagerDuty Operations Cloud automatically creates the incident and notifies the Linux on-call engineer.

    What is most valuable?

    According to PagerDuty Operations Cloud, the best features offered are intelligent incident management, AI Ops and alert noise reduction, automation and runbook execution, cloud and hybrid infrastructure monitoring, ChatOps integration, AI-powered operations, service ownerships, and business visibility.

    PagerDuty Operations Cloud has positively impacted my organization by reducing mean time to resolution, automatically routing incidents to the correct on-call engineers, reducing alert fatigue, whereas PagerDuty Operations Cloud AI Ops correlates duplicate and related alerts. It also provides 24/7 operational coverage, faster incident management, increased automation, better visibility and accountability, and improved knowledge management.

    What needs improvement?

    PagerDuty Operations Cloud can be improved by eliminating alert storms, automating common incident resolutions, improving major incident management, using AI for faster troubleshooting, and improving operational metrics.

    I could add a feature to automate common incident resolution where engineers perform repetitive actions like restarting servers, clearing disk space, collecting logs, or running diagnostics.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for almost six plus years.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is stable.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is highly scalable and is desired for organizations ranging from small operation teams to larger enterprises managing thousands of services, responders, and alerts.

    How are customer service and support?

    PagerDuty's customer support is generally considered strong from enterprise customers, particularly those running mission-critical operations and requiring 24/7 incident management.

    Which solution did I use previously and why did I switch?

    Before adopting PagerDuty Operations Cloud, I primarily relied on monitoring platforms such as Datadog and SolarWinds and native cloud monitoring solutions for alert generation.

    What was our ROI?

    Organizations commonly see ROI from PagerDuty Operations Cloud in both time-saving and downtime reduction, especially when integrated with monitoring platforms and automation workflows.

    What's my experience with pricing, setup cost, and licensing?

    From my experience evaluating and implementing PagerDuty Operations Cloud for incident management and on-call operation, PagerDuty Operations Cloud offers tiered plans that cost varies based on features such as incident management, AI Ops, event orchestration, automation, and AI capabilities.

    Which other solutions did I evaluate?

    During the evaluation phase, I considered other incident management platforms such as ServiceNow, Opsgenie, etc., before selecting PagerDuty Operations Cloud.

    What other advice do I have?

    PagerDuty Operations Cloud helps my team focus on core tasks, which means it helps engineers spend less time on operations overhead and more time on engineering work.

    My team has leveraged automation within PagerDuty Operations Cloud by integrating it with monitoring platforms such as Datadog, SolarWinds, and cloud monitoring tools to automate the end-to-end incident management process. Alerts are automatically correlated, routed through the appropriate on-call engineers, escalated when required, and tracked through resolution. The team also leverages automated runbooks and AI-assisted triage to reduce manual intervention during incidents.

    PagerDuty Operations Cloud's alert reduction feature has significantly reduced the risk of costly outages in my organization by ensuring critical issues are identified, routed, and escalated to the appropriate responders in real-time. Automation, notifications, intelligent alert grouping, and escalation policies help prevent incidents from being overlooked, reducing downtime and business impact.

    I use workflows to standardize the incident response process by automatically engaging the correct support teams, creating collaboration channels, notifying stakeholders, and tracking incidents through resolution. I have also utilized event orchestration to filter, enrich, suppress, and route alerts from monitoring tools such as Datadog and SolarWinds.

    Organizations get the highest value from PagerDuty Operations Cloud when they use it for full operation platforms combining incident management, event orchestration, AI Ops, automation, governance, and service ownership.

    Regarding governance and security, PagerDuty Operations Cloud provides strong controls that help the cloud operations team maintain compliance, accountability, and operational resilience.

    For cloud operations, AI is valuable only if it is accurate, reliable, and actionable. PagerDuty Operations Cloud improves this through an AI-first operation platform, which is built on large-scale operational data, event correlation, and incident response workflows.

    I rate this product a 10 out of 10.

    Alexey Dubrovin

    Centralized incident hub has improved real-time visibility and accelerated team response

    Reviewed on Sep 21, 2026
    Review from a verified AWS customer

    What is our primary use case?

    My main use case for PagerDuty Operations Cloud is to use it as a main hub to see incidents, to get proper responses to incidents, and to log resolutions and post-mortems.

    A recent specific example of how I use PagerDuty Operations Cloud for incidents or post-mortems involves wanting global visibility of an incident to react. When something happens, such as an external provider of the data that we integrate with starts giving us an error on our call, we immediately see it on the PagerDuty dashboard and in our Slack channel alerts, and somebody from the Ops teams is also notified immediately on their phone. We then assign a responsible person, which can be someone from Ops, dev, or product, depending on the system that fails, and whoever is responsible starts resolving the incident. If it is critical, we react immediately; if it is not critical for revenue-generating streams, it may wait and be prioritized in the following sprints.

    Everything is logged in PagerDuty, allowing all interested parties, such as product owners and top management, to track progress.

    What is most valuable?

    PagerDuty Operations Cloud integrates easily into our communication tools, such as Slack, which is super convenient. I get notified there and on my phone, allowing me to quickly click on the link, get to the dashboard, and see what is going on.

    Instant notifications help reduce response times, as I do not need to always open the PagerDuty Operations Cloud portal. I see the incident in Slack and can quickly assess if I need to react. When I do, it is super easy to get the information I need.

    PagerDuty Operations Cloud automates notifications about incidents, preventing overwhelm with different kinds of incidents.

    PagerDuty Operations Cloud has impacted our organization positively by enabling faster reactions to incidents and providing greater visibility, as all parties that need to be notified are notified automatically, creating high visibility.

    What needs improvement?

    While configuring PagerDuty Operations Cloud with Ops, it was not super intuitive to figure out which fields to use to assign a responsible person and who is writing the post-mortem, but once everything was figured out, it works really well.

    PagerDuty Operations Cloud can be improved by continuing to focus on user experiences. The initial setup can be streamlined better, as it feels overwhelming with many configurations and integrations, though once we got the hang of it, it became much easier to understand, and now the user experience for people such as me is absolutely great.

    Integrations with external tools such as Tableau could be useful, as I have not explored that yet myself.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud at CARFAX Europe for maybe a year and a half to two years.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is quite stable.

    What do I think about the scalability of the solution?

    I cannot judge PagerDuty Operations Cloud's scalability, but we are not growing rapidly. Our growth is pretty organic, with user count and services increasing by maybe 20 percent a year, and so far I have no complaints.

    Which solution did I use previously and why did I switch?

    We previously used a different solution, a pretty lightweight status page that was static, requiring us to visit a website for a list of services and their statuses, and manually check which service had issues, but I do not remember the name.

    How was the initial setup?

    Take your time to set it up properly.

    What about the implementation team?

    Since I was not involved in the configuration, I cannot really tell if the documentation is good or not, so I do not have any further comments.

    What was our ROI?

    Although I do not think we measured return on investment, we definitely saved time, and time is money, so I guess we saved some amount of money for sure.

    What's my experience with pricing, setup cost, and licensing?

    I was not dealing with pricing, setup cost, or licensing, as that was the responsibility of the Operations team in our company, so I cannot comment.

    Which other solutions did I evaluate?

    I think we evaluated other options before choosing PagerDuty Operations Cloud, but I was not a part of the evaluation.

    If public cloud, private cloud, or hybrid cloud, which cloud provider do you use?

    Amazon Web Services (AWS)
    MayankGarg

    Centralized incident routing has streamlined on-call response and reduced downtime

    Reviewed on Sep 18, 2026
    Review provided by PeerSpot

    What is our primary use case?

    I primarily use PagerDuty Operations Cloud for managing incidents and monitoring all critical services on our application. I rely on it most for real-time alerts, automated incident routing, on-call coordination, and quickly bringing the right team members into an issue when something goes wrong with our production applications.

    A recent issue occurred when one of our production services started showing a spike in errors. PagerDuty Operations Cloud triggered an alert and automatically routed the incident to our on-call engineer. We were able to acknowledge and quickly coordinate the response while keeping track of updates in one place, making it easy to identify issues and avoid delays in determining who should handle the problem and how to reduce the time to restore service to normal so we could maintain the same user experience.

    PagerDuty Operations Cloud fits well into my daily workflow because it provides one place to manage all alerts, incidents, and on-call responsibilities. I especially appreciate having clear ownership and escalation when an issue needs immediate attention. It reduces manual coordination and makes incident response more organized and structured.

    What is most valuable?

    Several features stand out for me, including automated alerting, intelligent incident routing, on-call scheduling, and escalation policies. I find it especially useful that alerts reach the right person automatically instead of relying on manual communication. The incident timeline and collaboration features make it easier to understand what happened and keep everyone aligned during an outage.

    Escalation policies make the biggest difference for my team because they ensure that if the first person does not respond, the incident automatically moves to the next person instead of getting missed. This provides us more confidence during critical issues, especially outside normal working hours.

    PagerDuty Operations Cloud has made our incident response more organized, well-structured, and consistent. Instead of relying on manual communications, alerts are automatically routed to the appropriate team and escalation ensures issues do not get overlooked. It has also improved collaboration during an incident because everyone can see the status and ownership clearly. Overall, my team spends less time coordinating and more time actually resolving problems.

    What needs improvement?

    One area I would like to see improved is making complex alerts configuration and routing rules easier to set up and manage.

    PagerDuty Operations Cloud's flexibility is useful, but some advanced configuration can take time to understand. Simplifying the setup of routing rules and escalation policies with clear guidance for complex workflows would make the platform easier for new team members to adopt.

    The platform feels a little complex when setting up advanced routing rules, escalation policies, and integrations. For new users, it takes time to understand how everything fits together. I would like to see simpler configuration workflows and clear guidance, particularly for new users or new team members. Once everything is configured, the day-to-day incident management experience is fairly smooth.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for two years.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is very stable.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is very efficient, robust, and scalable.

    How are customer service and support?

    Customer support is wonderful and always helpful.

    Which solution did I use previously and why did I switch?

    We are using this for monitoring all our production applications.

    What was our ROI?

    We see a good return on investment. It saves our time as well as money.

    What's my experience with pricing, setup cost, and licensing?

    Pricing is good and not too cheap.

    What other advice do I have?

    I would always recommend PagerDuty Operations Cloud to any organization because it monitors all production applications, enhances user experience, and reduces application downtime.

    I found the AI output to be generally accurate and reliable for common incident management tasks, especially when summarizing incidents and highlighting relevant information. It gives us a useful starting point and can save time when reviewing an issue. I still treat AI-generated recommendations as a support tool and verify important details before taking action, particularly during critical incidents.

    I give this product a rating of 9 out of 10.

    Kishan Chand

    Incident automation has reduced alert noise and now improves response times for critical issues

    Reviewed on Sep 17, 2026
    Review from a verified AWS customer

    What is our primary use case?

    I use PagerDuty Operations Cloud as my primary incident management system and automated escalation platform because it integrates directly with my monitoring tools like CloudWatch and Datahog to reduce aggregate alerts and alert noise through duplication. It centralizes real-time alert triage and automatic incident response across my infrastructure and microservices. It automatically logs issues and transfers them to my IT team to resolve as soon as possible. Most significantly, it is lowering my mean time to acknowledge (MTTA) and mean time to resolve (MTTR).

    In a recent situation, I had a production service where response time suddenly increased. PagerDuty Operations Cloud detected the issue and triggered an alert to the right engineer over a call. I saw the incident details and escalation workflow to quickly identify the affected service and bring the right engineer team to troubleshoot the issue so that user experience would not be hampered. The issue was resolved before it had a major customer impact. I found it particularly useful that PagerDuty reduced the time between detecting the problem, notifying the right person, and getting it resolved as soon as possible.

    What is most valuable?

    I see many valuable features, but the most useful are incident alerting, incident escalation, on-call scheduling, integration with my monitoring tools, and better collaboration. There is a centralized dashboard where all incidents are logged, where I can track the responses and see who is handling the issue. The mobile application is also very useful because it automatically alerts me when I am away from my desk so that I can fix and resolve issues from anywhere.

    The mobile application is particularly helpful when I am away from my laptop. For example, when a critical incident occurred outside normal working hours, PagerDuty notified me on my phone and I acknowledged the alert and checked the incident details right away. I was able to share the details with my IT team without waiting for someone to return to their desk. This ability to stay updated and respond from the mobile app is especially useful for on-call situations.

    PagerDuty Operations Cloud is creating a positive impact on my enterprise by making incident response more structured and organized. Alerts are automatically routed to the right engineer and escalation policies ensure that important incidents are not missed. It has overall improved communication during incidents because everyone can see the status and take ownership. It has helped me reduce delays in responding to operational issues and has improved coordination across all teams.

    What needs improvement?

    I feel the user interface is not optimal for new users and needs some optimization because some configuration takes time to understand, especially when adding rules and workflows. The platform is otherwise good and the experience is wonderful and easier for my team.

    The setup could be simpler, especially for new users who want to add integrations. Better guidance during the initial configuration would make it easier to get started. Some improvements could be made to the administrator features and to make the platform easier to manage as the team grows or increases.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for the last two years.

    What do I think about the stability of the solution?

    I have experienced stability issues.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is very scalable.

    How are customer service and support?

    Customer support is excellent.

    How was the initial setup?

    Pricing is reasonable and good for my organization. It is very easy to set up.

    What was our ROI?

    I see a good ROI of approximately 50 to 60 percent because it makes my application more robust and reduces the need for my IT team to do manual work and troubleshoot issues.

    What other advice do I have?

    AI capabilities are always useful in terms of governance and security, and they give me clear control around all data access, permissions, and auditability. AI gives me incident information quickly. In my experience, I am using AI for recommendations on troubleshooting issues. The AI feature helps me handle issues much quicker and faster.

    AI is generally useful for getting better incident response and identifying the next steps I can take to resolve issues. Recommendations are always helpful when I am spending time troubleshooting issues and monitoring incident details. The value is mainly in reducing the time spent analyzing the issue while keeping human review as a part of the process.

    Alert reduction features have helped me cut down duplicate or low priority notifications so that my team can focus on alerts that actually need attention and have high priorities. It reduces unnecessary escalation so that I can spend less time investigating issues and respond to genuine incidents sooner.

    PagerDuty Operations Cloud helps me reduce the time spent on incident triage and follow-ups. Alerts are easier to prioritize so that my team can identify the right responder faster. I have seen less time spent manually coordinating incidents and keeping track of routine alerts. It also gives me more flexibility so that I can customize reporting and create my own dashboards and workflows to easily detect and resolve issues as soon as possible.

    I have heavily utilized event orchestration and automatic incident workflows to streamline my process. By setting up rules that automatically suppress alerts for non-critical issues, I can resolve genuine high-severity issues more effectively. I integrated automated diagnostic runbooks that execute immediately upon incident creation, gathering essential logs and service health metrics before an engineer even logs in. The automation eliminated repetitive manual overhead during outages and removed significant alert fatigue across my team, allowing them to redirect their focus toward planning project work and system resilience. This enables my engineers to work simultaneously to resolve issues more quickly and efficiently.

    I use PagerDuty SRE the most because it helps automate incident response, identify issues quickly, and reduce the time needed to resolve incidents. It makes on-call work more efficient and helps my team focus on high-priority tasks.

    I always suggest using PagerDuty Operations Cloud to monitor all production applications. I rate this solution a 10 out of 10.

    Harjoth Sudan

    Structured alerting has reduced recurring incidents and enables faster resolution for our team

    Reviewed on Sep 04, 2026
    Review provided by PeerSpot

    What is our primary use case?

    My main use case for PagerDuty Operations Cloud is to manage alerts and to monitor AI-powered workflows. For example, whenever a service goes down, it needs attention, and PagerDuty Operations Cloud helps in keeping track of all these things.

    PagerDuty Operations Cloud helps me when a service goes down by notifying everyone. Prior to this, we were using Slack, but we were relying on messages, and this platform gives more clear visibility to the team rather than a message.

    PagerDuty Operations Cloud identifies the right person and notifies that person, so it is way faster and efficient.

    What is most valuable?

    The best features PagerDuty Operations Cloud offers for us include incident management and automated alerting, which are helping us. These are the only reasons we started and onboarded this.

    Automated alerting and incident management features make my team's life easier on a day-to-day basis. We have set the hierarchy, and if the first person does not respond, it moves to the second person, so that way we do not miss any escalations or issues.

    PagerDuty Operations Cloud has positively impacted my organization by making us more structured. All the production-related issues which we were facing before are sorted.

    Since using PagerDuty Operations Cloud, I have noticed specific outcomes such as faster incident resolution. Recurring issues were monitored properly, and recurring incidents were corrected and have stopped.

    What needs improvement?

    PagerDuty Operations Cloud can be improved. The product itself is a bit complicated, and the initial configuration takes time, making it very difficult for new members to understand.

    The product is strong, but making the setup simpler and easy to operate is an area that can be improved.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for the last year.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is stable.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is highly scalable.

    How are customer service and support?

    Customer support is good. I would rate customer support nine out of ten.

    Which solution did I use previously and why did I switch?

    We have not used any different solution previously.

    What was our ROI?

    I have seen a return on investment. The ROI is great for us as repetitive incidents have decreased, the time in solving those issues has also decreased, and the team needs more projects and fewer team members.

    What's my experience with pricing, setup cost, and licensing?

    My experience with pricing, setup cost, and licensing revealed that licensing is fine. However, the pricing was a bit high, and the setup, as I mentioned, is a bit complicated, costing us four weeks and roughly one hundred dollars.

    Which other solutions did I evaluate?

    Before choosing PagerDuty Operations Cloud, one of the team heads suggested this to us, and we implemented it because they were using this in their previous organization.

    What other advice do I have?

    Regarding PagerDuty Operations Cloud's AI capabilities, I think its governance and security are trying to cope up. For now, this has not been a challenge for us, but we keep an eye always.

    Regarding PagerDuty Operations Cloud's AI capabilities, I find its accuracy and reliability of output to be very accurate and quite reliable.

    My advice to others looking into using PagerDuty Operations Cloud is that it helps a lot with production issues and solving them on time. I suggest they go ahead and set this up properly, as initially it takes time but will provide ROI.

    PagerDuty Operations Cloud is one of the reliable solutions that help in solving production issues. I have rated this review eight out of ten.

    adithya k.

    Streamlined Alert Management with Seamless Integrations

    Reviewed on Sep 01, 2026
    Review provided by G2
    What do you like best about the product?
    I like the flexible scheduling, rotations, and override rules in PagerDuty that ensure the right person is always notified. I appreciate how it integrates with numerous monitoring tools to suppress noise and group related events. The ChatOps feature is a highlight for me, and Slack enables teams to acknowledge, reassign, and resolve issues via chat commands. Additionally, the deep integration with various tools makes the experience seamless. The initial setup for PagerDuty is generally straightforward, taking about 15 to 30 minutes for basic configuration.
    What do you dislike about the product?
    All good
    What problems is the product solving and how is that benefiting you?
    I use PagerDuty for flexible scheduling and to integrate with monitoring tools, suppressing noise and grouping events. ChatOps and deep integration let our team acknowledge incidents and resolve issues via chat.
    Gowhar S.

    Reliable Incident Management and Alert Response

    Reviewed on Sep 01, 2026
    Review provided by G2
    What do you like best about the product?
    What I like most about PagerDuty is how it turns alerts into a clear, structured incident response process. Real-time notifications, flexible on-call scheduling, and well-defined escalation policies make it easier to ensure critical issues reach the right person quickly, without unnecessary delays. I also appreciate the integrations with monitoring and collaboration tools, since they let teams manage incidents in one place instead of constantly switching between platforms. Overall, PagerDuty improves visibility and accountability, and it helps teams respond faster when incidents happen.
    What do you dislike about the product?
    PagerDuty can feel a bit complex to configure, especially for new users. If the initial setup isn’t done thoughtfully, managing multiple alerts and notifications can quickly become overwhelming. Pricing may also be a concern for smaller teams.
    What problems is the product solving and how is that benefiting you?
    PagerDuty helps us manage critical alerts by routing incidents to the right teams and reducing our response times. It also improves overall visibility, supports timely escalation, and makes it less likely that important issues will be missed.
    Aaron Held

    Automated alerting has streamlined incident response and improved team coordination

    Reviewed on Aug 26, 2026
    Review provided by PeerSpot

    What is our primary use case?

    The usual use cases for PagerDuty Operations Cloud include incident management. I ran the infrastructure teams and SRE, so the primary use case was incident management and spinning up RCAs, tracking remediation. We also used it for identifying key stakeholders of projects. We relied on the escalation chains and the groups. Even if it wasn't an incident, if somebody had a question or they needed to get a hold of the CMS team, they would go into PagerDuty. We had a lot of integrations into Teams chat.

    The heart of it was obviously kicking off pages from alerts within PagerDuty Operations Cloud. It was fully automated so that when an anomaly was detected, it automatically created an incident in PagerDuty. That automation was just table stakes. We also had some good automations to follow up and close out Jira tickets that would integrate with the PagerDuty incident. The primary one was simply calling the people when something happens.

    What is most valuable?

    My favorite feature about PagerDuty Operations Cloud is that it just worked. It was one of the tools we didn't have to worry or think about. It was very well polished. It did what they said they were going to do, and it didn't require a lot of customization. I think it facilitated us organizing the teams, groups, and escalation chains really well. However, there is no feature that stands out significantly.

    PagerDuty Operations Cloud absolutely requires maintenance on my end. Escalation chains and people get stale very quickly. I wish it would integrate into org charts or an HR system so that when people get terminated or shift roles, PagerDuty knows about the changes. We had one escalation chain where it went up to an executive that wasn't in the company anymore.

    What needs improvement?

    The users would complain that they always had a hard time maintaining their schedules, their escalations, and installing the app. I got a lot of complaints from the people that responded to pages rather than from the people who administered PagerDuty.

    The complaints I remember were about how users wanted to be notified. It was very common for somebody to misconfigure their own communication chain and not test it, so when an alert went off, their excuse was they didn't get the page or they didn't get the SMS. I got a lot of feedback from users that they struggled to do their shift when they needed someone to cover their shift. Often that wouldn't get done. Anything that can be done to make that as easy and simple as possible would be helpful.

    For how long have I used the solution?

    I have probably used PagerDuty Operations Cloud for about five years total.

    What do I think about the stability of the solution?

    The stability of PagerDuty Operations Cloud has been great. I have not experienced any stability issues.

    What do I think about the scalability of the solution?

    I have no problems with scalability. I have used it for like 500 engineers and experienced no issues.

    How are customer service and support?

    I did not have any partnerships with them. I was just a customer and user.

    Which solution did I use previously and why did I switch?

    I have used some alternatives. I used Grafana incident management a lot and we used something called RingCentral.

    How was the initial setup?

    The initial deployment for a new client when starting with PagerDuty Operations Cloud was moderate. I would say it was more difficult because the biggest problem we had was just identifying who the responders were. It does take a bit of setup. We also did the Teams integration, which requires security and permission. The setup is always hard, though this is not necessarily a criticism on PagerDuty.

    What's my experience with pricing, setup cost, and licensing?

    I am not current on what the current pricing is for PagerDuty Operations Cloud, but I assume it is per seat and per responder.

    Which other solutions did I evaluate?

    If I were to pick one, PagerDuty Operations Cloud is the one I would pick if there were no budget concerns. If I was budget-constrained, I would pick Grafana.

    What other advice do I have?

    PagerDuty Operations Cloud did influence the volume and nature of alerts. We had more of them. I have implemented this three major times in my career, and the number of incidents immediately shot up and then came down. What happens is you identify the incidents more aggressively and then you actually fix the problems. I think that is a good thing.

    I have not used Incident Workflows or Event Orchestration to automate toil from the incident management process. We played with it, but I did not use them in practice.

    I did not use the AI agents at all. I would rate this review a 9.

    Aryan Dwivedi

    Real-time alerts have protected client sites and now keep our team focused on critical issues

    Reviewed on Aug 25, 2026
    Review provided by PeerSpot

    What is our primary use case?

    PagerDuty Operations Cloud is used for real-time alerts when something goes wrong with our system so the team can fix it right away, and for setting up on-call schedules to ensure the right person is always ready to answer the problem at any time of the day. We also use it for scaling up discovery calls with our clients and prospects.

    Last month, one of our main web servers stopped working in the middle of the day, and the system sent an instant notification to the team, who automatically used our on-call schedule to contact the right engineer on duty, saving us precious time instead of guessing who to call.

    We have used PagerDuty Operations Cloud to control the external noise from the alerts.

    What is most valuable?

    The real-time notification is top-notch, and critical warnings reach the right person right away without delay. Smart tools that reduce extra alert noise are great for keeping the team focused on real-time problems. The large number of connections to over 700 different software tools allows us to bring all work together in one workflow.

    PagerDuty Operations Cloud has helped fix critical problems much faster, including the ads we are running that do not crash, the websites we run for our clients, and our CRM website.

    I cannot say that PagerDuty Operations Cloud has reduced downtime or improved our team's response time because we have been using it for around two months only. However, we have seen reduced downtime and quick resolution of issues because everyone is notified when something is damaged, whereas earlier we would only find out about issues when the client called us. This also gave us a demerit because the client thought we could not do the work properly, resulting in a significant reduction in our monthly retainer. Now we are stable with our monthly retainer.

    The security governance of PagerDuty Operations Cloud is pretty good.

    Regarding its accuracy and reliability of output, I would say it is pretty much accurate.

    PagerDuty's generative AI is pretty good, although sometimes the decision can be a bit generic and not very reliable, but they are improving it.

    What needs improvement?

    The learning curve for PagerDuty Operations Cloud is a bit steep, with its complicated rules and schedules, and we have to dedicate a proper team to use that tool.

    For how long have I used the solution?

    I have been using PagerDuty Operations Cloud for two months.

    What do I think about the stability of the solution?

    PagerDuty Operations Cloud is pretty much stable and I have not seen it crashing.

    What do I think about the scalability of the solution?

    PagerDuty Operations Cloud is pretty much scalable.

    How are customer service and support?

    I have not used customer support until now, but I believe they will be great.

    Which solution did I use previously and why did I switch?

    We were not using any different solution; this is the first time we are going for an on-cloud platform.

    What about the implementation team?

    I cannot share how my team has leveraged automation within PagerDuty Operations Cloud because it is a bit confidential.

    What was our ROI?

    It helps us save manual hours of work every week.

    For me as an employee, I would say the time has been reduced and the downtime has been reduced.

    Which other solutions did I evaluate?

    I was not on the team that evaluated options before choosing PagerDuty Operations Cloud, so I cannot tell you about this.

    What other advice do I have?

    We are still in the process of implementing AI and automation through PagerDuty for incident response.

    I cannot answer that question right now because we are still testing PagerDuty's autonomous AI agents.

    The alert reduction feature has been reporting to us issues such as a client's website being down or the CRM website being down, which I would say has saved us.

    I would say that if your team can purchase PagerDuty Operations Cloud, you can definitely use it, although it is a bit costly for a startup, as was pointed out by my team too, but we have saved a lot of time.

    I gave this review a rating of 9 out of 10.

    Raj kuruhuri

    Alert noise has been cut and on-call teams respond faster to critical incidents

    Reviewed on Aug 10, 2026
    Review from a verified AWS customer

    What is our primary use case?

    Our main use case for using PagerDuty Operations Cloud is because our L1 team was under constant stress from alert storms and missed escalations that were threatening our service reliability. We chose it to streamline the escalation matrices, and it has been excellent for us. It has streamlined our escalation metrics and drastically reduced erroneous notifications, meaning my engineering teams are now only waking up in the middle of the night for actual revenue-impacting emergencies. It has brought stability to our operations and allowed us to respond accurately and swiftly.

    The main use case is to ensure that our services run uninterrupted while optimizing our operational cost. We are using it as the central hub for IT alerting, incident management, and automated on-call scheduling across the organization.

    How has it helped my organization?

    PagerDuty Operations Cloud has definitely been a game changer for the organization. The return on investment has been significant and highly visible. It has helped us reduce alert noise and surface resolution insight. We have saved an average of 30 minutes per incident, which totals a time savings of at least 70% to 80% per year. On a broader operational scale, the built-in ticketing management, the macros, and the automation have saved us the equivalent of at least three to four full-time employees, which has minimized our cost.

    Most importantly, it has minimized our downtime by catching signals earlier and translated that to productive revenue, which has avoided an estimated loss worth one million dollars for our company.

    What is most valuable?

    PagerDuty Operations Cloud has given us very clear AI-driven alert grouping. The autonomous AI agent calculated alerts and suppressed the noise. We used to get temporary traffic spikes or minor upgrade dips, and it has ensured that wherever there is a need for a security alert, it provides the security alert in real time, 24/7.

    PagerDuty Operations Cloud includes automated escalation, which helps us set up a call and join the right person via SMS, mobile push, or direct phone call. Another valuable feature is historical data insight, which gives clear insight during an incident and surfaces the historical data on forward-looking resolution from similar past events. This helps the team resolve issues faster.

    These features have given us real-time insight and a historical and present-time alert system via SMS, mobile push, or direct phone call. Even if the team has missed an alert at the regular time, it provides historical data and insight to ensure we understand why an incident has been raised or why a security breach has occurred.

    What needs improvement?

    PagerDuty Operations Cloud is working as expected, but if I were to suggest one improvement, I believe that the AI could be better. The accuracy can be improved because during consistent critical events, it occasionally groups unrelated alerts or misses a correlation, which forces our on-call team to fall back to manual triage.

    For how long have I used the solution?

    We have been using PagerDuty Operations Cloud in our organization for the last three and a half years.

    What do I think about the stability of the solution?

    We experienced 100% stability. We did not face any challenges.

    What do I think about the scalability of the solution?

    It is definitely a scalable product. We can easily scale up and down as per our needs. It has great scale.

    How are customer service and support?

    Customer support is a 10 out of 10. We had an interaction with them for one of the incidents, and they resolved it within 30 minutes.

    What's my experience with pricing, setup cost, and licensing?

    Our experience with the pricing, setup cost, and licensing of PagerDuty Operations Cloud has been excellent. We have chosen the pay-as-you-go licensing through AWS Marketplace, which has been a great option for us.

    Which other solutions did I evaluate?

    Before choosing PagerDuty Operations Cloud, we definitely evaluated other options and alternatives to ensure we were selecting the right solution. Some solutions we considered included Opsgenie by Atlassian, Better Stack, xMatters, and Splunk On-Call.

    What other advice do I have?

    The alert reduction feature of PagerDuty Operations Cloud has given us at least 30 minutes per incident and has reduced the overall tasks and work that was being done. We have seen a reduction in time and a reduction in the incidents, and we have saved a lot of money as well.

    The team has leveraged PagerDuty Operations Cloud to optimize the incident workflows in a positive way. It has given us clear insight into all incidents. We can easily track them in the history section. The impact is significant because it has optimized our workflows, optimized team capacity, and optimized the quality of incidents. The major impact is on time. The team has been allocated tasks on time, and they are able to deliver them on time.

    I give PagerDuty Operations Cloud a rating of nine out of ten. Since it has provided us exceptional results, I have given it a nine, and one point is deducted for the improvement I have suggested. Anyone who is looking for a dedicated platform for alerting and automation-related incidents should definitely go with PagerDuty Operations Cloud. They are the best in this category.

    Which deployment model are you using for this solution?

    Public Cloud

    If public cloud, private cloud, or hybrid cloud, which cloud provider do you use?

    Amazon Web Services (AWS)