# What&#39;s the most reliable incident response platform based on reviews from teams running large-scale deployments based on user reviews?

<p class="elv-tracking-normal elv-text-default elv-font-figtree elv-text-base elv-leading-base elv-font-normal" elv="true">Looking for input from G2 reviewers and engineering managers, SRE leads, and IT operations directors at organisations running incident response programmes across hundreds of services, dozens of on-call rotations, or multiple geographic teams in the <a class="a a--md" elv="true" href="https://www.g2.com/categories/incident-response">Incident Response Software category</a>.</p><p class="elv-tracking-normal elv-text-default elv-font-figtree elv-text-base elv-leading-base elv-font-normal" elv="true">The platforms with the strongest large-scale deployment reliability evidence:</p><ul>
<li>
<a class="a a--md" elv="true" href="https://www.g2.com/products/pagerduty/reviews"><strong>PagerDuty</strong></a>: The Event Intelligence and AIOps capabilities that correlate and deduplicate alerts across high-volume environments are credited for preventing alert fatigue from degrading on-call responsiveness at scale. </li>
<li>
<a class="a a--md" elv="true" href="https://www.g2.com/products/datadog/reviews"><strong>Datadog Incident Management</strong></a>: The correlation between an incident declaration and the underlying service metrics, traces, and logs without requiring a context switch between tools is credited for reducing the time between alert and diagnosis during active incidents. </li>
<li>
<a class="a a--md" elv="true" href="https://www.g2.com/products/rootly/reviews"><strong>Rootly</strong></a>: The automation that handles the routine coordination steps during an incident, freeing responders to focus on technical diagnosis and resolution, is credited for reducing the cognitive overhead of incident management at scale. </li>
<li>
<a class="a a--md" elv="true" href="https://www.g2.com/products/firehydrant/reviews"><strong>FireHydrant</strong></a>: The automated status page updates and stakeholder communication that run alongside the technical resolution workflow are credited for reducing the communication overhead that falls on incident commanders during active incidents. </li>
<li>
<a class="a a--md" elv="true" href="https://www.g2.com/products/atlassian-opsgenie/reviews"><strong>Atlassian Opsgenie</strong></a>: The global alert routing across complex on-call rotation structures and escalation chains is credited for managing the multi-team coordination requirements of large engineering organisations. </li>
</ul><p class="elv-tracking-normal elv-text-default elv-font-figtree elv-text-base elv-leading-base elv-font-normal" elv="true">For SRE leads and engineering managers who have run an incident response platform across more than 50 services in production: what was the first at-scale failure mode your platform exhibited that did not appear in small-scale evaluation, and what configuration or architecture change resolved it?</p>

##### Post Metadata
- Posted at: 17 days ago
- Author title: Marketing Executive
- Net upvotes: 1


## Comments
### Comment 1

Splunk Enterprise Security and SentinelOne are both heavily reviewed and proven at large-scale deployments, with the integration breadth and reliability that large security teams depend on when things go wrong.

##### Comment Metadata
- Posted at: 9 days ago





## Related discussions
- [How well does Trello scale into a larger team?](https://www.g2.com/discussions/1-how-well-does-trello-scale-into-a-larger-team)
  - Posted at: over 13 years ago
  - Comments: 6
- [Can we please add a new section](https://www.g2.com/discussions/2-can-we-please-add-a-new-section)
  - Posted at: over 13 years ago
  - Comments: 0
- [Quantifiable benefits from implementing your CRM](https://www.g2.com/discussions/quantifiable-benefits-from-implementing-your-crm)
  - Posted at: over 13 years ago
  - Comments: 4


