Stripe

Stripe

Posted via Greenhouse

Incident Response Manager

Posted Sep 30, 2026

Role at a glance

Salary
Not Disclosed
Location
Ireland, Dublin Ireland Locations
Experience
5+ years of demonstrable major incident experience for organizations that run mission critical applications or always-on Saas environments.

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Incident Response Manager drives response to user-facing incidents across Stripe, coordinating teams to mitigate impact and communicate with users. The role also contributes to root cause analysis, remediation, and improvements to incident response operations.

What You'll Do

  • Act as an on-call Incident Commander, coordinating cross-functional responders and driving incident resolution.
  • Lead user-facing incidents across reliability, technical, security, and data privacy domains.
  • Assess user impact, provide situation reports, facilitate communications bridges, and support timely external communications.
  • Contribute to root cause analysis, post-mortems, remediation identification, and follow-up problem management.
  • Use incident trends and data to improve incident handling processes, metrics, and tooling.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

5+ years of demonstrable major incident experience in organizations running mission-critical applications or always-on SaaS environments; ability to lead multiple incidents concurrently and influence responders to resolve ambiguous problems and drive to root cause; intermediate understanding of application development, application architectures, and applications deployed in cloud environments; good understanding of physical, virtual, and container-based compute platforms; quantitative and analytical skills in data manipulation using SQL, Splunk, or other tools; strong task management, attention to detail, composure, and ability to think methodically and quickly in a high-pressure environment; exceptional written and verbal English communication skills, including translating complex technical issues for internal and external stakeholders.

Required

  • 5+ years of demonstrable major incident experience for organizations that run mission critical applications or always-on Saas environments.
  • Demonstrated ability to lead multiple incidents concurrently with authority and influence responders with agency and reasoning skills to...
  • Intermediate understanding of application development, application architectures, and applications deployed in cloud environments.
  • Good understanding of infrastructure, including physical, virtual, and container-based compute platforms.
  • Demonstrated quantitative and analytical skills in data manipulation using SQL, Splunk or other tools.
  • Excellent task management skills; detail-oriented, with the ability to remain composed, methodical, and think fast in a high-pressured...
  • Exceptional written and verbal English communication skills, with the ability to translate complex technical issues for internal and...

Preferred

  • Domain expertise in classes of incidents such as technical, privacy, security or crisis, with a strong desire to continuously learn...
  • Ability to review complex technical details regarding ongoing issues/events and convey key details to senior stakeholders to facilitate...
  • Experience with broad user-facing communications (e.g. status pages, tweets) and/or targeted communications (e.g. direct emails, support...
  • Familiarity operating or managing distributed architectures with the ability to correlate system behaviors based on known...
  • Demonstrated understanding of full stack development and support.

Original job description

Content provided by the employer

About the team

The Incident Ops team is a global 24/7 team responsible for driving incident response and management from detection to resolution. Stripe is proud of its five 9s reliability and this team is at the forefront of ensuring we keep it that way - working hand-in-hand with Reliability Eng and across the Tech Org. This team of incident response managers (IRM) is defined by our sense of ownership and how we drive incidents to resolution - marshaling the necessary cross-functional resources to respond to and resolve service outages, critical bugs, security attacks and anything that significantly impacts the users of our products. The team is user-first and ensures appropriate external communications from Stripe and senior management to keep our users informed of disruption to their experience of Stripe. The team is skilled in communications, incident handling and technical adeptness as incidents can arise from anywhere and cut across products and orgs in Stripe.

What you’ll do

As an Incident Response Manager (IRM), you’ll play the key role in driving the right level of response from Stripes to incidents, determining impact, rallying Stripes to mitigate, communicating to users and ensuring appropriate remediations and orchestrate the Root Cause Analysis (RCA) process. You’ll work hand-in-hand with IRMs and engineers globally to ensure solid 24/7 coverage on how we monitor, detect, respond, communicate and mitigate incidents. When not managing incidents, you'll help scale our ability to respond to incidents, improve our operations, analyze data to provide insights and deepen our technical expertise in products. As a result, you’ll be seen as the protector of our users - in minimizing the impact of incidents on their business and ensuring that Stripe is always thinking of our users.

Responsibilities

  • Act as an on-call Incident Commander, responsible for driving and managing incident resolution with a high level of urgency, cross-functional collaboration, and accuracy, while partnering with a global and diverse set of teams, including Engineering, Product, Policy, Risks, PR, Legal, Execs, etc.
  • Lead all user-facing incidents across domains at Stripe - including reliability, technical, security, and data privacy
  • "User First" approach to determine impact, providing accurate situation reports, facilitating comms bridges, and ensuring useful and timely external communications to users
  • Proactively update internal stakeholders, make decisions through data and influence by partnering with Engineering, Sales, Support and other cross-functional teams
  • Contribute to the root cause analysis process while conducting post-mortems, remediations identification, and ensure problem management tasks meet SLA and user expectations
  • Drive improvements in the incident handling process and incident management metrics and tooling based on trends and data of Stripe's incidents in collaboration with engineering, product and operations teams
  • Contribute to processes, projects, forums, or groups that impact and grow team culture positively

Minimum Requirements

  • 5+ years of demonstrable major incident experience for organizations that run mission critical applications or always-on Saas environments.
  • Demonstrated ability to lead multiple incidents concurrently with authority and influence responders with agency and reasoning skills to resolve ambiguous problems and drive to root cause.
  • Intermediate understanding of application development, application architectures, and applications deployed in cloud environments
  • Good understanding of infrastructure, including physical, virtual, and container-based compute platforms
  • Demonstrated quantitative, and analytical skills in data manipulation using SQL, Splunk or other tools.
  • Excellent task management skills, must be detail-oriented with ability to remain composed, methodical, and think fast in a high-pressured environment.
  • Exceptional written and verbal English communication skills, with the ability to translate complex technical issues for internal and external stakeholders.

Preferred qualifications

    • Domain expertise in classes of incidents such as technical, privacy, security or crisis with a strong desire to continuously learn about Stripe's products, technical issues and systems.
    • Ability to review complex technical details regarding ongoing issues/events and convey the key details to senior stakeholders to facilitate real-time decision making.
    • Experience with broad user-facing communications (e.g. status pages, tweets) and/or targeted communications (e.g. direct emails, support ticket responses).
    • Familiarity operating or managing distributed architectures with the ability to correlate system behaviors based on known inter-dependencies.
    • Demonstrated understanding of full stack development and support
Stripe

About the company

Stripe

Large Enterprise

Stripe is a leading financial technology company that provides a robust platform for online payment processing and ecommerce solutions. Founded in 2010, Stripe enables businesses of all sizes to accept payments, send payouts, and manage their online transactions seamlessly. With a focus on innovation and user experience, Stripe offers a range of APIs and tools that empower developers to integrate payment functionality quickly and efficiently. Trusted by millions of businesses worldwide, Stripe continues to shape the future of online commerce and financial services.