Incident and Escalations Manager 3 - US-West

Datadog · Denver, Colorado, USA; San Francisco, California, USA · Support Engineering · listed October 9, 2026

The shape of it

Seniority
Manager
Experience asked
2–5 years
Where
Hybrid
Stated pay
$104,000 – $138,000 USD
Requirements listed
8
Length
933 words

In the posting’s own words

The Incident and Escalation Management team (IEM) is part of Datadog's Global Support Engineering (GSE) organization. IEM exists to continuously improve Datadog's overall customer experience during incidents and other critical moments. We're looking for experts with a background in incident management and escalation handling to provide fast incident response, clear ownership, and calm, confident communication for our global customer base.

What it asks for · 8

  • 5+ years of related professional experience, including experience in incident management and customer escalations, with demonstrable ownership of complex incidents and critical customer situations; experience in a SaaS, cloud, or observability environment is a plus
  • 2+ years of experience in a hands-on technical role at a software or cloud company, such as Support Engineering, Site Reliability Engineering, Solutions Architecture, or a similar technical function
  • Strong familiarity with cloud computing and modern software architectures; scripting ability in Python, JavaScript, or shell is a plus
  • Ability to correlate behaviors across known system interdependencies, assess customer and technical impact, and make sound decisions while remaining calm in high-pressure and ambiguous situations
  • Experience independently coordinating cross-functional stakeholders and driving complex incidents and escalations toward resolution
  • Experience driving projects from conception through delivery, with strong problem-solving skills and the ability to operate effectively in a fast-paced environment
  • Strong written and verbal English communication skills, with the ability to communicate clearly with both technical and customer-facing audiences
  • Bachelor's degree in Computer Science, Information Science/Technology, Engineering, or equivalent practical experience

What the job covers

  • Drive incidents and escalations to resolution from start to finish, including triage, coordination with our support, engineering and product teams. Own the response and keep customers and internal teams informed throughout.
  • Independently coordinate cross-functional stakeholders through complex incidents and escalations, making the call on prioritization, response strategy, and when to pull in additional teams or leadership.
  • Own communications during active incidents, including sensitive, large-scale, and publicly visible events, adapting the message for customers, internal teams, and leadership.
  • Lead projects that deliver operational improvements to how IEM detects, manages, and communicates during incidents and escalations as Datadog and its products evolve.
  • Create, maintain, and improve incident and escalation documentation, runbooks, and training materials, and lead retrospectives that turn individual incidents into durable process change.
  • Mentor peers on incident and escalation management best practices, and contribute to interviewing and hiring.
  • Design and implement real-time and proactive monitoring of customer infrastructure to surface customer-impacting risks before they escalate.
  • Identify recurring cloud infrastructure and product issues across incidents and escalations, and partner with Engineering to feed findings back and reduce recurrence.

Degree language

  • Bachelor's degree in Computer Science, Information Science/Technology, Engineering, or equivalent practical experience

Tools and skills named

Cloud & infra
  • Datadog5×
  • Observability
  • Site reliability
Ways of working
  • Cross-functional2×
  • Technical writing
Go to market
  • SaaS
  • Solutions architecture
Languages
  • JavaScript
  • Python

Words the posting leans on

  • incidents17×
  • customer11×
  • escalations10×
  • experience8×
  • engineering6×
  • management6×
  • technical5×
  • cloud4×
  • communication4×
  • complex incidents3×
  • global3×
  • iem3×
  • incident management3×
  • product3×
  • response3×
  • skills3×

Counted from the posting after the mission statement and the legal notices are set aside. The ones near the top are the ones a screener is looking for.

The posting, your resume, and the gaps between them. One click loads all three.

More open at Datadog

every open role at Datadog →

How this page was made

An automated read of a public job posting, fetched October 9, 2026 and last changed by Datadog on October 9, 2026. Every list above is pulled from the posting’s own sentences — nothing rewritten, nothing added, no judgment about the role or the company. Counts and seniority are read off the text by rule, so they can be wrong where the posting is unusual. The original is the only thing that binds. Openings close without warning; check the source before spending an evening on it.