Senior Incident Manager

Lambda - Remote, USA, United States, San Jose Office (Zanker)

Posted Jun 3, 2026

Benefits

Parental leave: Not verified
Non-birth-parent leave: Not verified
Family-building benefits: Fertility benefits: Not verified
Adoption assistance: Not verified
Surrogacy assistance: Not verified
Mental health support: Not verified
Relocation assistance: Not verified
Childcare support: Not verified
Learning budget: Not verified
Verification: Not verified
Salary: Not verified
401(k) match: Not verified not verified - source URL not recorded; timestamp not recorded

Was this benefit information wrong? Tell us.

Market context

Median wage (BLS OEWS): $61,842 national median
Projected growth (BLS Employment Projections): +1.9% - Slower

Matched to SOC 11-1021 - Operations aggregate by role bucket.

Source: U.S. Bureau of Labor Statistics, OEWS, May 2024 and Employment Projections, 2024-2034.

Schedule

Shift type: Not verified
Weekend work: Not verified

Application

Cover letter: Not verified
Assessment: Not verified
Deadline: Not stated

Where they hire

State eligibility is not yet verified.

About this role

Senior Incident Manager Remote, USA, United States, San Jose Office (Zanker) Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. We are seeking a Senior Incident Manager to lead critical incident response across our AI data center infrastructure. This role is responsible for coordinating rapid resolution of service-impacting events, improving operational resilience, and driving incident management best practices across infrastructure, networking, platform engineering, and data center operations. Role Overview The Senior Incident Manager is responsible for leading the end-to-end lifecycle of operational incidents impacting AI infrastructure and data center services. This individual acts as the central command point during major incidents, ensuring rapid triage, cross-team coordination, effective communication, and structured post-incident analysis. This role requires deep operational expertise in high-availability infrastructure, large-scale GPU clusters, networking, and cloud platforms , along with strong leadership and communication skills. What You'll Do Incident Leadership - Lead the response to critical (SEV-1 / SEV-2) incidents impacting AI infrastructure, GPU clusters, networking, storage, and data center operations. - Serve as the Incident Commander during major outages, coordinating engineering, networking, facilities, and vendor teams. - Act as the liaison between leadership and external teams during incidents / post-incidents to provide updates and status summaries. -

Read the full description at jobs.ashbyhq.com. FewerJobs shows a source-linked preview and links to the original posting.

Apply at jobs.ashbyhq.com

Apply link not verified; last-live date unavailable.

What verified means

Verified means a displayed claim has a recorded source field, a source URL when available, and a timestamp showing when FewerJobs checked or enriched the evidence.

Related jobs

Project Manager

Northrop Grumman - United States-Maryland-Linthicum
Insurance Operations Senior Associate

NB Bancorp INC - Chicago, IL
Sr Financial Planner

UMB Financial CORP - Kansas City MO
Operations Project Manager 2

Northrop Grumman - United States-Maryland-Baltimore

Senior Incident Manager

Benefits

Market context

Schedule

Application

Where they hire

About this role

What verified means

Related jobs

Project Manager

Insurance Operations Senior Associate

Sr Financial Planner

Operations Project Manager 2