Senior Incident Manager
Lambda - Remote, USA, United States, San Jose Office (Zanker)
Posted Jun 3, 2026
Benefits
- Parental leave
- Not verified
- Non-birth-parent leave
- Not verified
- Family-building benefits
-
- Fertility benefits: Not verified
- Adoption assistance: Not verified
- Surrogacy assistance: Not verified
- Mental health support
- Not verified
- Relocation assistance
- Not verified
- Childcare support
- Not verified
- Learning budget
- Not verified
- Verification
- Not verified
- Salary
- Not verified
- 401(k) match
- Not verified not verified - source URL not recorded; timestamp not recorded
Was this benefit information wrong? Tell us.
Market context
- Median wage (BLS OEWS)
- $61,842 national median
- Projected growth (BLS Employment Projections)
- +1.9% - Slower
Matched to SOC 11-1021 - Operations aggregate by role bucket.
Source: U.S. Bureau of Labor Statistics, OEWS, May 2024 and Employment Projections, 2024-2034.
Schedule
- Shift type
- Not verified
- Weekend work
- Not verified
Application
- Cover letter
- Not verified
- Assessment
- Not verified
- Deadline
- Not stated
Where they hire
State eligibility is not yet verified.
About this role
Senior Incident Manager Remote, USA, United States, San Jose Office (Zanker) Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU. If you'd like to build the world's best AI cloud, join us. We are seeking a Senior Incident Manager to lead critical incident response across our AI data center infrastructure. This role is responsible for coordinating rapid resolution of service-impacting events, improving operational resilience, and driving incident management best practices across infrastructure, networking, platform engineering, and data center operations. Role Overview The Senior Incident Manager is responsible for leading the end-to-end lifecycle of operational incidents impacting AI infrastructure and data center services. This individual acts as the central command point during major incidents, ensuring rapid triage, cross-team coordination, effective communication, and structured post-incident analysis. This role requires deep operational expertise in high-availability infrastructure, large-scale GPU clusters, networking, and cloud platforms , along with strong leadership and communication skills. What You'll Do Incident Leadership - Lead the response to critical (SEV-1 / SEV-2) incidents impacting AI infrastructure, GPU clusters, networking, storage, and data center operations. - Serve as the Incident Commander during major outages, coordinating engineering, networking, facilities, and vendor teams. - Act as the liaison between leadership and external teams during incidents / post-incidents to provide updates and status summaries. -
Read the full description at jobs.ashbyhq.com. FewerJobs shows a source-linked preview and links to the original posting.
Apply link not verified; last-live date unavailable.
What verified means
Verified means a displayed claim has a recorded source field, a source URL when available, and a timestamp showing when FewerJobs checked or enriched the evidence.
Related jobs
-
Project Manager
Northrop Grumman - United States-Maryland-Linthicum
-
Insurance Operations Senior Associate
NB Bancorp INC - Chicago, IL
-
Sr Financial Planner
UMB Financial CORP - Kansas City MO
-
Operations Project Manager 2
Northrop Grumman - United States-Maryland-Baltimore