FewerJobs.
All jobs

Sr. SDM, AI Inference Technology, Neuron SDK

Amazon - Seattle, Washington, USA

Posted Jun 16, 2025

Benefits

Parental leave
6 weeks From the posting source
Non-birth-parent leave
6 weeks From the posting source
Family-building benefits
  • Fertility benefits: Not verified
  • Adoption assistance: Not verified
  • Surrogacy assistance: Not verified
Mental health support
Not verified
Relocation assistance
Not verified
Childcare support
Not verified
Learning budget
Not verified
Verification
Source-linked checked Jun 7, 2026
Salary
$253K-$342K From the posting source
401(k) match
Reported from DOL Form 5500 industry filing (not employer-specific)

Was this benefit information wrong? Tell us.

Market context

U.S. role benchmark (BLS OEWS)
$102,662 U.S. median for this role
Projected growth (BLS Employment Projections)
+5.4% - Faster than average

190% above the BLS role benchmark for product management aggregate.

Matched to SOC 11-1021 - Product Management aggregate by role bucket.

Source: U.S. Bureau of Labor Statistics, OEWS, May 2024 and Employment Projections, 2024-2034.

Role

Role function
Product From the posting source
Seniority
Senior From the posting source

Schedule

Shift type
Not verified
Weekend work
Not verified

Company

Equity
Offered Verified - SEC 10-K source

Application

Cover letter
Not verified
Assessment
Not verified
Deadline
Not stated

Where they hire

State eligibility is not yet verified.

About this role

Sr. SDM, AI Inference Technology, Neuron SDK Seattle, Washington, USA AWS Utility Computing (UC) provides product innovations - from foundational services such as Amazon Elastic Compute Cloud (EC2), to new product innovations that continue to set AWS's services and features apart in the industry. Come develop inference acceleration for AWS Neuron, the complete software stack for Trainium, Amazon's custom cloud-scale machine learning accelerators that power the latest AI models As the Sr. SDM for the Inference Technology Team, you will lead a strong team of managers and engineers to build fundamental inference technology building blocks and libraries to enable AI developers to optimize model for inference on Trainium and Inferentia devices. You will be responsible for the full development life cycle of inference library and feature development, including reliability and scalability. You will develop the Neuronx_Distributed Inference Libraries and contribute to other popular open source Inference Libraries, enabling customers to optimize LLMs, multimodal, and generative models. The ideal candidate will have an established background in delivering AI feature support for demanding, fast-changing priorities or delivering high-performance models using distributed inference libraries. The ideal candidate should have a strong technical ability to understand and manage a vertically integrated system stack that consisting of hardware, frameworks, and workflows. A day in the life You will work with the executive leadership and other senior management and technical leaders to define product directions and deliver them to customers. We build massive-scale distributed training and inference solutions, developing the full stack of software, servers and

Read the full description at www.amazon.jobs. FewerJobs shows a preview and links to the original posting.

Apply at amazon.jobs

Apply link not verified; last-live date unavailable.

What verified means

Verified means a displayed claim has field-level provenance to a source FewerJobs pulled: a government or employer source, or the original job posting. Posting-sourced facts are employer-stated and are labeled separately from government records.

Related jobs