Sr. SDM, AI Inference Technology, Neuron SDK
Amazon - Seattle, Washington, USA
Posted Jun 16, 2025
Benefits
- Parental leave
- 6 weeks From the posting source
- Non-birth-parent leave
- 6 weeks From the posting source
- Family-building benefits
-
- Fertility benefits: Not verified
- Adoption assistance: Not verified
- Surrogacy assistance: Not verified
- Mental health support
- Not verified
- Relocation assistance
- Not verified
- Childcare support
- Not verified
- Learning budget
- Not verified
- Verification
- Source-linked checked Jun 7, 2026
- Salary
- $253K-$342K From the posting source
- 401(k) match
- Reported from DOL Form 5500 industry filing (not employer-specific)
Was this benefit information wrong? Tell us.
Market context
- U.S. role benchmark (BLS OEWS)
- $102,662 U.S. median for this role
- Projected growth (BLS Employment Projections)
- +5.4% - Faster than average
190% above the BLS role benchmark for product management aggregate.
Matched to SOC 11-1021 - Product Management aggregate by role bucket.
Source: U.S. Bureau of Labor Statistics, OEWS, May 2024 and Employment Projections, 2024-2034.
Schedule
- Shift type
- Not verified
- Weekend work
- Not verified
Company
- Equity
- Offered Verified - SEC 10-K source
Application
- Cover letter
- Not verified
- Assessment
- Not verified
- Deadline
- Not stated
Where they hire
State eligibility is not yet verified.
About this role
Sr. SDM, AI Inference Technology, Neuron SDK Seattle, Washington, USA AWS Utility Computing (UC) provides product innovations - from foundational services such as Amazon Elastic Compute Cloud (EC2), to new product innovations that continue to set AWS's services and features apart in the industry. Come develop inference acceleration for AWS Neuron, the complete software stack for Trainium, Amazon's custom cloud-scale machine learning accelerators that power the latest AI models As the Sr. SDM for the Inference Technology Team, you will lead a strong team of managers and engineers to build fundamental inference technology building blocks and libraries to enable AI developers to optimize model for inference on Trainium and Inferentia devices. You will be responsible for the full development life cycle of inference library and feature development, including reliability and scalability. You will develop the Neuronx_Distributed Inference Libraries and contribute to other popular open source Inference Libraries, enabling customers to optimize LLMs, multimodal, and generative models. The ideal candidate will have an established background in delivering AI feature support for demanding, fast-changing priorities or delivering high-performance models using distributed inference libraries. The ideal candidate should have a strong technical ability to understand and manage a vertically integrated system stack that consisting of hardware, frameworks, and workflows. A day in the life You will work with the executive leadership and other senior management and technical leaders to define product directions and deliver them to customers. We build massive-scale distributed training and inference solutions, developing the full stack of software, servers and
Read the full description at www.amazon.jobs. FewerJobs shows a preview and links to the original posting.
Apply link not verified; last-live date unavailable.
What verified means
Verified means a displayed claim has field-level provenance to a source FewerJobs pulled: a government or employer source, or the original job posting. Posting-sourced facts are employer-stated and are labeled separately from government records.
Related jobs
-
Senior AI/ML Platform Engineer (LLM/SLM Inference)
Cisco - San Jose, California, US
-
Sr. Software Engineer I, Applied AI
Axon Enterprise - Seattle, Washington, United States
-
Sr. Staff Edge AI Applied Machine Learning Engineer
Ambiq Micro INC - Austin, Texas, United States
-
Senior Product Manager, AI Literacy
Blackbaud INC - Remote - Anywhere - USA