# FewerJobs export - 100 curated jobs
Generated: 2026-08-10T05:18:00.875Z
Source: https://fewerjobs.com

## Filters applied
- **q**: Neural Magic
- **quality_floor**: default
- **match_401k_strict**: true
- **parental_strict**: true
- **non_birth_strict**: true
- **pto_strict**: true
- **include_older**: false
- **apply_url_verified**: false
- **page**: 1
- **per_page**: 100
- **sort**: relevance

## Jobs
### ITストラテジー コンサルタント - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-21
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/IT-_R-056989-5
- Excerpt: ITストラテジー コンサルタント Tokyo posted: Posted 22 Days Ago

### Forward Deployed Engineer, AI Inference (vLLM and Kubernetes) - Neural Magic
- Location: 6 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-19
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-US-WA/Forward-Deployed-Engineer--AI-Inference--vLLM-and-Kubernetes-_R-053737-1
- Excerpt: Forward Deployed Engineer, AI Inference (vLLM and Kubernetes) 6 Locations posted: Posted 24 Days Ago

### Consultant - OpenShift - Neural Magic
- Location: Canberra (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Canberra/Consultant---OpenShift_R-057082-1
- Excerpt: Consultant - OpenShift Canberra posted: Posted Today

### Consultant - Neural Magic
- Location: Canberra (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Canberra/Consultant_R-054263-1
- Excerpt: Consultant Canberra posted: Posted 30+ Days Ago

### Architect - Neural Magic
- Location: Remote UK (remote)
- Salary: Not disclosed
- Posted: 2026-06-02
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-UK/Architect_R-057427-1
- Excerpt: Architect Remote UK posted: Posted 10 Days Ago

### Senior AI Architect, APAC - Neural Magic
- Location: Singapore (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-14
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Singapore/Senior-AI-Architect--APAC_R-056987-1
- Excerpt: Senior AI Architect, APAC Singapore posted: Posted 29 Days Ago

### Senior Consultant - Neural Magic
- Location: Canberra (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Canberra/Senior-Consultant_R-040324-1
- Excerpt: Senior Consultant Canberra posted: Posted 30+ Days Ago

### Project Manager - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Project-Manager_R-056366-1
- Excerpt: Project Manager Tokyo posted: Posted 30+ Days Ago

### Principal Architect - Neural Magic
- Location: Remote Australia (remote)
- Salary: Not disclosed
- Posted: 2026-06-11
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-Australia/Principal-Architect_R-054717-1
- Excerpt: Principal Architect Remote Australia posted: Posted Yesterday

### Software Engineer - Neural Magic
- Location: Raleigh (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Raleigh/Software-Engineer_R-056802
- Excerpt: Software Engineer Raleigh posted: Posted 30+ Days Ago

### Cloud Architect - OpenShift and AI - Neural Magic
- Location: Puteaux (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Puteaux/Cloud-Architect---OpenShift-and-AI_R-053695-2
- Excerpt: Cloud Architect - OpenShift and AI Puteaux posted: Posted Today

### Junior Solution Architect - Neural Magic
- Location: Sao Paulo (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-09
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Sao-Paulo/Junior-Solution-Architect_R-057636
- Excerpt: Junior Solution Architect Sao Paulo posted: Posted 3 Days Ago

### Sales Specialist - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-21
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Sales-Specialist-4_R-054328-2
- Excerpt: Sales Specialist Tokyo posted: Posted 22 Days Ago

### Senior Manager, AI Inference - Neural Magic
- Location: Boston (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-19
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Boston/Senior-Manager--AI-Inference_R-056089
- Excerpt: Senior Manager, AI Inference Boston posted: Posted 24 Days Ago

### Consultant, OpenShift - Neural Magic
- Location: 5 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-07
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-US-DC/Consultant--OpenShift_R-057616-1
- Excerpt: Consultant, OpenShift 5 Locations posted: Posted 5 Days Ago

### Junior Consultant - Neural Magic
- Location: New Delhi (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-08
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/New-Delhi/Junior-Consultant_R-054554-1
- Excerpt: Junior Consultant New Delhi posted: Posted 4 Days Ago

### Middleware Consultant - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Middleware-Consultant_R-055803-1
- Excerpt: Middleware Consultant Tokyo posted: Posted 30+ Days Ago

### Manager Solution Architecture - Neural Magic
- Location: Santiago (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-01
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Santiago/Manager-Solution-Architecture_R-057364-1
- Excerpt: Manager Solution Architecture Santiago posted: Posted 11 Days Ago

### Architect, OpenShift Virtualization - Neural Magic
- Location: 3 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-10
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-CA-AB/Architect--OpenShift-Virtualization_R-057823-1
- Excerpt: Architect, OpenShift Virtualization 3 Locations posted: Posted 2 Days Ago

### Principal Software Engineer- AI - Neural Magic
- Location: Boston (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-15
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Boston/Principal-Software-Engineer--AI_R-055185-1
- Excerpt: Principal Software Engineer- AI Boston posted: Posted 28 Days Ago

### Senior AI Architect, APAC - Neural Magic
- Location: 3 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Bangkok---MSO---Gaysorn/Senior-AI-Architect--APAC_R-056542-1
- Excerpt: Senior AI Architect, APAC 3 Locations posted: Posted 30+ Days Ago

### Edge Specialist Solution Architect - Neural Magic
- Location: Singapore (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-27
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Singapore/Edge-Specialist-Solution-Architect_R-056922-1
- Excerpt: Edge Specialist Solution Architect Singapore posted: Posted 16 Days Ago

### Senior Software Engineer, AI Inference - Neural Magic
- Location: 2 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-19
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Boston/Senior-Software-Engineer_R-053075-1
- Excerpt: Senior Software Engineer, AI Inference 2 Locations posted: Posted 24 Days Ago

### AI Sales Specialist - Singapore - Neural Magic
- Location: Singapore (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-21
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Singapore/AI-Sales-Specialist--Southeast-Asia--Singapore-_R-054963
- Excerpt: AI Sales Specialist - Singapore Singapore posted: Posted 22 Days Ago

### Neuroengineer Intern - Neuralink
- Location: South San Francisco, California, United States (unspecified)
- Salary: Not disclosed
- Posted: 2025-09-30
- 401(k) match: listed (not source-backed)
- Apply: https://boards.greenhouse.io/neuralink/jobs/7483748003?gh_jid=7483748003
- Excerpt: Neuroengineer Intern South San Francisco, California, United States About Neuralink: We are creating devices that enable a bi-directional interface with the brain. These devices allow us to restore movement to the paralyzed, restore sight to the blind, and revolutionize how humans interact with their digital world. Team Description: The Next Gen team at Neuralink is developing the next generation of brain-computer interfaces. We are laying the groundwork for intuitive, high-dimensional, and bidirectional interfaces between brains and machines, with the goal of helping people challenged by a variety of neurological disorders and conditions. Our team consists of scientists and engineers, working closely together to define the engineering requirements for these future products. As a Neuroengineer Intern on this team, you will contribute across a wide range of projects, from designing novel experimental preparations, to interpreting neural signals and behavioral data, to developing novel BCI paradigms. Successful candidates will be highly adaptable, able to deploy their core technical and creative skills to tackle a wide range of problems, and have a keen sense of urgency. Job Description and Responsibilities: - Investigate and map neural correlates of sensory and internally-generated states. - Build and iterate on advanced machine learning architectures to map neural activity to complex, multi-dimensional outputs. - Design, execute, interpret, and communicate the results of experiments, ranging from non-human primate (NHP) electrophysiology to human fMRI studies. - Present results in a fast-paced, highly collaborative setting. Required Qualifications: - Evidence of exceptional ability in neuroscience, machine learning, biomedical engineering, or a related

### Machine Learning Engineer - Neuralink
- Location: Austin, Texas, United States; South San Francisco, California, United States (unspecified)
- Salary: $199K-$331K
- Posted: 2023-06-29
- 401(k) match: listed (not source-backed)
- Apply: https://boards.greenhouse.io/neuralink/jobs/5663271003?gh_jid=5663271003
- Excerpt: Machine Learning Engineer Austin, Texas, United States; South San Francisco, California, United States About Neuralink: We are creating devices that enable a bi-directional interface with the brain. These devices allow us to restore movement to the paralyzed, restore sight to the blind, and revolutionize how humans interact with their digital world. About the Team: The BCI team develops the software and systems that communicate with the brain. These systems decode raw neural signals into useful actions, such as moving a cursor, typing, or actuating a robotic arm. Additionally, real-world data, such as video feeds, can be encoded into neural data to project images into the visual cortex. We also work closely with users to gather feedback, make improvements, and fundamentally reshape the user experience and interface of the BCI. About the Role: Engineers on the BCI team utilize signal processing and machine learning to communicate with the brain. You will have access to the most cutting-edge neural interface hardware and develop state-of-the-art neural encoders and decoders. No prior knowledge of neuroscience is required; we value simple solutions grounded in first principles. Neuralink designs all hardware in-house, from custom ASICs to thin-film arrays. There is no part of the technical design that cannot change. Learnings from your work will directly influence next-generation device architecture. Job Responsibilities: - Telepathy Product: Develop and refine models that decode neural data, enabling individuals with paralysis to reliably type at 35 words per minute or control robotics arms for activities of daily living. - Blindsight Product:

### Neuroscience Specialist - Freelance AI Trainer Project - Agency
- Location: United States of America (unspecified)
- Salary: $6-$65/hr
- Posted: 2026-06-07
- Apply: https://job-boards.eu.greenhouse.io/agency/jobs/4603886101
- Excerpt: Neuroscience Specialist - Freelance AI Trainer Project United States of America Are you a neuroscience expert eager to shape the future of AI? Large‑scale language models are evolving from clever chatbots into powerful engines of scientific discovery. With high‑quality training data, tomorrow's AI can democratize world‑class education, keep pace with cutting‑edge research, and streamline lab work for scientists everywhere. That training data begins with you-we need your expertise to help power the next generation of AI. We're looking for neuroscience specialists who live and breathe neuroanatomy, neurophysiology, neurochemistry, cognitive neuroscience, behavioral neuroscience, neuroimaging, neurodegenerative diseases, neural networks, sensory processing, and synaptic plasticity. You'll challenge advanced language models on topics like brain structure and function, neuroplasticity, neurogenesis, neurotransmitter systems, brain disorders, and neurodevelopmental processes-documenting every failure mode so we can harden model reasoning. On a typical day, you will converse with the model on lab scenarios and theoretical neuroscience questions, verify factual accuracy and logical soundness, capture reproducible error traces, and suggest improvements to our prompt engineering and evaluation metrics. A master's or PhDs in neuroscience or a closely related life‑science field is ideal; peer‑reviewed publications, wet‑lab or field research, or hands‑on neuroimaging projects signal fit. Clear, metacognitive communication-“showing your work”-is essential. Ready to turn your neuroscience expertise into the knowledge base for tomorrow's AI? Apply today and start teaching the model that will teach the world. We offer a pay range of $6-to- $65 per hour, with the exact rate determined after evaluating your experience, expertise, and geographic location. Final

### Machine Learning Engineer, AWS Neuron Inference, Annapurna ML - Amazon
- Location: Seattle, Washington, USA (unspecified)
- Salary: $144K-$194K
- Posted: 2025-12-22
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/3147714/machine-learning-engineer-aws-neuron-inference-annapurna-ml
- Excerpt: Machine Learning Engineer, AWS Neuron Inference, Annapurna ML Seattle, Washington, USA AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the Trn2 and future Trn3 servers that use them. This role is for a software engineer in the Machine Learning Applications (ML Apps) team for AWS Neuron. This role develops, enables and performance tunes building blocks for all key ML model families, including Llama3, GPT OSS, Qwen3, DeepSeek and beyond. The Neuron Inference Technology team works side by side with the Inference Model Enablement, compiler runtime engineers to create, build and tune high-performance distributed inference solutions for the latest generation Trainium accelerators. Experience optimizing LLM inference performance with kernels, Python, PyTorch or JAX is a must. Key job responsibilities This team develops optimized building blocks for the Neuron distributed inference library, tuning them to ensure highest performance and maximize efficiency running on Trn2 and Trn3 servers. A day in the life As you develop technology components, you'll create metrics, implement automation and other improvements, and resolve the root cause of software defects. You'll also participate in design discussions, code review, and communicate with internal and external stakeholders. You will work cross-functionally with teams across Neufon in a fast-paced startup-like development environment, where we constantly stay on top of the latest priorities as the AI landscape evolves. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an

### Software Development Engineer - AI/ML, Amazon Neuron, Multimodal Inference - Amazon
- Location: Seattle, Washington, USA (unspecified)
- Salary: $144K-$194K
- Posted: 2026-05-08
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/10415417/software-development-engineer-ai-ml-amazon-neuron-multimodal-inference
- Excerpt: Software Development Engineer - AI/ML, Amazon Neuron, Multimodal Inference Seattle, Washington, USA The Annapurna Labs team at Amazonbuilds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch and JAX enabling unparalleled ML inference and training performance. The Inference Enablement and Acceleration team is at the forefront of running a wide range of models and supporting novel architecture alongside maximizing their performance for AWS's custom ML accelerators. Working across the stack from PyTorch till the hardware-software boundary, our engineers build systematic infrastructure, innovate new methods and create high-performance kernels for ML functions, ensuring every compute unit is fine tuned for optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. As part of the broader Neuron organization, our team works across multiple technology layers - from frameworks and kernels and collaborate with compiler to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work at the intersection

### Research Scientist, Robotics, Embodied AI, DeepMind - DeepMind
- Location: Mountain View, CA, USA (unspecified)
- Salary: $147K-$211K
- Posted: 2026-06-04
- Parental leave: 18 weeks (not source-backed)
- Non-birth-parent leave: 18 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Apply: https://www.google.com/about/careers/applications/signin?jobId=CiUAL2FckQpFyCqW3nIHT9aA1sg6WL2fBdc1uJ1iApABvQQsEBx7EjsACxwdTK3_2ZHCRHY_Rvd9MZCRMeJxkJXR-lrlIMtQBlFZ5uuKJ8bYNER81VNL2MJPmAx3R734ZRvklw%3D%3D_V2&loc=US&title=Research+Scientist
- Excerpt: Research Scientist, Robotics, Embodied AI, DeepMind Mountain View, CA, USA We believe there are many problems in the world in which robotics could play a significant role in making it easier, faster and safer for people to get things done. We're looking for roboticists, designers, hardware and software engineers to help us explore these possibilities, develop breakthrough technologies, and build new products that could help millions of people. At DeepMind Robotics we are pioneering bringing AI to the physical world. We are powering an era of physical agents, enabling robots to perceive, plan, think, use tools, and act to better solve tasks. In this role, you will develop Vision Language Action (VLA) models that combine Gemini's world understanding with physical actions to directly control robots. You will include Gemini Robotics (our Gemini model for the physical world), and Gemini Robotics On-Device (our Gemini model that runs without a data network). You will also develop reasoning and agentic systems for the physical world, including Gemini Robotics-ER, a Gemini agent with spatial understanding. You will enable robots to perform a range of tasks, respond interactively to their environment, achieve dexterity, and reason over long multi-step tasks. You will focus on advancing in areas of general purpose robotics, including real world understanding, action generalization, human robot interaction, whole-body control, and continual learning. Additionally, you will partner with robotics companies to bring this intelligence to applications at scale. Artificial intelligence will be one of humanity's most transformative inventions. At Google DeepMind, we are a

### Solution Architect - Neural Magic
- Location: Mumbai (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-13
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Mumbai/Solution-Architect_R-052980-1
- Excerpt: Solution Architect Mumbai posted: Posted 30 Days Ago

### Senior Technical Writer - Neural Magic
- Location: Raleigh (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Raleigh/Senior-Technical-Writer_R-053114-1
- Excerpt: Senior Technical Writer Raleigh posted: Posted Today

### AI Specialist Solution Architect, Southeast Asia (Singapore) - Neural Magic
- Location: Singapore (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-15
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Singapore/AI-Specialist-Solution-Architect--Southeast-Asia--Singapore-_R-055339
- Excerpt: AI Specialist Solution Architect, Southeast Asia (Singapore) Singapore posted: Posted 28 Days Ago

### Principal Solution Architect - Neural Magic
- Location: 3 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-31
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-US-VA/Principal-Solution-Architect_R-056757-1
- Excerpt: Principal Solution Architect 3 Locations posted: Posted 12 Days Ago

### Principal Software Engineer - Agentic AI - Neural Magic
- Location: 2 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-10
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Bangalore---Carina/Principal-Software-Engineer---Agentic-AI_R-056934-1
- Excerpt: Principal Software Engineer - Agentic AI 2 Locations posted: Posted 2 Days Ago

### OpenStack Consultant - Neural Magic
- Location: Puteaux (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-03
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Puteaux/OpenStack-Consultant_R-056721-1
- Excerpt: OpenStack Consultant Puteaux posted: Posted 9 Days Ago

### Senior ML Kernel Performance Engineer - Amazon
- Location: Toronto, Ontario, CAN (unspecified)
- Salary: CAD 151K-252K
- Posted: 2025-08-14
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/3059954/senior-ml-kernel-performance-engineer
- Excerpt: Senior ML Kernel Performance Engineer Toronto, Ontario, CAN The Annapurna Labs team at Amazon builds Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The Acceleration Kernel Library team is at the forefront of maximizing performance for Amazon's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions, ensuring every FLOP counts in delivering optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. The Amazon Neuron SDK, developed by the Annapurna Labs team at Amazon, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference and training performance. As part of the broader Neuron Compiler organization, our team works across multiple technology layers - from frameworks and compilers to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work at the intersection of machine learning, high-performance computing, and distributed architectures, where you'll help shape the future of AI acceleration technology This is an opportunity to work on cutting-edge products at the intersection of machine-learning, high-performance computing, and distributed

### Consultant in Adelaide - Neural Magic
- Location: Remote Australia (remote)
- Salary: Not disclosed
- Posted: 2026-06-11
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-Australia/Consultant-in-Adelaide_R-056665-1
- Excerpt: Consultant in Adelaide Remote Australia posted: Posted Yesterday

### Senior Software Engineer - AI/ML, AWS Neuron Inference - Amazon
- Location: Seattle, Washington, USA (unspecified)
- Salary: $168K-$227K
- Posted: 2026-05-18
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Surrogacy assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/10422684/senior-software-engineer-ai-ml-aws-neuron-inference
- Excerpt: Senior Software Engineer - AI/ML, AWS Neuron Inference Seattle, Washington, USA AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators. This role is for a senior software engineer in the Machine Learning Inference Applications team. This role is responsible for development and performance optimization of core building blocks of LLM Inference - Attention, MLP, Quantization, Speculative Decoding, Mixture of Experts, etc. The team works side by side with chip architects, compiler engineers and runtime engineers to deliver performance and accuracy on Neuron devices across a range of models. Key job responsibilities Responsibilities of this role include adapting latest research in LLM optimization to Neuron chips to extract best performance from both open source as well as internally developed models. Working across teams and organizations is key. About the team Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we're building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future. Basic Qualifications: - 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience - Bachelor's degree in computer science or equivalent - 5+ years of programming using a

### Applied Scientist, Neuron ARG, Annapurna ML - Amazon
- Location: Seattle, Washington, USA (unspecified)
- Salary: $143K-$193K
- Posted: 2026-05-13
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Surrogacy assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/10418948/applied-scientist-neuron-arg-annapurna-ml
- Excerpt: Applied Scientist, Neuron ARG, Annapurna ML Seattle, Washington, USA The Automated Reasoning Group in the Amazon Neuron team is looking for an Applied Scientist to work on the intersection of Artificial Intelligence and program analysis to raise the code quality bar in our state-of-the-art deep learning compiler stack. This stack is designed to optimize application models across diverse domains, including Large Language and Vision, originating from leading frameworks such as PyTorch and JAX. Your role will involve working closely with our custom-built Machine Learning accelerator, Trainium, which represents the forefront of innovation for advanced ML capabilities, and is the underpinning of Generative AI. In this role as an Applied Scientist, you'll be instrumental in designing, developing, and deploying analyzers for ML compiler stages and compiler IRs. You will architect and implement business-critical tooling, publish research, and mentor a brilliant team of experienced scientists and engineers. You will need to be technically capable, credible, and curious in your own right as a trusted AWS Neuron engineer, innovating on behalf of our customers. Your responsibilities will involve tackling crucial challenges alongside a talented engineering team, contributing to leading-edge design and research in compiler technology and deep-learning systems software. Strong experience in programming languages, compilers, program analyzers, theorem provers, and program synthesis engines will be a benefit in this role. A background in machine learning and AI accelerators is preferred but not required. Basic Qualifications: - PhD, or Master's degree and 4+ years of CS, CE, ML or related field experience - Experience

### Telco Consultant - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Telco-Consultant_R-055801-1
- Excerpt: Telco Consultant Tokyo posted: Posted 30+ Days Ago

### Software Development Engineer, AI/ML, AWS Neuron, Model Inference - Amazon
- Location: Cupertino, California, USA (unspecified)
- Salary: $165K-$224K
- Posted: 2025-11-19
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/3128924/software-development-engineer-ai-ml-aws-neuron-model-inference
- Excerpt: Software Development Engineer, AI/ML, AWS Neuron, Model Inference Cupertino, California, USA The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch and JAX enabling unparalleled ML inference and training performance. The Inference Enablement and Acceleration team is at the forefront of running a wide range of models and supporting novel architecture alongside maximizing their performance for AWS's custom ML accelerators. Working across the stack from PyTorch till the hardware-software boundary, our engineers build systematic infrastructure, innovate new methods and create high-performance kernels for ML functions, ensuring every compute unit is fine tuned for optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. As part of the broader Neuron organization, our team works across multiple technology layers - from frameworks and kernels and collaborate with compiler to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work

### Senior Agile Practitioner - Neural Magic
- Location: Raleigh (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Raleigh/Senior-Agile-Practitioner_R-057301
- Excerpt: Senior Agile Practitioner Raleigh posted: Posted 30+ Days Ago

### North America Public Sector (NAPS) Customer Success Executive - Intelligence Community - Neural Magic
- Location: Remote US DC (remote)
- Salary: Not disclosed
- Posted: 2026-06-08
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-US-DC/North-America-Public-Sector--NAPS--Customer-Success-Executive---Intelligence-Community_R-056059
- Excerpt: North America Public Sector (NAPS) Customer Success Executive - Intelligence Community Remote US DC posted: Posted 4 Days Ago

### Junior Solution Architect - Neural Magic
- Location: Sao Paulo (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-11
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Sao-Paulo/Junior-Solution-Architect_R-057637
- Excerpt: Junior Solution Architect Sao Paulo posted: Posted Yesterday

### Principal Software Engineer - AI Experiment Tracking (Ireland) - Neural Magic
- Location: 4 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-09
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-Ireland/RHOAI-Senior-Customer-Experience-Engineer_R-050522
- Excerpt: Principal Software Engineer - AI Experiment Tracking (Ireland) 4 Locations posted: Posted 3 Days Ago

### Agile Development Coach - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-10
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Agile-Development-Lead_R-055737-1
- Excerpt: Agile Development Coach Tokyo posted: Posted 2 Days Ago

### Machine Learning Engineer, Distributed vLLM - Neural Magic
- Location: Boston (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-27
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Boston/Machine-Learning-Engineer--Distributed-vLLM_R-057174-1
- Excerpt: Machine Learning Engineer, Distributed vLLM Boston posted: Posted 16 Days Ago

### Strategic Account manager - Neural Magic
- Location: Mumbai (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-02
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Mumbai/Strategic-Account-manager_R-053304
- Excerpt: Strategic Account manager Mumbai posted: Posted 10 Days Ago

### Staff Python / PyTorch Developer — Frontend Inference Compiler – Dubai - Cerebras Systems
- Location: Europe; Remote, California, United States; UAE (remote)
- Salary: Not disclosed
- Posted: 2025-10-31
- 401(k) match: listed (not source-backed)
- Apply: https://job-boards.greenhouse.io/cerebrassystems/jobs/7513711003
- Excerpt: Staff Python / PyTorch Developer — Frontend Inference Compiler – Dubai Europe; Remote, California, United States; UAE Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About the Role: Would you like to participate in creating the fastest Generative Models inference in the world? Join the Cerebras Inference Team to participate in development of unique Software and Hardware combination that sports best inference characteristics in the market while running largest models available. Cerebras wafer scale inference platform allows running Generative models with unprecedented speed thanks to unique hardware architecture that provides fastest access to local memory, ultra-fast interconnect and huge amount

### Account Executive - Banking - Neural Magic
- Location: Istanbul - MSO (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-11
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Istanbul---MSO/Account-Executive---Banking_R-056238-1
- Excerpt: Account Executive - Banking Istanbul - MSO posted: Posted Yesterday

### ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs - Amazon
- Location: Cupertino, California, USA (unspecified)
- Salary: $165K-$224K
- Posted: 2026-03-27
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/10378255/ml-kernel-performance-engineer-aws-neuron-annapurna-labs
- Excerpt: ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs Cupertino, California, USA The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The Acceleration Kernel Library team is at the forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions, ensuring every FLOP counts in delivering optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference and training performance. As part of the broader Neuron Compiler organization, our team works across multiple technology layers - from frameworks and compilers to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work at the intersection of machine learning, high-performance computing, and distributed architectures, where you'll help shape the future of AI acceleration technology This is an opportunity to work on cutting-edge products at the

### Ansible Consultant - Neural Magic
- Location: Brasilia - MSO (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-03
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Brasilia---MSO/Ansible-Consultant_R-055157-1
- Excerpt: Ansible Consultant Brasilia - MSO posted: Posted 9 Days Ago

### Senior Product Manager, AWS Neurosymbolic AI - Amazon
- Location: Boston, Massachusetts, USA (unspecified)
- Salary: $152K-$206K
- Posted: 2026-05-14
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/10420150/senior-product-manager-aws-neurosymbolic-ai
- Excerpt: Senior Product Manager, AWS Neurosymbolic AI Boston, Massachusetts, USA The AWS Neurosymbolic AI team is pioneering the integration of formal reasoning and neural approaches to build AI systems that are not only powerful, but provably correct. We sit at one of the most compelling frontiers in computer science: the convergence of neural networks and symbolic reasoning, where large language models meet theorem provers, and where probabilistic intelligence meets mathematical certainty. Our mission is to make AI trustworthy at scale. We develop technology that enables AI systems to reason rigorously, verify their own outputs, and provide mathematical guarantees about their behavior. This is a fundamental shift in how AI systems are built, and we believe it's on the critical path to the next generation of safe, reliable AI-powered applications. We are one of the strongest concentrations of neurosymbolic AI talent in industry. Our team includes original contributors to the Lean theorem prover and is advised by Lean's Chief Architect. We bring together researchers and engineers from both the AI and formal methods communities, a combination that is extraordinarily rare and increasingly essential. We build on Amazon's 10+ year track record of bringing automated reasoning to production at scale. AWS pioneered the use of formal methods in cloud infrastructure, from network reachability analysis to cryptographic protocol verification to access policy reasoning, systems that serve hundreds of millions of customers today. Now we're taking the next giant leap: fusing that heritage with frontier AI to make every AI system verifiable, trustworthy, and safe.

### Senior Project Manager - Neural Magic
- Location: Canberra (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Canberra/Senior-Project-Manager_R-049907-2
- Excerpt: Senior Project Manager Canberra posted: Posted 30+ Days Ago

### Software Engineer II- AI/ML, AWS Neuron - Amazon
- Location: Seattle, Washington, USA (unspecified)
- Salary: $144K-$194K
- Posted: 2026-03-24
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/10375046/software-engineer-ii-ai-ml-aws-neuron
- Excerpt: Software Engineer II- AI/ML, AWS Neuron Seattle, Washington, USA The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch and JAX enabling unparalleled ML inference and training performance. The Training Enablement and Foundation team is at the forefront of running a wide range of models and supporting novel architecture alongside maximizing their performance for AWS's custom ML accelerators. Working across the stack from PyTorch till the hardware-software boundary, our engineers build systematic infrastructure, innovate new methods and create high-performance kernels for ML functions, ensuring every compute unit is fine tuned for optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. As part of the broader Neuron organization, our team works across multiple technology layers - from frameworks and kernels and collaborate with compiler to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work at the intersection of

### Consultant, Ansible - Neural Magic
- Location: 4 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-08
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-US-NY/Consultant--Anasible_R-057562-2
- Excerpt: Consultant, Ansible 4 Locations posted: Posted 4 Days Ago

### Electrical Engineer, Compute Architecture - Neuralink
- Location: Austin, Texas, United States; South San Francisco, California, United States (unspecified)
- Salary: $105K-$176K
- Posted: 2026-05-27
- 401(k) match: listed (not source-backed)
- Apply: https://boards.greenhouse.io/neuralink/jobs/7750309003?gh_jid=7750309003
- Excerpt: Electrical Engineer, Compute Architecture Austin, Texas, United States; South San Francisco, California, United States About Neuralink: We are creating devices that enable a bi-directional interface with the brain. These devices allow us to restore movement to the paralyzed, restore sight to the blind, and revolutionize how humans interact with their digital world. Team Description: The Surgery & Robotics Hardware Team is looking for Electrical Engineers who want to advance the future of healthcare solutions with technology. At Neuralink we recognize that increasing healthcare costs, lack of access, and insufficient numbers of neurosurgeons globally require new approaches to enable greater access to our devices. To address these challenges we design fully autonomous robotics systems, human assistive surgical tools, and design entire surgical suites in order to provide better experiences for our customers. Our team works on our surgical robot system which aims to fully automate the implantation of the Neuralink implant. We design and integrate multi-axis robotic arms, perception systems including custom imaging hardware, optical coherence tomography tissue imaging, power electronics, and safety systems such as 3D spatial mapping/object avoidance sensors. Robotics is part of a larger system of tools we work on. For example, using pre-op MRI scans we generate patient anatomy data that we later use in surgery to register precise surgery site locations. In addition to research and development of modern surgical technology, we also are responsible for bringing these technologies to market. We perform full lifecycle testing for all devices we design and ultimately manufacture them for

### Sales Account Manager - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Sales-Account-Manager_R-053936-1
- Excerpt: Sales Account Manager Tokyo posted: Posted 30+ Days Ago

### Senior Software Development Engineer, AI/ML, AWS Neuron, Model Inference - Amazon
- Location: Cupertino, California, USA (unspecified)
- Salary: $193K-$262K
- Posted: 2025-10-01
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/3098687/senior-software-development-engineer-ai-ml-aws-neuron-model-inference
- Excerpt: Senior Software Development Engineer, AI/ML, AWS Neuron, Model Inference Cupertino, California, USA The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch and JAX enabling unparalleled ML inference and training performance. The Inference Enablement and Acceleration team is at the forefront of running a wide range of models and supporting novel architecture alongside maximizing their performance for AWS's custom ML accelerators. Working across the stack from PyTorch till the hardware-software boundary, our engineers build systematic infrastructure, innovate new methods and create high-performance kernels for ML functions, ensuring every compute unit is fine tuned for optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. As part of the broader Neuron organization, our team works across multiple technology layers - from frameworks and kernels and collaborate with compiler to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to

### Software Quality Engineer - Neural Magic
- Location: Raleigh (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-21
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Raleigh/Software-Quality-Engineer_R-057367
- Excerpt: Software Quality Engineer Raleigh posted: Posted 22 Days Ago

### LLM Inference Performance & Evals Engineer - Cerebras Systems
- Location: Toronto, Ontario, Canada (unspecified)
- Salary: Not disclosed
- Posted: 2025-07-24
- 401(k) match: listed (not source-backed)
- Apply: https://job-boards.greenhouse.io/cerebrassystems/jobs/6658665003
- Excerpt: LLM Inference Performance & Evals Engineer Toronto, Ontario, Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Role Join the inference model team dedicated to bring up the state-of-the-art models, numerically validating and accelerating new model ideas on wafer-scale hardware. You will prototype architectural tweaks, build performance-eval pipelines, and turn hard numbers into changes that land in production. Key Responsibilities - Prototype and benchmark cutting-edge ideas: new attentions, MoE, speculative decoding, and many more innovations as they emerge. - Develop agent-driven automation that designs experiments, schedules runs, triages regressions, and drafts pull-requests. - Work closely with compiler, runtime,

### Consultant, Ansible - Neural Magic
- Location: 4 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-07
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-US-DC/Consultant--Ansible_R-057335-1
- Excerpt: Consultant, Ansible 4 Locations posted: Posted 5 Days Ago

### Consultant - Strasbourg - Neural Magic
- Location: Remote France (remote)
- Salary: Not disclosed
- Posted: 2026-06-02
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-France/Consultant---Strasbourg_R-055946
- Excerpt: Consultant - Strasbourg Remote France posted: Posted 10 Days Ago

### Business Analyst Intern (m/f/d) - Neural Magic
- Location: Munich (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-11
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Munich/Business-Analyst-Intern--m-f-d-_R-050330
- Excerpt: Business Analyst Intern (m/f/d) Munich posted: Posted Yesterday

### Software Engineer – AI Context Engineering - Neural Magic
- Location: 2 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-05
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Pune/Software-Engineer---AI-Context-Engineering_R-056995-2
- Excerpt: Software Engineer – AI Context Engineering 2 Locations posted: Posted 7 Days Ago

### Electrical Engineer, Power Systems - Neuralink
- Location: Austin, Texas, United States; South San Francisco, California, United States (unspecified)
- Salary: $105K-$176K
- Posted: 2026-02-23
- 401(k) match: listed (not source-backed)
- Apply: https://boards.greenhouse.io/neuralink/jobs/7641728003?gh_jid=7641728003
- Excerpt: Electrical Engineer, Power Systems Austin, Texas, United States; South San Francisco, California, United States About Neuralink: We are creating devices that enable a bi-directional interface with the brain. These devices allow us to restore movement to the paralyzed, restore sight to the blind, and revolutionize how humans interact with their digital world. Team Description: The Surgery & Robotics Hardware Team is looking for Electrical Engineers who want to advance the future of healthcare solutions with technology. At Neuralink we recognize that increasing healthcare costs, lack of access, and insufficient numbers of neurosurgeons globally require new approaches to enable greater access to our devices. To address these challenges we design fully autonomous robotics systems, human assistive surgical tools, and design entire surgical suites in order to provide better experiences for our customers. Our team works on our surgical robot system which aims to fully automate the implantation of the Neuralink implant. We design and integrate multi-axis robotic arms, perception systems including custom imaging hardware, optical coherence tomography tissue imaging, power electronics, and safety systems such as 3D spatial mapping/object avoidance sensors. Robotics is part of a larger system of tools we work on. For example, using pre-op MRI scans we generate patient anatomy data that we later use in surgery to register precise surgery site locations. In addition to research and development of modern surgical technology, we also are responsible for bringing these technologies to market. We perform full lifecycle testing for all devices we design and ultimately manufacture them for

### Business Value Director - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Business-Value-Director_R-053794-1
- Excerpt: Business Value Director Tokyo posted: Posted 30+ Days Ago

### Principal Software Engineer - Neural Magic
- Location: Pune (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-25
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Pune/Principal-Software-Engineer_R-055144
- Excerpt: Principal Software Engineer Pune posted: Posted 18 Days Ago

### Principal Product Manager, AWS Neurosymbolic AI - Amazon
- Location: Seattle, Washington, USA (unspecified)
- Salary: $181K-$245K
- Posted: 2026-05-13
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/10419273/principal-product-manager-aws-neurosymbolic-ai
- Excerpt: Principal Product Manager, AWS Neurosymbolic AI Seattle, Washington, USA The AWS Neurosymbolic AI team is pioneering the integration of formal reasoning and neural approaches to build AI systems that are not only powerful, but provably correct. We sit at one of the most compelling frontiers in computer science: the convergence of neural networks and symbolic reasoning, where large language models meet theorem provers, and where probabilistic intelligence meets mathematical certainty. Our mission is to make AI trustworthy at scale. We develop technology that enables AI systems to reason rigorously, verify their own outputs, and provide mathematical guarantees about their behavior. This is a fundamental shift in how AI systems are built, and we believe it's on the critical path to the next generation of safe, reliable AI-powered applications. We are one of the strongest concentrations of neurosymbolic AI talent in industry. Our team includes original contributors to the Lean theorem prover and is advised by Lean's Chief Architect. We bring together researchers and engineers from both the AI and formal methods communities, a combination that is extraordinarily rare and increasingly essential. We build on Amazon's 10+ year track record of bringing automated reasoning to production at scale. AWS pioneered the use of formal methods in cloud infrastructure, from network reachability analysis to cryptographic protocol verification to access policy reasoning, systems that serve hundreds of millions of customers today. Now we're taking the next giant leap: fusing that heritage with frontier AI to make every AI system verifiable, trustworthy, and safe.

### Neuroengineer, Next Gen - Neuralink
- Location: Fremont, California, United States (unspecified)
- Salary: $122K-$226K
- Posted: 2025-12-23
- 401(k) match: listed (not source-backed)
- Apply: https://boards.greenhouse.io/neuralink/jobs/7571489003?gh_jid=7571489003
- Excerpt: Neuroengineer, Next Gen Fremont, California, United States About Neuralink: We are creating devices that enable a bi-directional interface with the brain. These devices allow us to restore movement to the paralyzed, restore sight to the blind, and revolutionize how humans interact with their digital world. Team Description: The Next Gen team at Neuralink is developing the next generation of brain-computer interfaces. We are laying the groundwork for intuitive, high-dimensional, and bidirectional interfaces between brains and machines, with the goal of helping people challenged by a variety of neurological disorders and conditions. Our team consists of scientists and engineers, working closely together to define the engineering requirements for these future products. As a Neuroengineer on this team, you will contribute across a wide range of projects, from designing novel experimental preparations, to interpreting neural signals and behavioral data, to developing novel BCI paradigms. Successful candidates will be highly adaptable, able to deploy their core technical and creative skills to tackle a wide range of problems, and have a keen sense of urgency. Job Description and Responsibilities: - Develop computational models of, and design encoding strategies for, electrical stimulation - Design, execute, interpret, and communicate the results of neural recording and stimulation experiments - Implement end-to-end hardware and software solutions for prosthetic vision, including machine vision algorithms, smart glasses, and eye tracking technology - Run and support research sessions in coordination with software engineers, field engineers, and animal specialists - Present results in a collaborative setting Required Qualifications: - Evidence of exceptional

### Senior Middleware Consultant - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Senior-Middleware-Consultant_R-052765-1
- Excerpt: Senior Middleware Consultant Tokyo posted: Posted 30+ Days Ago

### Agile Development Coach - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-22
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Agile-Development-Coach_R-057529-2
- Excerpt: Agile Development Coach Tokyo posted: Posted 21 Days Ago

### Machine Learning Engineer - Perception - Zoox
- Location: Foster City, CA (unspecified)
- Salary: Not disclosed
- Posted: 2025-07-09
- Apply: https://jobs.lever.co/zoox/50b011cc-1023-4d9a-b8be-c938e2c12f13
- Excerpt: Machine Learning Engineer - Perception Foster City, CA The Perception team at Zoox is at the forefront of leveraging GenAI to create synthetic data, unlocking scalable training and evaluation for our autonomous system's perception and entire stack. As a Generative AI Engineer, you will develop and train cutting-edge models for sensor-level scenario generation, utilizing world models and radiance fields techniques with large-scale proprietary data. This role directly impacts the productivity, safety, and capabilities of Zoox's autonomous system by validating algorithms in real-world conditions. The Perception team at Zoox is at the forefront of leveraging GenAI to create synthetic data, unlocking scalable training and evaluation for our autonomous system's perception and entire stack. As a Generative AI Engineer, you will develop and train cutting-edge models for sensor-level scenario generation, utilizing world models and radiance fields techniques with large-scale proprietary data. This role directly impacts the productivity, safety, and capabilities of Zoox's autonomous system by validating algorithms in real-world conditions. In this role, you will: - Design, develop, train and evaluate multi-sensor fusion based deep learning models to understand obstacles and environmental context - Understand and curate real and synthetic datasets to improve our models - Perform latency optimization and deploy models to our robot fleet - Build a deep understanding of Perception gaps and behavioral issues around difficult obstacle types in order to help plan and prioritize our work - Collaborate with Prediction/Planner team to deploy fully autonomous vehicles in environments with difficult and rare obstacles, extreme weather conditions, and complex

### Senior Deep Learning Engineer (Perception) - Seneca Foods CORP
- Location: Sausalito (unspecified)
- Salary: $160K-$220K
- Posted: 2025-10-19
- 401(k) match: listed (not source-backed)
- Apply: https://jobs.ashbyhq.com/seneca/202dc06f-5aff-4998-9cda-b3790f0ac4de
- Excerpt: Senior Deep Learning Engineer (Perception) Sausalito The Job: We are seeking an exceptional and highly motivated Deep Learning Engineer to join our fast-moving and innovative team. In this role, you will leverage your expertise to research, develop, and deploy cutting-edge perception algorithms for our next-generation autonomous systems. You will play a pivotal part in building the "eyes" of our technology, enabling it to robustly perceive and understand the surrounding world. This is a hands-on, systems-level role where you will bridge fundamental research and real-world deployment, tackling complex challenges. As a key early member of our technical team, you will have significant ownership and the opportunity to directly shape our product's perception capabilities. What You'll Do: Algorithm development: Research, design, and implement advanced deep learning algorithms for core perception tasks, such as 2D/3D object detection, semantic segmentation, tracking, and classification. Multimodal sensor fusion: Develop and optimize algorithms to fuse data from multiple sensor modalities, including cameras, LiDAR, and radar, to build a robust and comprehensive perception system. Data-centric development: Take a rigorous, data-driven approach to model development. This includes designing training and validation pipelines, curating large-scale datasets, and prioritizing data collection and labeling efforts. System integration: Work cross-functionally with embedded, hardware, and systems engineers to seamlessly integrate perception software into the larger autonomous stack. Performance analysis and debugging: Analyze logged field data to identify performance bottlenecks and edge cases, and rapidly iterate on solutions to improve model accuracy and robustness in challenging real-world conditions. Stay current on research: Keep abreast of

### Consultant - Middleware - Neural Magic
- Location: Remote Malaysia (remote)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-Malaysia/Consultant---Middleware_R-055074-1
- Excerpt: Consultant - Middleware Remote Malaysia posted: Posted 30+ Days Ago

### Software Development Engineer AI/ML, Inference Serving, AWS Neuron - Amazon
- Location: Cupertino, California, USA (unspecified)
- Salary: $193K-$262K
- Posted: 2025-09-19
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/3089270/software-development-engineer-ai-ml-inference-serving-aws-neuron
- Excerpt: Software Development Engineer AI/ML, Inference Serving, AWS Neuron Cupertino, California, USA AWS Neuron is the software stack powering AWS Inferentia and Trainium machine learning accelerators, designed to deliver high-performance, low-cost inference at scale. The Neuron Serving team develops infrastructure to serve modern machine learning models-including large language models (LLMs) and multimodal workloads-reliably and efficiently on AWS silicon. We are seeking a Software Development Engineer to lead and architect our next-generation model serving infrastructure, with a particular focus on large-scale generative AI applications. Key job responsibilities * Architect and lead the design of distributed ML serving systems optimized for generative AI workloads * Drive technical excellence in performance optimization and system reliability across the Neuron ecosystem * Design and implement scalable solutions for both offline and online inference workloads * Lead integration efforts with frameworks such as vLLM, SGLang, Torch XLA, TensorRT, and Triton * Develop and optimize system components for tensor/data parallelism and disaggregated serving * Implement and optimize custom PyTorch operators and NKI kernels * Mentor team members and provide technical leadership across multiple work streams * Drive architectural decisions that impact the entire Neuron serving stack * Collaborate with customers, product owners, and engineering teams to define technical strategy * Author technical documentation, design proposals, and architectural guidelines A day in the life You'll lead critical technical initiatives while mentoring team members. You'll collaborate with cross-functional teams of applied scientists, system engineers, and product managers to architect and deliver state-of-the-art inference capabilities. Your day might involve: * Leading

### Specialist Solution Architect - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-03
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Specialist-Solution-Architect_R-051518-2
- Excerpt: Specialist Solution Architect Tokyo posted: Posted 9 Days Ago

### Staff Inference ML Runtime Engineer - Cerebras Systems
- Location: Sunnyvale CA or Toronto Canada (unspecified)
- Salary: Not disclosed
- Posted: 2025-11-25
- 401(k) match: listed (not source-backed)
- Apply: https://job-boards.greenhouse.io/cerebrassystems/jobs/7523546003
- Excerpt: Staff Inference ML Runtime Engineer Sunnyvale CA or Toronto Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Role The Inference ML Engineering team at Cerebras Systems is dedicated to enabling our fast generative inference solution through simple APIs powered by a distributed runtime that runs on large clusters of our own hardware. Our mission is to empower enterprises, developers, and researchers to unlock the full potential of our platform, leveraging its performance, scalability, and flexibility. The team works closely with cross-functional groups, including compiler developers, cluster orchestrators, ML scientists, cloud architects, and product teams, to deliver high-impact

### Senior Consultant - Ansible - Neural Magic
- Location: Singapore (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Singapore/Senior-Consultant---Ansible_R-053753-1
- Excerpt: Senior Consultant - Ansible Singapore posted: Posted 30+ Days Ago

### Staff Kernel Optimzation Engineer - Cerebras Systems
- Location: Remote, California, United States (remote)
- Salary: Not disclosed
- Posted: 2026-02-05
- 401(k) match: listed (not source-backed)
- Apply: https://job-boards.greenhouse.io/cerebrassystems/jobs/7620254003
- Excerpt: Staff Kernel Optimzation Engineer Remote, California, United States Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. About The Role As a Kernel Engineer on our team, you will develop high-performance software solutions at the intersection of hardware and software, developing high-performance software for cutting-edge AI and HPC workloads. Your focus will be on implementing, optimizing, and scaling deep learning operations to fully leverage our custom, massively parallel processor architecture. You will be part of a world-class team responsible for the design, performance tuning, and validation of foundational ML and HPC kernels. This includes building a library of parallel and distributed

### Principal Data Scientist - Neural Magic
- Location: 2 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Bangalore---Carina/Principal-Data-Scientist_R-048286-1
- Excerpt: Principal Data Scientist 2 Locations posted: Posted Today

### Executive Policy Engagement Lead, DeepMind - DeepMind
- Location: London, UK (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-29
- Parental leave: 18 weeks (not source-backed)
- Non-birth-parent leave: 18 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Apply: https://www.google.com/about/careers/applications/signin?jobId=CiUAL2FckVxvMY3uqanu9SOD6Deq2TEevdb3KGvWewG65rnof2OKEjoACxwdTA0VtcFhp5zH6C7kUB6wAAqNLhPvBIAXOyp92zaFdhjk9gbGLcUkfMDVw4J690QIM_IuIAq-_V2&loc=GB&title=Executive+Policy+Engagement+Lead
- Excerpt: Executive Policy Engagement Lead, DeepMind London, UK In this role, you will be an experienced leader to co-design and scale DeepMind's Executive Policy Engagement function. You will help to set and drive the strategy for how DeepMind's executives build trust with governments and policymakers. You will own long-term engagement plans, brief and staff our executives for high-stakes events, and help to build an executive policy engagement function. You are an autonomous leader with advisory experience, operating expertly at the intersection of policy, communications, and executive management. You will collaborate closely and build trust with leaders in partner teams, including Communications and multiple executive offices. Artificial intelligence will be one of humanity's most transformative inventions. At DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority. We are pushing the boundaries across multiple domains. Our global teams offer learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort. Partner with leadership to build the executive policy engagement function. Own, design, and execute bespoke, long-term engagement strategies for our growing executive portfolio. Serve as the primary on-the-ground counsel for executives during international travel, providing real-time political insights and navigating stakeholder dynamics at major global forums. Draft researched briefing materials, speech content, and policy narratives to ensure

### Senior Architect - Neural Magic
- Location: Remote Malaysia (remote)
- Salary: Not disclosed
- Posted: 2026-05-26
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-Malaysia/Senior-Architect_R-057075-1
- Excerpt: Senior Architect Remote Malaysia posted: Posted 17 Days Ago

### Senior Principal Machine Learning Engineer, vLLM - Neural Magic
- Location: Boston (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-19
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Boston/Senior-Principal-Machine-Learning-Engineer_R-053063-1
- Excerpt: Senior Principal Machine Learning Engineer, vLLM Boston posted: Posted 24 Days Ago

### Software Engineer, Telco - Neural Magic
- Location: Raleigh (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-18
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Raleigh/Software-Engineer--Telco_R-055569
- Excerpt: Software Engineer, Telco Raleigh posted: Posted 25 Days Ago

### Junior Solution Architect- Openshift - Neural Magic
- Location: Singapore (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-11
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Singapore/Junior-Solution-Architect--Openshift_R-054546
- Excerpt: Junior Solution Architect- Openshift Singapore posted: Posted Yesterday

### AI Solution Specialist, ANZ (Sydney/Melbourne-based) - Neural Magic
- Location: 3 Locations (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-30
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-Australia/AI-Solution-Specialist--ANZ--Sydney-Melbourne-based-_R-057115
- Excerpt: AI Solution Specialist, ANZ (Sydney/Melbourne-based) 3 Locations posted: Posted 13 Days Ago

### Sr. SDM, AI Inference Technology, Neuron SDK - Amazon
- Location: Seattle, Washington, USA (unspecified)
- Salary: $253K-$342K
- Posted: 2025-06-16
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/3009701/sr-sdm-ai-inference-technology-neuron-sdk
- Excerpt: Sr. SDM, AI Inference Technology, Neuron SDK Seattle, Washington, USA AWS Utility Computing (UC) provides product innovations - from foundational services such as Amazon Elastic Compute Cloud (EC2), to new product innovations that continue to set AWS's services and features apart in the industry. Come develop inference acceleration for AWS Neuron, the complete software stack for Trainium, Amazon's custom cloud-scale machine learning accelerators that power the latest AI models As the Sr. SDM for the Inference Technology Team, you will lead a strong team of managers and engineers to build fundamental inference technology building blocks and libraries to enable AI developers to optimize model for inference on Trainium and Inferentia devices. You will be responsible for the full development life cycle of inference library and feature development, including reliability and scalability. You will develop the Neuronx_Distributed Inference Libraries and contribute to other popular open source Inference Libraries, enabling customers to optimize LLMs, multimodal, and generative models. The ideal candidate will have an established background in delivering AI feature support for demanding, fast-changing priorities or delivering high-performance models using distributed inference libraries. The ideal candidate should have a strong technical ability to understand and manage a vertically integrated system stack that consisting of hardware, frameworks, and workflows. A day in the life You will work with the executive leadership and other senior management and technical leaders to define product directions and deliver them to customers. We build massive-scale distributed training and inference solutions, developing the full stack of software, servers and

### Agile Development Coach - Neural Magic
- Location: Tokyo (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-18
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Tokyo/Agile-Development-Coach_R-052763-1
- Excerpt: Agile Development Coach Tokyo posted: Posted 25 Days Ago

### Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs - Amazon
- Location: Cupertino, California, USA (unspecified)
- Salary: $193K-$262K
- Posted: 2025-08-14
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/3059992/sr-ml-kernel-performance-engineer-aws-neuron-annapurna-labs
- Excerpt: Sr. ML Kernel Performance Engineer, AWS Neuron, Annapurna Labs Cupertino, California, USA The Annapurna Labs team at Amazon Web Services (AWS) builds AWS Neuron, the software development kit used to accelerate deep learning and GenAI workloads on Amazon's custom machine learning accelerators, Inferentia and Trainium. The Acceleration Kernel Library team is at the forefront of maximizing performance for AWS's custom ML accelerators. Working at the hardware-software boundary, our engineers craft high-performance kernels for ML functions, ensuring every FLOP counts in delivering optimal performance for our customers' demanding workloads. We combine deep hardware knowledge with ML expertise to push the boundaries of what's possible in AI acceleration. The AWS Neuron SDK, developed by the Annapurna Labs team at AWS, is the backbone for accelerating deep learning and GenAI workloads on Amazon's Inferentia and Trainium ML accelerators. This comprehensive toolkit includes an ML compiler, runtime, and application framework that seamlessly integrates with popular ML frameworks like PyTorch, enabling unparalleled ML inference and training performance. As part of the broader Neuron Compiler organization, our team works across multiple technology layers - from frameworks and compilers to runtime and collectives. We not only optimize current performance but also contribute to future architecture designs, working closely with customers to enable their models and ensure optimal performance. This role offers a unique opportunity to work at the intersection of machine learning, high-performance computing, and distributed architectures, where you'll help shape the future of AI acceleration technology This is an opportunity to work on cutting-edge products at

### Cloud Architect - Neural Magic
- Location: Mexico City (unspecified)
- Salary: Not disclosed
- Posted: 2026-06-07
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Mexico-City/Cloud-Architect_R-054214-1
- Excerpt: Cloud Architect Mexico City posted: Posted 5 Days Ago

### Director, Solution Architecture - Neural Magic
- Location: Kuala Lumpur (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-13
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Kuala-Lumpur/Director--Solution-Architecture_R-049624-1
- Excerpt: Director, Solution Architecture Kuala Lumpur posted: Posted 30 Days Ago

### Machine Learning: Whole-Body Control - The Bot Company
- Location: San Francisco (unspecified)
- Salary: Not disclosed
- Posted: 2026-02-25
- Apply: https://jobs.ashbyhq.com/thebotcompany/69c93a4a-161a-49c9-89bc-4c3e0c4716e6
- Excerpt: Machine Learning: Whole-Body Control San Francisco The Bot Company We're building a helpful robot for every home. We're a small team of engineers, designers, and operators based in San Francisco. Our team comes from Tesla, Cruise, OpenAI, Google, Pixar, and many other great companies. In the past we've shipped to hundreds of millions of users and know what it takes to build amazing products and experiences. Our team is deliberately lean to promote rapid decision making and do away with bureaucracy and hierarchy. Everyone is an IC and is empowered with massive scope, radical ownership, and direct responsibility. We work across the stack with a culture built for rapid iteration and fast execution. What we look for in all candidates All roles at The Bot Company demand extreme sharpness and the ability to move fast in high-intensity environments. Throughout the process, we expect candidates to demonstrate: - Exceptional mental acuity: you think quickly, learn instantly, and reason across unfamiliar domains. - Engineering curiosity: you naturally dig into how systems work, even outside your specialty. - High performance mindset: you move fast, handle ambiguity, and excel when the environment is demanding. Machine Learning: Whole-Body Control We are building high-performance whole-body controllers that produce robust, agile motion and manipulation in the real world. You will train low-level control policies in simulation and own the stack from environment design to large-scale training and sim-to-real deployment. What You'll Do - Train whole-body policies for locomotion, manipulation, and coordinated motion. - Build scalable simulation environments

### Senior Resource Deployment Specialist - Neural Magic
- Location: Melbourne (unspecified)
- Salary: Not disclosed
- Posted: 2026-05-22
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Melbourne/Senior-Resource-Deployment-Specialist_R-057466
- Excerpt: Senior Resource Deployment Specialist Melbourne posted: Posted 21 Days Ago

### Research Scientist, Robotics, DeepMind - DeepMind
- Location: Cambridge, MA, USA (unspecified)
- Salary: $147K-$211K
- Posted: 2026-06-12
- Parental leave: 18 weeks (not source-backed)
- Non-birth-parent leave: 18 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Apply: https://www.google.com/about/careers/applications/signin?jobId=CiUAL2FckQx_i4WU6SWHgKivWRKzYfFCgj8ZZIPOO_y1cUTWJgOxEjsACxwdTP5fMhuSb6SqeNplEvnfwqezB3rl4Dhvc4cTAJgTwpznNnctmCpHXAo5VDBnfbxcGLIiPInmpg%3D%3D_V2&loc=US&title=Research+Scientist
- Excerpt: Research Scientist, Robotics, DeepMind Cambridge, MA, USA We believe there are many problems in the world in which robotics could play a significant role in making it easier, faster and safer for people to get things done. We're looking for roboticists, designers, hardware and software engineers to help us explore these possibilities, develop breakthrough technologies, and build new products that could help millions of people. At Google DeepMind Robotics, we are pioneering the integration of AI into the physical world to power an era of physical agents. By enabling robots to perceive, plan, think, use tools, and act, we empower them to solve increasingly tasks. In this role, you will focus on building foundation models, such as advanced vision-language-action (VLA) models, which combine Gemini's world understanding with physical actions to provide direct robotic control. These include Gemini Robotics, our most advanced Gemini model for the physical world, and Gemini Robotics On-Device, our fastest Gemini model that functions without a data network. These models allow robots to perform a broad range of tasks, respond interactively to their environment, achieve high dexterity, and reason over long, multi-step sequences. Beyond model building, we are committed to advancing general-purpose robotics, specifically in areas like agentic reasoning, real-world understanding, action generalization, human-robot interaction, dexterity, whole-body control, and continual learning. To deploy these innovations at scale, you will partner with key robotics companies to bring this intelligence to the physical world across a broad array of applications. Artificial intelligence will be one of humanity's most transformative

### Principal Engineer, AI Inference Reliability - Cerebras Systems
- Location: Remote, California, United States; Sunnyvale CA or Toronto Canada (remote)
- Salary: Not disclosed
- Posted: 2025-10-29
- 401(k) match: listed (not source-backed)
- Apply: https://job-boards.greenhouse.io/cerebrassystems/jobs/7511484003
- Excerpt: Principal Engineer, AI Inference Reliability Remote, California, United States; Sunnyvale CA or Toronto Canada Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs. Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation. In late 2024, we launched Cerebras Inference, the fastest Generative AI inference service in the world, over 10 times faster than GPU-based hyperscale cloud inference. Since launch, we've scaled to meet the surging demand from AI labs, enterprises, and a thriving developer community. In October 2025, we announced our series G funding, raising $1.1 billion USD to accelerate the expansion of our products and services to meet global AI demand. About the team The Cerebras Inference team's mission

### Technical Translator, Korean Language - Neural Magic
- Location: Remote Australia (remote)
- Salary: Not disclosed
- Posted: 2026-06-12
- Apply: https://redhat.wd5.myworkdayjobs.com/jobs/job/Remote-Australia/Technical-Translator--Korean-Language_R-057938-1
- Excerpt: Technical Translator, Korean Language Remote Australia posted: Posted Today

### Principal Applied Scientist, Neuro-Symbolic AI Labs - Amazon
- Location: Boston, Massachusetts, USA (unspecified)
- Salary: $199K-$269K
- Posted: 2026-04-27
- Parental leave: 6 weeks (not source-backed)
- Non-birth-parent leave: 6 weeks (not source-backed)
- 401(k) match: listed (not source-backed)
- Fertility benefits: yes (not source-backed)
- Adoption assistance: yes (not source-backed)
- Mental health support: yes (not source-backed)
- Childcare support: yes (not source-backed)
- Apply: https://www.amazon.jobs/en/jobs/10404529/principal-applied-scientist-neuro-symbolic-ai-labs
- Excerpt: Principal Applied Scientist, Neuro-Symbolic AI Labs Boston, Massachusetts, USA We are looking for researchers who aim to build super-intelligent AI systems that leverage proof assistants to guide learning and reasoning. Our neuro-symbolic AI technology is applied across a wide range of science and engineering domains within Amazon, and you will join the team at the forefront of this research. As a Principal Applied Scientist, you will play a pivotal role in shaping the definition, vision, and development of product features from beginning to end. You will: - Define and implement new neuro-symbolic applications that employ scalable and efficient approaches to solve complex problems. - Work in an agile, startup-like development environment, where you are always working on the most important stuff. - Deliver high-quality scientific artifacts. About the team We work closely with academia. Our team includes an Amazon Scholar in mathematics, and we maintain active research collaborations with faculty at leading CS departments (MIT, Berkeley, CMU). Basic Qualifications: - Experience programming in Java, C++, Python or related language - Experience with leading experienced scientists as well as having a record of developing junior members from academia or industry to a career track in a business environment - PhD, or Master's degree and 8+ years of applied research experience Preferred Qualifications: - 10+ years of relevant work in industry or academia experience - Experience creating novel algorithms and advancing the state of the art - Have peer-reviewed scientific contributions in premier journals and conferences Amazon is an equal opportunity employer

### Deep Learning Engineer - Carbon Robotics
- Location: Seattle, WA (unspecified)
- Salary: $140K-$220K
- Posted: 2026-04-17
- Mental health support: yes (not source-backed)
- Apply: https://carbonrobotics.com/job-openings?gh_jid=4673637006
- Excerpt: Deep Learning Engineer Seattle, WA The Carbon Robotics LaserWeeder™ leverages advanced robotics, computer vision, AI/deep learning, and lasers to eliminate weeds with sub-millimeter accuracy-all without herbicides. This innovative solution reduces environmental impact, promotes soil health, and helps farmers address labor shortages and rising costs. Designed in Seattle and built at our cutting-edge manufacturing facility in Richland, Washington, the LaserWeeder is setting a new standard for automated weed control. With $157 million in funding from prominent investors such as BOND, NVentures (NVIDIA's venture arm), Anthos Capital, Fuse Venture Capital, Ignition Partners, Revolution, Sozo Ventures, and Voyager Capital , Carbon Robotics is driving innovation. As a no-nonsense team with a bias for action, we take pride in executing our ideas. Whether it's designing transformative technology or visiting farms to ensure our products are reliable and safe, we do whatever it takes to deliver for our customers. Working here means tackling big problems with big impact. You'll find opportunities to grow professionally, solve complex challenges, and make meaningful contributions to a mission that matters. At Carbon Robotics, we trust our team to act independently and make practical, real-world decisions. Join us as we innovate, execute, and build the future of farming together. YouTube | X | Instagram | LinkedIn | News Deep Learning Engineer As a Deep Learning Engineer at Carbon Robotics, you will contribute to designing, developing, and deploying novel deep learning systems that power our autonomous laser weeding robots in the field. What You'll Do - Lead the design and execution

---
Source-backed benefit claims include source links; other benefit values are labeled separately.