Demo

Machine Learning Engineer, AWS Neuron Inference, AWS Neuron

Amazon Web Services (AWS)
Seattle, WA Full Time
POSTED ON 12/26/2025
AVAILABLE BEFORE 4/9/2026
Description

AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine

learning accelerators and the Trn1 and Inf1 servers that use them. This role is for a software engineer in the Machine Learning Applications (ML Apps) team for AWS Neuron.

This role is responsible for development, enablement and performance tuning of a wide variety of ML model families, including massive scale large language models like Llama2, GPT2, GPT3 and beyond, as well as stable diffusion, Vision Transformers and many more.

The ML Apps team works side by side with compiler engineers and runtime engineers to create, build and tune distributed inference solutions with Trn1. Experience optimizing inference performance for both latency and throughput on these large models using Python, Pytorch or JAX is a must. Deepspeed and other distributed inference libraries are central to this and extending all of this for the Neuron based system is key.

Key job responsibilities

This role will help lead the efforts building distributed inference support into Pytorch, Tensorflow using XLA and the Neuron compiler and runtime stacks. This role will help tune these models to ensure highest performance and maximize the efficiency of them running on the customer AWS Trainium and Inferentia silicon and the TRn1 , Trn2 servers. Strong software development using Python/C and ML knowledge are both critical to this role.

A day in the life

As You Design And Code Solutions To Help Our Team Drive Efficiencies In Software Architecture, You’ll Create Metrics, Implement Automation And Other Improvements, And Resolve The Root Cause Of Software Defects. You’ll Also

Build high-impact solutions to deliver to our large customer base.

Participate in design discussions, code review, and communicate with internal and external stakeholders.

Work cross-functionally to help drive business decisions with your technical input.

Work in a startup-like development environment, where you’re always working on the most important stuff.

About The Team

Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we’re building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future.

Basic Qualifications

  • 3 years of non-internship professional software development experience
  • 2 years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience
  • Experience programming with at least one software programming language

Preferred Qualifications

  • 3 years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience
  • Bachelor's degree in computer science or equivalent

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

Our compensation reflects the cost of labor across several US geographic markets. The base pay for this position ranges from $129,300/year in our lowest geographic market up to $223,600/year in our highest geographic market. Pay is based on a number of factors including market location and may vary depending on job-related knowledge, skills, and experience. Amazon is a total compensation company. Dependent on the position offered, equity, sign-on payments, and other forms of compensation may be provided as part of a total compensation package, in addition to a full range of medical, financial, and/or other benefits. For more information, please visit https://www.aboutamazon.com/workplace/employee-benefits. This position will remain posted until filled. Applicants should apply via our internal or external career site.


Company - Annapurna Labs (U.S.) Inc.

Job ID: A3008876

Salary : $129,300 - $223,600

If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Machine Learning Engineer, AWS Neuron Inference, AWS Neuron?

Sign up to receive alerts about other jobs on the Machine Learning Engineer, AWS Neuron Inference, AWS Neuron career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$97,257 - $120,701
Income Estimation: 
$123,167 - $152,295
Income Estimation: 
$101,387 - $124,118
Income Estimation: 
$119,030 - $151,900
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Amazon Web Services (AWS)

  • Amazon Web Services (AWS) Sparks, NV
  • Description AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we’re the peop... more
  • 13 Days Ago

  • Amazon Web Services (AWS) Canton, MS
  • Description Join our dynamic AWS team and become a critical guardian of global cloud infrastructure! You'll play a pivotal role in maintaining the heartbea... more
  • 13 Days Ago

  • Amazon Web Services (AWS) Canton, MS
  • Description The Data Center Global Controls team is looking for exceptional individuals to join our Controls organization as a Controls Technician for Serv... more
  • 13 Days Ago

  • Amazon Web Services (AWS) Umatilla, OR
  • Description AWS Infrastructure Services owns the design, planning, delivery, and operation of all AWS global infrastructure. In other words, we’re the peop... more
  • 13 Days Ago


Not the job you're looking for? Here are some other Machine Learning Engineer, AWS Neuron Inference, AWS Neuron jobs in the Seattle, WA area that may be a better fit.

  • Amazon Web Services (AWS) Seattle, WA
  • Description AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the Trn2 and future... more
  • 20 Days Ago

  • Amazon Seattle, WA
  • Description AWS Neuron is the complete software stack for the AWS Inferentia and Trainium cloud-scale machine learning accelerators and the Trn1 and Inf1 s... more
  • 7 Days Ago

AI Assistant is available now!

Feel free to start your new journey!