1,092,897open jobs
63,657companies
186,338added this week
Browse all
Salary
$165k – $224k per year
Location
In office (Cupertino)
Seniority
Middle · 3+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 2, 2026. First seen by Alion on Sep 17, 2026. Amazon scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Amazon is an American technology and retail conglomerate founded by Jeff Bezos in 1994 as an online bookstore and headquartered in Seattle, Washington. It operates the world's largest online marketplace together with a global logistics network, physical grocery stores and a third-party seller platform that accounts for most units sold. Amazon Web Services, launched in 2006, is the leading public cloud provider and generates the majority of the group's operating profit, while advertising, Prime Video, Alexa devices and Kuiper satellite broadband round out the business.

Do you want to be part of AI revolution? At Amazon our vision is to make deep learning pervasive for everyday developers and to democratize access to AI hardware and software infrastructure. In order to deliver on that vision, we’ve created innovative software and hardware solutions that make it possible. AWS Neuron is the SDK that optimizes the performance of complex ML models executed on AWS Inferentia and Trainium, our custom chips designed to accelerate deep-learning workloads.

This role is for a software engineer in the Compiler team for AWS Neuron. As part of this role, you will be responsible for building next generation Neuron compiler which transforms ML models written in ML frameworks (e.g, PyTorch, TensorFlow, and JAX) to be deployed AWS Inferentia and Trainium based servers in the Amazon cloud. You will be responsible for solving hard compiler optimization problems to achieve optimum performance for variety of ML model families including massive scale large language models like Llama, Deepseek, and beyond as well as stable diffusion, vision transformers and multi-model models. You will be required to understand how these models work inside-out to make informed decisions on how to best coax the compiler to generate optimal implementation instruction. You will leverage your technical communications skill to partner with internal and external customers/stakeholders and will be involved in pre-silicon design, bringing new products/features to market, ultimately, making Neuron compiler highly performant and easy-to-use.

Experience in object-oriented languages like C++/Java is a must, experience with compilers or building ML models using ML frameworks on accelerators (e.g., GPUs) is preferred but not required.

Experience with technologies like OpenXLA, StableHLO, MLIR will be added bonus!

Explore the product and our history! https://awsdocs-neuron.readthedocs-hosted.com/en/latest/neuron-guide/neuron-cc/index

htmlhttps://aws.amazon.com/machine-learning/neuron/

https://github.com/aws/aws-neuron-sdk https://www.amazon.science/how-silicon-innovation-became-the-secret-sauce-behind-awss-success

AWS Utility Computing (UC) provides product innovations - from foundational services such as Amazon’s Simple Storage Service (S3) and Amazon Elastic Compute Cloud (EC2), to consistently released new product innovations that continue to set AWS’s services and features apart in the industry. As a member of the UC organization, you’ll support the development and management of Compute, Database, Storage, Internet of Things (Iot), Platform, and Productivity Apps services in AWS, including support for customers who require specialized security solutions for their cloud services.

Key job responsibilities

You will design, implement, test, deploy and maintain innovative software solutions to transform Neuron compiler’s performance, stability and user-interface. You will work side by side with chip architects, runtime/OS engineers, scientists and ML Apps teams to seamlessly deploy state of the art ML models from our customers on AWS accelerators with optimal cost/performance benefits. You will have opportunity to work with open-source software (e.g., StableHLO, OpenXLA, MLIR) to pioneer optimizing advanced ML workloads on AWS software and hardware. You will also work on building innovative features that will deliver best possible experiences for our customers - developers across the globe.

A day in the life

As you design and code solutions to help our team drive efficiencies in compiler architecture, you’ll create compiler optimization and verification passes, build features surface features and peculiarities of AWS accelerators to developers, implement tools to analyze numerical errors, and resolve the root cause of compiler defects. You’ll also participate in design discussions, code review, and communicate with internal (other Neuron SDK and Amazon wide teams) and external stakeholders (open-source communities). Lastly, work in a startup-like development environment, where you’re always working on the most important stuff.

About the team

About the Team

Our team is dedicated to supporting new members. We have a broad mix of experience levels and tenures, and we’re building an environment that celebrates knowledge-sharing and mentorship. Our senior members enjoy one-on-one mentoring and thorough, but kind, code reviews. We care about your career growth and strive to assign projects that help our team members develop your engineering expertise so you feel empowered to take on more complex tasks in the future.

Diverse Experiences

Amazon values diverse experiences. Even if you do not meet all of the qualifications and skills listed in the job description, we encourage candidates to apply. If your career is just starting, hasn’t followed a traditional path, or includes alternative experiences, don’t let it stop you from applying.

Inclusive Team Culture

Here at Amazon, it’s in our nature to learn and be curious. Our employee-led affinity groups foster a culture of inclusion that empower us to be proud of our differences. Ongoing events and learning experiences, including our Conversations on Race and Ethnicity (CORE) and AmazeCon conferences, inspire us to never stop embracing our uniqueness.

Work/Life Balance

We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home, which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home, there’s nothing we can’t achieve in the cloud.

Basic qualifications

- 3+ years of non-internship professional software development experience

- 2+ years of non-internship design or architecture (design patterns, reliability and scaling) of new and existing systems experience

- Experience programming with at least one software programming language

Preferred qualifications

- Master's degree or PhD in Computer Science, or a related technical field.

- 3+ years of experience writing production grade code in object-oriented languages such as C++/Java.

- Experience in compiler design for CPU/GPU/Vector engines/ML-accelerators.

- Experience with OpenSource compiler toolset like LLVM/MLIR.

- Experience with the following technologies: PyTorch, OpenXLA, StableHLO, JAX, TVM, deep learning models, and algorithms.

- Experience with modern build systems like Bazel/CMake.

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Los Angeles County applicants: Job duties for this position include: work safely and cooperatively with other employees, supervisors, and staff; adhere to standards of excellence despite stressful conditions; communicate effectively and respectfully with employees, supervisors, and staff to ensure exceptional customer service; and follow all federal, state, and local laws and Company policies. Criminal history may have a direct, adverse, and negative relationship with some of the material job duties of this position. These include the duties and responsibilities listed above, as well as the abilities to adhere to company policies, exercise sound judgment, effectively manage stress and work safely and respectfully with others, exhibit trustworthiness and professionalism, and safeguard business operations and the Company’s reputation. Pursuant to the Los Angeles County Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.

USA, CA, Cupertino - 165,200.00 - 223,600.00 USD annually

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,092,897 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Cupertino
≈ $134k – $254k per year (Estimated) • Hybrid • 5+ years exp • Bachelor's Degree • San Diego
Python
JavaScript
TypeScript
C#
AI/ML
Copilot
LangGraph
LangChain
Claude
Model Context Protocol
Prompt Engineering
AI Agents
Semantic Kernel
RAG
OpenAI
DevOps
Azure
CI/CD
Cybersecurity
CAPA
Management
Power Automate
Apply
≈ $119k – $225k per year (Estimated) • Hybrid • 5+ years exp • Bachelor's Degree • Marlborough
Python
JavaScript
TypeScript
C#
AI/ML
Copilot
LangGraph
LangChain
Claude
Model Context Protocol
Prompt Engineering
AI Agents
Semantic Kernel
RAG
OpenAI
DevOps
Azure
CI/CD
Cybersecurity
CAPA
Management
Power Automate
Apply
$162k – $259k per year • In office • North Reading
Python
AI/ML
Model Context Protocol
AI Agents
RAG
DevOps
AIOps
Apply
AI/ML Engineer 2 days ago
≈ $91k – $210k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Houston
Python
SQL
AI/ML
LangGraph
LangChain
Scikit-learn
Prompt Engineering
AI Agents
TensorFlow
PyTorch
LLM
RAG
LLMOps
Machine Learning
DevOps
Azure
AWS
Management
Agile
Apply
Senior AI Engineer 2 hours ago
≈ $125k – $237k per year (Estimated) • In office • 8+ years exp • Middleton
AI/ML
Copilot
RAG
OpenAI
Machine Learning
DevOps
Azure
FinOps
Analytics
Tableau
Power BI
Apply
$130k – $150k per year • In office • TS/SCI • 8+ years exp • Bachelor's Degree • Fairborn
C
C++
C
Valgrind
C++
CMake
LLVM
Abseil
DevOps
GCP
Azure
CI/CD
AWS
Docker
Configuration Management
Linux
Apply
≈ $21k – $47k per year (Estimated) • In office • Full-Time • 3+ years exp • Henderson
Python
JavaScript
Java
TypeScript
C#
Node JS
Java
Spring Boot
C#
.NET
Frontend
Vue.js
Angular
React.js
DevOps
Rest API
GCP
Azure
CI/CD
AWS
Incident Management
Management
Agile
Scrum
Apply
$64k per year • Hybrid • Internship • Bachelor's Degree • Chicago
Python
JavaScript
SQL
C++
DevOps
Linux
Management
Agile
Apply
$64k per year • Hybrid • Internship • Bachelor's Degree • Irving
Python
JavaScript
SQL
C++
DevOps
Linux
Management
Agile
Apply
$88k – $109k per year • In office • 8+ years exp • Bachelor's Degree • Calgary
JavaScript
Java
TypeScript
SQL
COBOL
Java
Spring Boot
COBOL
IBM MQ
Databases
Oracle
RabbitMQ
Apache Kafka
AI/ML
Copilot
Cursor
Claude Code
Prompt Engineering
Frontend
Angular
DevOps
Rest API
Splunk
Dynatrace
Prometheus
Azure
CI/CD
Jenkins
AWS
Docker
Kubernetes
Grafana
Platform Engineering
API Gateway
Management
Agile
Scrum
QA
Swagger
Apply
$158k – $214k per year • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • New York
Java
C#
C++
Perl
AI/ML
Computer Vision
TPU
Machine Learning
Apply
$230k – $311k per year • Equity • In office • Full-Time • 15+ years exp • PhD • Arlington
AI/ML
AWS Bedrock
Amazon SageMaker
DevOps
AWS
Amazon EC2
Cybersecurity
Zero Trust
Apply
$192k – $260k per year • Equity • In office • Full-Time • 6+ years exp • Master's Degree • Sunnyvale
Python
Java
C++
C++
TensorFlow C++
AI/ML
Hadoop
Spark
Scikit-learn
SciPy
TensorFlow
NumPy
Machine Learning
Apply
$184k – $249k per year • Equity • In office • Full-Time • 6+ years exp • Master's Degree • New York
Python
Java
C++
C++
TensorFlow C++
AI/ML
Hadoop
Spark
Model Context Protocol
Scikit-learn
SciPy
AI Agents
NLP
TensorFlow
NumPy
Anomaly Detection
LLM Guardrails
Tool Use
Machine Learning
Apply
≈ $154k – $291k per year (Estimated) • Equity • In office • Full-Time • 7+ years exp • Master's Degree • Austin
Python
C++
MATLAB
AI/ML
AI Agents
Machine Learning
Apply
≈ $233k – $440k per year (Estimated) • Equity • In office • 8+ years exp • Master's Degree • Cupertino
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Gemma
Fine-tuning
Embeddings
Quantization
JAX
Multimodal AI
Knowledge Distillation
Function Calling
AI Agents
Llama
Mistral
TensorFlow
PyTorch
LLM
RAG
BERT
Reranking
Semantic Search
Human-in-the-Loop
Semantic Search
Edge AI
Recommender Systems
Tool Use
Model Distillation
Machine Learning
Apply
≈ $228k – $432k per year (Estimated) • Equity • In office • 2+ years exp • Master's Degree • Cupertino
Python
Java
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Hadoop
Spark
TensorFlow
PyTorch
Machine Learning
Apply
$124k – $146k per year • In office • 7+ years exp • Bachelor's Degree • Chicago • Atlanta • Minneapolis • Charlotte • Cupertino
Python
AI/ML
LangGraph
LangChain
Prompt Engineering
AI Agents
LLM
RAG
DevOps
Terraform
Azure
CI/CD
AWS
Docker
Kubernetes
Platform Engineering
Bicep
Apply
≈ $225k – $425k per year (Estimated) • Equity • In office • 2+ years exp • Bachelor's Degree • Cupertino
Python
AI/ML
Prompt Engineering
NLP
RAG
SFT
Machine Learning
Apply
$165k – $224k per year • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • Cupertino
Python
C++
C++
PyTorch C++
AI/ML
DeepSeek
vLLM
CUDA Toolkit
JAX
SGLang
TensorRT
Llama
PyTorch
LLM
CUDA
Triton
AWS Trainium
CUTLASS
Machine Learning
DevOps
CI/CD
AWS
GitHub
Management
Agile
Apply
See all jobs
This is one of many
1,092,897 more open roles from verified company boards, updated every day.