368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$215k – $285k per year
Location
In office (Sunnyvale)
Employment
Full-Time
Overview
Company
Impact
Profile match
Applied Intuition is a Silicon Valley software company founded in 2017 that develops digital infrastructure and simulation platforms for autonomous systems and physical AI. The platform provides advanced tools for simulation, synthetic data generation, and vehicle software management, allowing engineers to develop, test, and safely deploy self-driving technology. It serves leading global automotive manufacturers, defense organizations, and commercial transport companies across sectors like trucking, agriculture, and mining.

Applied Intuition, Inc. is powering the future of physical AI. Founded in 2017 and now valued at $15 billion, the Silicon Valley company is creating the digital infrastructure needed to bring intelligence to every moving machine on the planet. Applied Intuition services the automotive, defense, trucking, construction, mining and agriculture industries in three core areas: tools and infrastructure, operating systems, and autonomy. Eighteen of the top 20 global automakers, as well as the United States military and its allies, trust the company’s solutions to deliver physical intelligence. Applied Intuition is headquartered in Sunnyvale, California, with offices in Washington, D.C.; San Diego; Ft. Walton Beach, Florida; Ann Arbor, Michigan; London; Stuttgart; Munich; Stockholm; Bangalore; Seoul; and Tokyo. Learn more at applied.co.

We are an in-office company, and our expectation is that full-time employees primarily work from their Applied Intuition office 5 days a week. However, we also recognize the importance of flexibility and trust our employees to manage their schedules responsibly. This may include occasional remote work, starting the day with morning meetings from home before heading to the office, or leaving earlier when needed to accommodate family commitments. This in-office expectation does not apply to contractor positions

About the Role

We are looking for a performance engineer who specializes in making large-scale machine learning workloads fast and cost-efficient in the datacenter. This role is focused on distributed training runs spanning many nodes, and high-throughput batch inference sweeping petabytes of real-world autonomy logs for auto-labeling, data mining, ground-truth generation, and evaluation.

The optimization target here is not tail latency on a vehicle - it is throughput, cluster goodput, and cost per unit of data processed. A training run that wastes 30% of its GPU-hours on stalled data loaders, or an offline inference sweep that takes a week instead of a day, directly slows down how fast the whole company can iterate. You will own the gap between what our fleet of accelerators is theoretically capable of and what our workloads actually achieve: profiling across the stack, finding where the compute and the wall-clock time actually go, and closing the difference.

You will work at the intersection of accelerators, ML frameworks, and large-scale data infrastructure, partnering with the teams who own each layer to land wins that show up in training time-to-result and offline processing cost. At Applied, we encourage all engineers to take ownership over technical and product decisions, closely interact with users to collect feedback, and contribute to a thoughtful, dynamic team culture.

At Applied, you will:

  • Profile and optimize distributed training end to end - data loading and preprocessing, augmentation, kernel execution, gradient communication, and checkpointing

  • Optimize large-scale offline and batch inference over petabyte-scale sensor logs: batching and scheduling strategies, quantization and low-precision execution, graph optimization, and accelerator saturation across long-running sweeps

  • Establish roofline and performance models for our workloads, quantify the gap between achieved and theoretical performance, and stack-rank optimization opportunities by impact and effort

  • Improve multi-node scaling efficiency: sharding and parallelism strategies, collective communication, interconnect utilization, and memory-bandwidth and kernel-fusion bottlenecks

  • Drive cluster goodput - reduce GPU idle time from input pipeline stalls, storage and network I/O, scheduling gaps, stragglers, and failure recovery on long-running jobs

  • Build the benchmarking, observability, and regression-detection tooling that keeps performance from silently degrading as models and code evolve

  • Collaborate with engineers across functions to solve complex data and compute problems at scale

  • Contribute to a team culture that values effective collaboration, technical excellence, and innovation

We're looking for someone who has:

  • Hands-on ML performance engineering experience: profiling, roofline analysis, throughput optimization, and root-cause investigation in production systems

  • Experience with distributed multi-node training at scale (FSDP, DeepSpeed, Megatron, NCCL, or equivalent), including diagnosing scaling inefficiency as node count grows

  • Deep familiarity with GPU or accelerator performance concepts - memory bandwidth, kernel launch overhead, occupancy, quantization, collective communication

  • Experience with high-throughput or batch inference systems (NVIDIA Triton Inference Server, TensorRT, ONNX Runtime, Ray, or similar)

  • Fluency in Python and proficiency in C++ or another systems language

  • Excellent debugging, analytical, and problem-solving skills

  • A deep understanding of machine learning foundations, and the ability to develop technical solutions for problems with no established playbook

Nice to have:

  • GPU kernel development experience: CUDA, Triton, CUTLASS, or hand-tuned attention implementations

  • Experience with profiling toolchains such as Nsight Systems/Compute, PyTorch Profiler, or perf

  • Experience with GPU scheduling and orchestration on Kubernetes, Slurm, or Ray, including multi-tenant cluster utilization

  • Experience with fault tolerance and elastic training for long-running jobs - checkpointing strategy, straggler mitigation, preemption recovery

  • Familiarity with autonomy or robotics data (ROS, OpenCV, multi-sensor log formats)

Don’t meet every single requirement? If you’re excited about this role but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyway. You may be just the right candidate for this or other roles.

Applied Intuition is an equal opportunity employer and federal contractor or subcontractor. Consequently, the parties agree that, as applicable, they will abide by the requirements of 41 CFR 60-1.4(a), 41 CFR 60-300.5(a) and 41 CFR 60-741.5(a) and that these laws are incorporated herein by reference. These regulations prohibit discrimination against qualified individuals based on their status as protected veterans or individuals with disabilities, and prohibit discrimination against all individuals based on their race, color, religion, sex, sexual orientation, gender identity or national origin. These regulations require that covered prime contractors and subcontractors take affirmative action to employ and advance in employment individuals without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status or disability. The parties also agree that, as applicable, they will abide by the requirements of Executive Order 13496 (29 CFR Part 471, Appendix A to Subpart A), relating to the notice of employee rights under federal labor laws.

FOR US-BASED ROLES:Applied Intuition is committed to providing an accessible and inclusive application and interview experience to applicants who are disabled veterans and other applicants with disabilities or medical conditions. Reasonable accommodations are available, requesting an accommodation will not affect your candidacy in any way, and you are not required to disclose the nature of your disability or medical condition in order to make a request.

If you require an accommodation please contact [email protected]. We will work with you!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Sunnyvale
$55k – $157k per year (Estimated) • Remote • Full-Time • Sydney
C++
Go
Lua
Python
C++
CMake
Databases
ActiveMQ
Aerospike
Apache Kafka
Cassandra
DevOps
Docker
gRPC
Kubernetes
Apply
In office • Full-Time • 7+ years exp • Bachelor's Degree • Ho Chi Minh City
JavaScript
Python
AI/ML
Copilot
Cursor
DevOps
Azure DevOps
Jenkins
Azure
GitLab
QA
JMeter
k6
Playwright
Postman
Rest-Assured
Selenium
TestRail
Apply
$28k – $65k per year (Estimated) • In office • Full-Time • 12+ years exp • Bengaluru
Bash
PowerShell
Python
Node JS
JavaScript
Node JS
Commander.js
AI/ML
AI Agents
DevOps
Amazon EC2
Amazon EKS
AWS
Azure
Kubernetes
Amazon ECS
IAM
Cybersecurity
Crowdstrike
Zero Trust
Apply
$72k – $201k per year (Estimated) • In office • Full-Time • 3+ years exp • PhD • Basel
Python
AI/ML
Multimodal AI
OpenCV
PyTorch
scikit-image
Scikit-learn
TensorFlow
AI Agents
DevOps
Git
Apply
$20k – $52k per year (Estimated) • In office • Full-Time • Gurgaon
SQL
Python
Python
pySpark
AI/ML
Hadoop
Spark
Analytics
Power BI
Tableau
Apply
System Designer 2 days ago
$170k – $200k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • Sunnyvale
AI/ML
AI Agents
Multimodal AI
Apply
$116k – $174k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Munich
Bash
C++
Go
Java
Python
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Apply
$150k – $220k per year • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree
Apply
$110k – $150k per year • Equity • In office • Full-Time • 2+ years exp • Ann Arbor
C++
Python
AI/ML
Claude
Claude Code
Copilot
Cursor
AI Agents
DevOps
Docker
GitHub
Robotics
Path Planning
Sensor Fusion
Apply
$26k – $61k per year (Estimated) • In office • Full-Time • 4+ years exp • Bengaluru
Design
SolidWorks
Apply
Sr. SDET 1 hour ago
$109k – $163k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Sunnyvale
Java
Python
SQL
Java
Spring Framework
Mobile
JUnit
DevOps
CI/CD
Kubernetes
QA
Cypress
JMeter
Postman
Apply
$168k – $356k per year • In office • Full-Time • 8+ years exp • Master's Degree • Sunnyvale
Python
AI/ML
AI Agents
ChatGPT
Gemini
LLM
NLP
RAG
Recommender Systems
Semantic Search
Semantic Search
Vertex AI
Apply
$81k – $155k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Sunnyvale
Apply
$148k – $267k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Sunnyvale • Boston • Austin • Toronto • New York
AI/ML
AI Agents
Claude
Claude Code
Lovable
Replit
Cybersecurity
BeyondTrust
Crowdstrike
CyberArk
Delinea
Microsoft Entra ID
Okta
PCI DSS
SOC 2
Threat Modeling
Apply
$162k – $180k per year • Equity • In office • 5+ years exp • Bachelor's Degree • Sunnyvale
AI/ML
AI Agents
Chain-of-Thought
Fine-tuning
LLM
Prompt Engineering
RAG
Robotics
Digital Twin
Analytics
A/B Testing
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.