1,454,406open jobs
86,438companies
225,454added this week
Browse all
Salary
$210k – $273k per year
Location
Remote (United States)
Seniority
Senior
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 11, 2026. First seen by Alion on Oct 7, 2026. Unity Technologies scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Headquartered in San Francisco, California, Unity Software Inc. (Unity Technologies) is a technology enterprise specializing in real-time 3D (RT3D) content development software and engine solutions. The company provides an integrated platform that enables developers, creators, and engineers to build, run, and monetize interactive 2D, 3D, virtual reality (VR), and augmented reality (AR) experiences across mobile, console, PC, and spatial computing hardware. By combining cross-platform runtime engines with developer services, cloud tools, and advertising monetization networks, it serves as a core infrastructure provider for the global gaming industry as well as industrial sectors such as automotive, architecture, and film production.

The Role

We are seeking a Senior ML engineer to design and evolve Unity Vector’s online model inference platform. This role focuses on building reliable infrastructure for serving machine learning models in production, optimizing inference performance, and enabling safe, efficient experimentation across high-traffic online systems.

You will work closely with ML engineers, platform teams, and product stakeholders to ensure models can be deployed, scaled, monitored, and iterated on efficiently. You will play a key role in shaping how models are packaged, served, validated, monitored, and optimized in production environments.

This role requires strong systems thinking, deep experience with production ML infrastructure, and the ability to drive architectural improvements across teams.

What you'll be doing

  • Design and operate large-scale online inference infrastructure that serves production ML models with low latency and high reliability, such as PyTorch, Triton Inference Server, Kubernetes, GKE, Ray, or similar distributed serving frameworks.
  • Develop infrastructure that supports distributed training workflows using technologies such as Pytorch, Ray Data, and Ray Train, etc.
  • Integrate ML pipelines with workflow orchestration systems (e.g., Flyte, Airflow, or similar) to enable reliable multi-stage training workflows
  • Optimize model performance through model compilation, GPU/CPU utilization improvements, request scheduling, kernel fusion, and runtime-level tuning.
  • Improve observability of ML systems through latency, throughput, error-rate, cost, saturation, and model-health monitoring.
  • Partner closely with ML engineers to support faster model iteration while maintaining production safety, scalability, and cost efficiency.
  • Improve the reliability and reproducibility of model serving workflows, including model packaging, artifact validation, compatibility testing, and deployment automation.
  • Lead architectural improvements that make the online ML platform more robust, user-friendly, scalable, and cost-efficient.

What we're looking for

  • Experience building and operating production-grade online ML inference systems, such as NVIDIA Triton Inference Server, TorchServe, Ray Serve, TensorFlow Serving, or similar systems.
  • Experience with model serving frameworks such as NVIDIA Triton Inference Server, TorchServe, Ray Serve, TensorFlow Serving, or similar systems.
  • Experience optimizing inference workloads using techniques such as dynamic batching, model compilation, quantization, GPU acceleration, GPU kernel optimization, caching, or runtime tuning.
  • Strong experience with distributed systems, Kubernetes, autoscaling, service reliability, and production observability.
  • Strong programming skills in Python, with practical experience working on production ML systems and high-scale services.
  • Experience with PyTorch and modern model deployment workflows, including model packaging, validation, and serving lifecycle management.
  • Experience designing infrastructure for safe model rollout, canary testing, A/B experimentation, and automated rollback.
  • Strong systems thinking, with the ability to reason about latency, throughput, reliability, scalability, and cost tradeoffs in online systems.
  • Proven ability to lead technical direction and influence architectural decisions across teams without formal authority.

Additional information

Zone A: $210,300 - $273,400

Zone B: $187,200 - $243,300

Zone C: $165,600 - $215,200

This range reflects the anticipated base salary for this position. Beyond base salary, this role may be eligible for equity awards and participation in our company incentive plans (such as annual discretionary bonuses or sales commissions). The final offer amount will depend on several factors, including geographic location and the candidate’s relevant experience, professional background, and skill set.

Benefits

At Unity, we want our team members to thrive. We offer a wide range of benefits designed to support well-being and work-life balance.

Please note: Benefits eligibility, specific offerings, and coverage vary based on the country and employment status.

While specific benefits vary, here are some of the ways we strive to take care of our eligible team members globally: Comprehensive health, life, and disability insurance | Commute subsidy | Employee stock ownership | Competitive retirement/pension plans | Generous vacation and personal days | Support for new parents through leave and family-care programs | Office food snacks | Mental Health and Wellbeing programs and support | Employee Resource Groups | Global Employee Assistance Program | Training and development programs | Volunteering and donation matching program

Life at Unity

Unity [NYSE: U] is the world’s leading game engine, powering play for more than 3 billion consumers each month. The top mobile games in the world, the most played PC indie titles, the most innovative console games, and virtually all of the top XR and Web Games are developed, deployed, and grown in Unity. Unity also enables teams across industries like automotive, manufacturing, and healthcare to design, simulate, and collaborate in 3D - closing the gap between ideas and reality. For more information, please visit www.unity.com.

Unity is a proud equal opportunity employer. We are committed to fostering an inclusive, innovative environment and celebrate our employees across age, race, color, ancestry, national origin, religion, disability, sex, gender identity or expression, sexual orientation, or any other protected status in accordance with applicable law. Our differences are strengths that enable us to support the growing and evolving needs of our customers, partners, and collaborators. If you have a disability that means there are preparations or accommodations we can make to help ensure you have a comfortable and positive interview experience, please fill out this form to let us know.

This position requires the incumbent to have a sufficient knowledge of English to have professional verbal and written exchanges in this language since the performance of the duties related to this position requires frequent and regular communication with colleagues and partners located worldwide and whose common language is English.

This posting is intended to fill an existing vacancy, and we are committed to providing applicants with updates throughout the hiring process in accordance with applicable law.

Headhunters and recruitment agencies may not submit resumes/CVs through this website or directly to managers. Unity does not accept unsolicited headhunter and agency resumes. Unity will not pay fees to any third-party agency or company that does not have a signed agreement with Unity.

Your privacy is important to us. Please take a moment to review ourProspect and Applicant Privacy Policies. Should you have any concerns about your privacy, please contact us at [email protected].

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,454,406 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Washington
≈ $140k – $265k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Houston
Python
C++
C++
PyTorch C++
AI/ML
vLLM
CUDA Toolkit
Quantization
PyTorch
RAG
CUDA
LLMOps
LLM Evaluation
Machine Learning
DevOps
AWS
Kubernetes
Management
Agile
Apply
≈ $133k – $253k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Houston • Tlaquepaque
Python
Java
C++
Python
FastAPI
C++
TensorFlow C++
PyTorch C++
AI/ML
Spark
Model Context Protocol
Scikit-learn
AWS Bedrock
TensorFlow
PyTorch
Amazon SageMaker
Edge AI
Machine Learning
DevOps
Terraform
Azure
CI/CD
AWS
Kubernetes
Management
Agile
Apply
$129k – $216k per year • Equity • In office • 4+ years exp • Bachelor's Degree • Santa Rosa
Python
JavaScript
TypeScript
C#
AI/ML
Copilot
Embeddings
AI Agents
TensorFlow
PyTorch
LLM
RAG
Frontend
GraphQL
Angular
React.js
DevOps
Rest API
VMWare
Azure
CI/CD
Jenkins
AWS
Docker
Kubernetes
Bitbucket
GitHub
Cybersecurity
Least Privilege
Management
Confluence
Jira
Apply
$129k – $216k per year • Equity • In office • Master's Degree • Calabasas
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Reinforcement Learning
Scikit-learn
TensorFlow
PyTorch
Machine Learning
Chips/EDA
Cadence Virtuoso
Management
Agile
Apply
≈ $132k – $251k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Colorado Springs
Python
C#
C++
AI/ML
LangGraph
AutoGen
LangChain
Prompt Engineering
Function Calling
AI Agents
Semantic Kernel
CrewAI
LLM
RAG
Tool Use
DevOps
CI/CD
Git
Bitbucket
Management
Confluence
Jira
Agile
Apply
$132k – $170k per year • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Washington • San Francisco
Apply
$122k – $146k per year • In office • Full-Time • Portland • San Francisco • Washington • Boston
AI/ML
Claude
Model Context Protocol
AI Agents
Gemini
AWS Bedrock AgentCore
Agentic Workflows
Multi-Agent Systems
DevOps
CI/CD
AWS
Apply
$79k – $118k per year • In office • 5+ years exp • Bachelor's Degree • Washington
DevOps
Windows
Apply
Data Scientist 1 day ago
$112k – $179k per year • In office • TS/SCI • 5+ years exp • Bachelor's Degree • Washington
JavaScript
SQL
AI/ML
Machine Learning
Analytics
Tableau
Management
ServiceNow
Apply
≈ $40k – $83k per year (Estimated) • In office • Top Secret • 2+ years exp • High School Diploma • Washington
Management
Microsoft Office
Apply
See all jobs
This is one of many
1,454,406 more open roles from verified company boards, updated every day.