1,401,056open jobs
81,239companies
208,212added this week
Browse all
Salary
≈ $234k – $446k per year (Estimated)
Location
In office (Santa Clara)
Seniority
Senior · 5+ years exp
Visa
H-1B filings in 12 months: 6,350 · for this role: 2,959 · green card filings: 37

Confirmed on the employer's own hiring board on Oct 9, 2026. First seen by Alion on Sep 14, 2026.

Overview
Company
Impact
Profile match
Apple is an American multinational technology company founded in 1976 by Steve Jobs, Steve Wozniak and Ronald Wayne, and headquartered in Cupertino, California. It designs and sells consumer hardware including the iPhone, Mac, iPad, Apple Watch, AirPods and Vision Pro, together with the operating systems and silicon that run them. A growing services division built around the App Store, iCloud, Apple Music, Apple TV+ and Apple Pay now contributes a large share of profit, making Apple one of the most valuable companies in the world.

At Apple, machine learning powers experiences that anticipate what people need before they ask. We're looking for a Senior Machine Learning Engineer to help build the next generation of intelligent search and AI experiences technology that understands user intent, context, and personal information while preserving privacy. In this role, you'll design, train, fine-tune, optimize, and deploy large language models, semantic retrieval systems, and ranking models that power relevant, personalized, and context-aware experiences across Apple's ecosystem.

Description

You'll design, train, fine-tune, and optimize transformer-based language models and foundation models for efficient on-device deployment, and build semantic retrieval, embedding, reranking, and retrieval-augmented generation systems that improve search quality and AI-powered experiences. You'll develop models for query understanding, intent prediction, personalization, retrieval, and ranking, while researching new approaches to LLM fine-tuning, knowledge distillation, model compression, quantization, and low-latency inference. You'll explore techniques for adapting large foundation models into smaller, highly capable models that can operate efficiently under on-device memory, compute, power, and latency constraints.

You'll partner with engineers, researchers, product managers, and designers to bring new AI capabilities from research into production, driving technical strategy and leading projects from early exploration through large-scale deployment. This is an opportunity to explore new applications of foundation models, multimodal AI, agentic retrieval, and personalized intelligence, shaping the next generation of proactive and intelligent user experiences.

Minimum Qualifications

Master degree in Computer Science, Machine Learning, Artificial Intelligence, or a related field.

5+ years of industry or research experience developing machine learning systems.

Background in machine learning, deep learning, natural language processing, information retrieval, search, recommender systems, or generative AI.

Experience training, fine-tuning, or deploying transformer-based models and large language models.

Experience with modern deep learning architectures and techniques, including transformers, embeddings, representation learning, and neural ranking.

Programming skills in Python and/or C/C++, with experience building production-quality software using modern machine learning frameworks such as PyTorch, JAX, or TensorFlow.

Ability to work onsite in Cupertino, California, in accordance with Apple's applicable work policies.

Preferred Qualifications

Master's or Ph.D. in Computer Science, Machine Learning, Artificial Intelligence, or a related field.

Experience optimizing machine learning models for resource-constrained environments, including knowledge distillation, model compression, quantization, and pruning.

Experience with on-device machine learning or edge AI, or mobile inference frameworks, including optimizing models for latency, memory, compute, and power constraints.

Experience distilling capabilities from large foundation models into small language models or task-specific models for efficient inference.

Experience building retrieval-augmented generation, vector search, embedding retrieval, neural reranking, or semantic search systems.

Experience with query understanding, query rewriting, intent classification, personalized retrieval, learning-to-rank, or recommendation models.

Experience working with transformer architectures and foundation model families such as BERT, T5, Llama, Gemma, Mistral, or related architectures.

Experience evaluating language models, designing AI quality metrics, and building automated and human-in-the-loop evaluation pipelines.

Experience building large-scale production search, recommendation, personalization, or generative AI systems.

Familiarity with multimodal foundation models, tool use, agentic AI, or agentic retrieval systems.

Strong understanding of the tradeoffs among model quality, latency, memory, power consumption, privacy, and reliability for production on-device AI systems.

Ability to prototype new ideas, conduct rigorous experiments, solve ambiguous technical problems, and translate research advances into production-quality machine learning solutions.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,401,056 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Santa Clara
≈ $116k – $221k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Reading
Python
JavaScript
SQL
Analytics
Power BI
Management
Agile
Apply
AI Vision Engineer 22 days ago
$75k – $165k per year • Hybrid • 8+ years exp • Bachelor's Degree • United States
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
OpenCV
Computer Vision
TensorFlow
PyTorch
OCR
Machine Learning
DevOps
Azure
AWS
Linux
Apply
$170k – $210k per year • Hybrid • Full-Time • 4+ years exp • Boston • San Francisco • Washington
Python
AI/ML
Weights & Biases
Model Context Protocol
vLLM
MLFlow
Fine-tuning
Embeddings
AI Agents
SGLang
TensorRT
TensorRT-LLM
PyTorch
LLM
RAG
Hybrid Search
Hugging Face
Context Engineering
LLM Evaluation
Speculative Decoding
DevOps
OpenTelemetry
CI/CD
Kubernetes
Platform Engineering
API Gateway
Apply
$208k – $312k per year • Hybrid • 5+ years exp • New York
Python
Go
JavaScript
TypeScript
AI/ML
Prompt Engineering
AI Agents
LLM
OpenAI
Feature Store
Machine Learning
DevOps
Vercel
Apply
≈ $149k – $297k per year (Estimated) • In office • Full-Time • 4+ years exp • San Francisco
AI/ML
Vision-Language-Action
DevOps
Terraform
GCP
AWS
Kubernetes
Platform Engineering
Apply
$176k – $308k per year • In office • Full-Time • Santa Clara
Python
Java
AI/ML
Machine Learning
Management
ServiceNow
Apply
$159k – $178k per year • In office • Full-Time • 1+ year exp • Master's Degree • Santa Clara
JavaScript
AI/ML
Multimodal AI
Computer Vision
AI Agents
Machine Learning
DevOps
Docker
Kubernetes
Management
ServiceNow
Apply
≈ $189k – $386k per year (Estimated) • In office • Santa Clara
SQL
Databases
ElasticSearch
Trino
AI/ML
Spark
DeepSeek
vLLM
Triton Inference Server
Quantization
Multimodal AI
Function Calling
AI Agents
SGLang
TensorRT
TensorRT-LLM
Llama
Transformers
RAG
Mixture of Experts
InfiniBand
NVLink
KV Cache
Multi-Agent Systems
Tool Use
DevOps
OpenShift
Kubernetes
Gateway API
Amazon S3
Cybersecurity
Zero Trust
Analytics
ETL/ELT
Apply
AI Researcher 1 day ago
≈ $120k – $264k per year (Estimated) • In office • Master's Degree • Plano • Santa Clara
AI/ML
JAX
PyTorch
Physical AI
Machine Learning
Apply
$76k – $188k per year • In office • Internship • PhD • Santa Clara
Python
AI/ML
Function Calling
AI Agents
PyTorch
Tool Use
Apply
See all jobs
This is one of many
1,401,056 more open roles from verified company boards, updated every day.