599,985open jobs
30,608companies
86,410added this week
Browse all
Location
In office
Employment
Full-Time
Overview
Company
Impact
Profile match

About the Role

We’re looking for a Machine Learning Engineer focused on model distillation to help us build smaller, faster, and more efficient models without sacrificing quality. You’ll work at the intersection of research and production-taking cutting-edge techniques and turning them into systems that scale.

This is a hands-on role with real ownership: you’ll design distillation pipelines, run large-scale experiments, and ship models used in production.

What You’ll Do

  • Design and implement knowledge distillation pipelines (teacher-student, self-distillation, multi-teacher, etc.)

  • Distill large foundation models into smaller, faster, and cheaper models for inference

  • Run and analyze large-scale training experiments to evaluate quality, latency, and cost tradeoffs

  • Collaborate with research to translate new distillation ideas into production-ready code

  • Optimize training and inference performance (memory, throughput, latency)

  • Contribute to internal tooling, evaluation frameworks, and experiment tracking

  • (Optional) Contribute back to open-source models, tooling, or research

What We’re Looking For

  • Strong background in machine learning or deep learning

  • Hands-on experience with model distillation (LLMs or other neural networks)

  • Solid understanding of training dynamics, loss functions, and optimization

  • Experience with PyTorch (or JAX) and modern ML tooling

  • Comfort running experiments on multi-GPU or distributed setups

  • Ability to reason about model quality vs. performance tradeoffs

  • Pragmatic mindset: you care about shipping, not just papers

Nice to Have

  • Experience distilling LLMs or large sequence models

  • Experience with inference optimization (quantization, pruning, kernels, etc.)

  • Familiarity with evaluation for language models

  • Open-source contributions or research publications

  • Experience in early-stage or fast-moving startups

Why Join

  • Work on core model quality and cost efficiency -not side projects

  • High ownership and direct impact on product and roadmap

  • Small, senior team with strong research + engineering culture

  • Competitive compensation + meaningful equity

  • Remote-friendly, async-first environment

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
599,985 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$152k – $294k per year (Estimated) • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree • Pleasanton
Python
Python
pySpark
AI/ML
Spark
Fine-tuning
AI Agents
TensorFlow
Pandas
PyTorch
RAG
Recommender Systems
DevOps
GCP
AWS
Docker
Kubernetes
Analytics
ETL/ELT
Apply
$70k – $196k per year (Estimated) • In office • Bachelor's Degree • Waxahachie
AI/ML
PyTorch
Ignite
Apply
$240k – $280k per year • Remote/Hybrid • Full-Time • 3+ years exp • San Francisco • New York • San Mateo
AI/ML
Cursor
Fine-tuning
Multimodal AI
PyTorch
Fireworks AI
DevOps
Vercel
Apply
In office • Contractor • Oshawa
AI/ML
PyTorch
Ignite
Apply
Remote/Hybrid • 10+ years exp
AI/ML
Quantization
Computer Vision
ONNX
TensorRT
PyTorch
Synthetic Data
Edge AI
Robotics
nuScenes SDK
Sim-to-Real
Apply
$46k – $93k per year (Estimated) • Remote • Full-Time • 1+ year exp • Atlanta • Los Angeles • Tampa • Miami • Orlando
AI/ML
OpenAI
Management
n8n
Marketing
HubSpot
Apply
$180k – $240k per year • Remote • Full-Time • Atlanta • Los Angeles • Tampa • Miami • Orlando
AI/ML
Claude
ChatGPT
Marketing
HubSpot
Apply
$150k – $190k per year • Remote • Full-Time • Atlanta • Tampa • Miami • Orlando • Ottawa
AI/ML
Claude
ChatGPT
Marketing
HubSpot
Apply
Chief of Staff 1 month ago
Remote/Hybrid • Full-Time • San Francisco
Management
Discord
Apply
Content Marketer 3 months ago
$42k – $56k per year • Remote • Full-Time
Apply
See all jobs
This is one of many
599,985 more open roles from verified company boards, updated every day.