DeepSpeed Jobs - Remote & On-site

DeepSpeed jobs curated from vetted startups and product companies, updated daily. Many DeepSpeed roles offer remote or hybrid schedules plus transparent pay ranges to streamline decision-making. Browse and apply today.

AI/ML
Data Science
Backend
Frontend
Mobile
DevOps
Web3
Games
Hardware
Robotics
Security
QA
Executive
Networking
Product
Design
Analytics
Support
Enterprise Apps
Quantum

Micron Technology

micron.com
Micron Technology is an American semiconductor company founded in Boise, Idaho in 1978 and one of only a handful of manufacturers of both DRAM and NAND flash memory. It designs and fabricates memory and storage products used in data centres, personal computers, smartphones, cars and industrial equipment, and sells them under the Micron and consumer-facing Crucial brands. High-bandwidth memory for AI accelerators has become the fastest-growing part of the business, and the company operates fabrication and assembly sites across the United States and Asia.
micron.com • HQ: Boise, United States • Data Centers • Semiconductor Manufacturing • Consumer Electronics • Hardware • Storage • Digital Storage • Semiconductors • 5000+ employees • Est. 1978
HQ: Boise, United States • Data Centers • Semiconductor Manufacturing • Consumer Electronics • Hardware • Storage • Digital Storage • Semiconductors • 5000+ employees • Est. 1978
Verified live · 3 hours ago 28 days ago

MTS, AI Engineering, SMAI

Senior • 5+ years expIn office (Taichung, Taoyuan, Taiwan) • PhD • Full-Time • Senior
C++
Java
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
AI Agents
AutoGen
Chain-of-Thought
Computer Vision
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
Function Calling
Kubeflow
LangChain
LangGraph
LlamaIndex
LLM
LoRA
OpenCL
PEFT
Prompt Engineering
PyTorch
QLoRA
Ray
RLHF
Scikit-learn
TensorFlow
TensorRT
TensorRT-LLM
Triton
vLLM
Transformers
FSDP
Megatron-LM
NVLink
SFT
Frontend
Sass
Mobile
Metal
DevOps
CI/CD
Docker
Git
Jenkins
Kubernetes
SLURM
HPC
Apply
Verified live · 3 hours ago 2 months ago

Member of Technical Staff (MTS), Machine Learning, SMAI

$136k – $298k per year (Estimated) • Staff+ • 10+ years expIn office (Singapore) • Full-Time • Staff
C++
Java
Python
C++
PyTorch C++
TensorFlow C++
Databases
Snowflake
AI/ML
AI Agents
AutoGen
Chain-of-Thought
Computer Vision
CrewAI
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
Function Calling
Kubeflow
LangChain
LangGraph
LlamaIndex
LLM
LoRA
PEFT
Prompt Engineering
PyTorch
QLoRA
Ray
RLHF
Scikit-learn
TensorFlow
TensorRT
TensorRT-LLM
Triton
vLLM
Transformers
FSDP
Megatron-LM
NVLink
SFT
DevOps
CI/CD
Docker
GCP
Git
Jenkins
Kubernetes
SLURM
HPC
Analytics
ETL/ELT
Apply
Verified live · 3 hours ago 3 months ago

Member of Technical Staff, AI Engineering

$162k – $328k per year (Estimated) • Staff+ • 10+ years expIn office (Boise, United States) • Bachelor's Degree • Full-Time • Staff
C++
Java
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
AI Agents
AutoGen
Chain-of-Thought
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
Function Calling
LangChain
LangGraph
LlamaIndex
LLM
LoRA
OpenCL
PEFT
Prompt Engineering
PyTorch
QLoRA
RLHF
Scikit-learn
TensorFlow
TensorRT
TensorRT-LLM
vLLM
Transformers
Edge AI
FSDP
Megatron-LM
NVLink
SFT
Computer Vision
Kubeflow
Ray
Triton
Frontend
Sass
DevOps
CI/CD
Docker
Git
Jenkins
Kubernetes
SLURM
HPC
Apply
Report

Cantina

cantina.com
Cantina is a social artificial intelligence technology company based in San Francisco, California, and founded in 2023. The company provides a platform where users can create, interact with, and share multimodal AI characters that feature unique personalities and the ability to generate video content. It operates primarily through a mobile application and web interface, focusing on the intersection of generative AI and social media for a global audience of creators and consumers.
cantina.com • HQ: San Francisco, United States • Content Creation • Generative Video • AI Agents • LLM & Generative AI • Media & Entertainment • Artificial Intelligence • Conversational AI • Est. 2023
HQ: San Francisco, United States • Content Creation • Generative Video • AI Agents • LLM & Generative AI • Media & Entertainment • Artificial Intelligence • Conversational AI • Est. 2023
Verified live · 2 days ago 28 days ago

Machine Learning Engineer - Voice Conversion

$200k – $220k per year • In office • Full-Time
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
DeepSpeed
Knowledge Distillation
Multimodal AI
PyTorch
Synthetic Data
Tokenization
DPO
FSDP
GRPO
LLM Guardrails
Text-to-Speech
Speech Recognition
Apply
Verified live · 2 days ago 2 months ago

Machine Learning Engineer, Speech - Joint Audio-Video Modeling

$200k – $220k per year • In office • Full-Time
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
DeepSpeed
Knowledge Distillation
Multimodal AI
PyTorch
Quantization
Synthetic Data
Tokenization
DPO
FSDP
GRPO
LLM Guardrails
Text-to-Speech
Apply
Report

Thinking Machines Lab

thinkingmachines.ai
Thinking Machines Lab is an artificial intelligence research and product company based in San Francisco and founded in 2025. The company develops multimodal AI systems and open-weights models, such as Inkling, alongside developer tools like Tinker for model fine-tuning. It operates as a public benefit corporation focused on human-AI collaboration and open science, supported by significant venture capital investment.
thinkingmachines.ai • HQ: San Francisco, United States • MLOps • Machine Learning • Multimodal AI • LLM & Generative AI • Artificial Intelligence • 201-500 employees • Est. 2025
HQ: San Francisco, United States • MLOps • Machine Learning • Multimodal AI • LLM & Generative AI • Artificial Intelligence • 201-500 employees • Est. 2025
Top 25% payVerified live · 1 min ago 28 days ago

Research Engineer, Infrastructure, Inference

$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time
JAX
PyTorch
Ray
SGLang
vLLM
Edge AI
DeepSpeed
DevOps
Kubernetes
SLURM
Apply
Top 25% payVerified live · 1 min ago 28 days ago

Research Engineer, Infrastructure, Numerics

$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time
JAX
PyTorch
DeepSpeed
Megatron-LM
Apply
Top 25% payVerified live · 1 min ago 28 days ago

Research Engineer, Infrastructure, Training Systems

$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time
JAX
PyTorch
DeepSpeed
Megatron-LM
Apply
Report

Edwards

edwards.com
2025 Corporate Impact report Our greatest impact is in improving the lives of patients and we strive to be a good corporate citizen for all of our stakeholders. Letter from our CEO At Edwards, improving lives is our purpose and guides every decision we make.
edwards.com • HQ: Irvine, United States • Health Care • Medical Devices • 1001-5000 employees
HQ: Irvine, United States • Health Care • Medical Devices • 1001-5000 employees
Verified live · 1 hour ago 30 days ago

Staff Data Scientist

$126k – $178k per year • Staff+ • 6+ years expIn office (Irvine, United States) • Relocation • Bachelor's Degree • Full-Time • Staff
C++
C++
PyTorch C++
AI/ML
Contrastive Learning
CUDA
CUDA Toolkit
DeepSpeed
Diffusion Models
Federated Learning
JAX
Multimodal AI
PyTorch
PyTorch Lightning
Ray
Self-Supervised Learning
LLM Evaluation
Apply
Report

Meesho

meesho.com
Meesho is an Indian e-commerce company headquartered in Bengaluru and founded in 2015. The company operates an online marketplace that enables small businesses and individual entrepreneurs to sell products such as apparel, home goods, and electronics directly to consumers. Originally established as a social commerce platform facilitating sales through networks like WhatsApp and Facebook, it has evolved into a large-scale retail ecosystem serving millions of users across India.
meesho.com • HQ: Bengaluru, India • B2B Commerce • Apparel & Fashion • Consumer Goods • Commerce • Marketplaces • Social Commerce • Est. 2015
HQ: Bengaluru, India • B2B Commerce • Apparel & Fashion • Consumer Goods • Commerce • Marketplaces • Social Commerce • Est. 2015
Verified live · 12 hours ago 2 months ago

Engineering Manager – AI Engineering

$42k – $93k per year (Estimated) • Lead • 9+ years expIn office (Bengaluru, India) • Bachelor's Degree • Full-Time • Staff
C++
Python
Rust
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
DeepSpeed
Flink
Knowledge Distillation
LLM
PyTorch
Quantization
Ray
SGLang
Spark
TensorRT
TensorRT-LLM
vLLM
AI Agents
FSDP
LLMOps
Megatron-LM
DevOps
FinOps
Google GKE
Kubernetes
GCP
Apply
Report

Together AI

together.ai
Together AI (Together Computer, Inc.) is a full-stack AI infrastructure and cloud platform headquartered in San Francisco, California. Founded in 2022 by prominent AI researchers and system engineers - including CEO Vipul Ved Prakash, CTO Ce Zhang, Chief Scientist Tri Dao (co-creator of FlashAttention), Chris Ré, and Percy Liang - the company operates as an "AI Native Cloud" designed to train, fine-tune, and deploy open-source generative AI models at scale with high performance and optimized unit economics.
together.ai • HQ: San Francisco, United States • Hardware • Events & Ticketing • Corporate Events • AI Infrastructure • Artificial Intelligence • 501-1000 employees • Est. 2022
HQ: San Francisco, United States • Hardware • Events & Ticketing • Corporate Events • AI Infrastructure • Artificial Intelligence • 501-1000 employees • Est. 2022
Top 25% payVerified live · 3 hours ago 2 months ago

Research Engineer, Large-Scale Training

$200k – $290k per year • In office (San Francisco, United States) • Full-Time
Python
AI/ML
Fine-tuning
PyTorch
Together AI
CUDA
CUDA Toolkit
DeepSpeed
Mamba
Triton
FSDP
Megatron-LM
NCCL
Apply
Report

Procore Technologies

procore.com
Procore is a software development company disrupting the construction industry by providing cloud-based construction software to clients in 150+ countries with a vast user base of 2 million+ across the globe.
procore.com • Bengaluru • AI Agents • Artificial Intelligence • Project Management • Property Development • Software • Est. 2002
Bengaluru • AI Agents • Artificial Intelligence • Project Management • Property Development • Software • Est. 2002
2 months ago

Lead AI Engineer

$38k – $90k per year (Estimated) • Lead • In office (Bengaluru, India) • Internship • Staff
C++
Python
C++
PyTorch C++
AI/ML
CLIP
Computer Vision
ControlNet
CUDA Toolkit
DeepSpeed
Diffusion Models
JAX
Knowledge Distillation
LLaVA
LoRA
Multimodal AI
ONNX
PyTorch
PyTorch Lightning
Ray
Synthetic Data
TensorRT
Stable Diffusion
PEFT
Edge AI
Robotics
Isaac Sim
MoveIt
Nav2
Open3D
Perception
ROS
ROS2
Apply
Report

ConveGenius

convegenius.com
ConveGenius is an education management company that provides services related to sustainable education for K-10 categories through impact-driven programs, an adaptive learning platform, and data dashboards.
convegenius.com • HQ: Noida, India • Primary Education • Est. 2013
HQ: Noida, India • Primary Education • Est. 2013
2 months ago

ML Engineer

$29k – $121k per year (Estimated)In office (Noida, India)
DeepSpeed
Fine-tuning
LoRA
PEFT
PyTorch
QLoRA
RLHF
Transformers
TRL
DPO
Hugging Face
Megatron-LM
SFT
Apply
$28k – $118k per year (Estimated)In office (Noida, India)
DeepSpeed
Fine-tuning
LoRA
PEFT
Perplexity
PyTorch
QLoRA
RLHF
Transformers
TRL
DPO
FSDP
Hugging Face
Megatron-LM
PPO
SFT
Apply
2 months ago

AI Platform Engineer

$26k – $109k per year (Estimated)In office (Noida, India)
Databricks
AI/ML
AWS Bedrock
DeepSpeed
Fine-tuning
LangSmith
LLM
LoRA
NLP
PyTorch
QLoRA
RAG
RLHF
vLLM
PEFT
Amazon SageMaker
DPO
Megatron-LM
AI Agents
Function Calling
TGI
DevOps
Amazon EKS
AWS
Azure
CI/CD
Docker
GCP
Google GKE
Kubernetes
OpenTelemetry
Vector
Apply
Report

Avomind

avomind.com
Avomind is a platform connecting graduates and senior alumni from leading academic institutions with fast-growing firms. Our mission is to connect high-caliber candidates with impactful opportunities.... Our global network is built up of partnersh...
avomind.com • HQ: London, United Kingdom • Professional Services • Human Resources • 501-1000 employees • Est. 2019
HQ: London, United Kingdom • Professional Services • Human Resources • 501-1000 employees • Est. 2019
Verified live · 3 hours ago 2 months ago

Principal Machine Learning Engineer

$119k – $284k per year (Estimated) • Lead • Singapore • Remote (Singapore) • Full-Time • Principal
DeepSpeed
Fine-tuning
JAX
Knowledge Distillation
LoRA
PyTorch
QLoRA
Quantization
Ray
PEFT
DPO
FSDP
Megatron-LM
SFT
Diffusion Models
LLM
Multimodal AI
Reinforcement Learning
RLHF
Spark
TensorRT
TensorRT-LLM
vLLM
PPO
Apply
Report

Mozn

mozn.ai
MOZN is an enterprise AI company that has helped 100+ organizations make critical and informed decisions through specialized AI, in two key areas: Financial Crime Prevention and Enterprise Knowledge Intelligence
mozn.ai • Riyadh • Cairo • Lagos • Dubai • Tunis • Law Enforcement • Fraud Detection • Artificial Intelligence
Riyadh • Cairo • Lagos • Dubai • Tunis • Law Enforcement • Fraud Detection • Artificial Intelligence
Verified live · 1 day ago 2 months ago

AI Infrastructure Engineer III

Middle • 4+ years expCairo • Remote (Egypt) • Full-Time • Middle
Python
Databases
OpenSearch
AI/ML
CUDA
CUDA Toolkit
KServe
Kubeflow
MLFlow
Ray
Ray Serve
Triton
Triton Inference Server
DeepSpeed
JAX
LLM
PyTorch
RAG
TensorFlow
Hugging Face
NCCL
DevOps
Ansible
AWS
Azure
CI/CD
GCP
GitOps
Grafana
Helm
Kubernetes
OpenTelemetry
Platform Engineering
Prometheus
Terraform
Vector
Apply
Report

Chattermill

chattermill.com
Chattermill is the AI-native customer intelligence platform that unifies fragmented feedback, support, reviews and social data to proactively surface the insights you need to act on first.
chattermill.com • London • Ride Hailing • Artificial Intelligence • Data & Analytics
London • Ride Hailing • Artificial Intelligence • Data & Analytics
$96k – $199k per year (Estimated) • Senior • Remote/Hybrid (London, United Kingdom) • Visa sponsorship • Bachelor's Degree • Cofounder • Senior
DeepSpeed
Embeddings
Fine-tuning
LLM
LoRA
PyTorch
Reranking
Semantic Search
Sentiment Analysis
vLLM
PEFT
Semantic Search
AI Agents
Reinforcement Learning
Apply
Verified live · 1 day ago 2 months ago

Senior Machine Learning Engineer

$98k – $200k per year (Estimated) • Senior • Remote (United Kingdom) • Master's Degree • Full-Time • Senior
DeepSpeed
Embeddings
Fine-tuning
LLM
LoRA
PyTorch
Reranking
Semantic Search
Sentiment Analysis
vLLM
PEFT
Semantic Search
AI Agents
Reinforcement Learning
Apply
Open 145 daysVerified live · 1 day ago 5 months ago

Senior Machine Learning Scientist

$98k – $200k per year (Estimated) • Senior • Remote (United Kingdom) • Master's Degree • Full-Time • Senior
DeepSpeed
Embeddings
Fine-tuning
LLM
LoRA
PyTorch
Reranking
Semantic Search
Sentiment Analysis
vLLM
PEFT
Semantic Search
AI Agents
Reinforcement Learning
Apply
Report

ZenteiQ

zenteiq.ai
ZenteiQ is a deep technology company headquartered in Bengaluru, India, and founded in 2022 out of research at the Indian Institute of Science. The company works on scientific machine learning, combining numerical simulation with neural networks for engineering and industrial modelling problems, and builds training programmes around those methods. It sells to industrial and research customers in India while running education initiatives that push scientific computing skills into the wider engineering workforce.
zenteiq.ai • HQ: Bengaluru, India • Predictive Analytics • Materials Science & Nanotechnology • Science & Engineering • EdTech Platforms • Education • Machine Learning • Artificial Intelligence • Est. 2022
HQ: Bengaluru, India • Predictive Analytics • Materials Science & Nanotechnology • Science & Engineering • EdTech Platforms • Education • Machine Learning • Artificial Intelligence • Est. 2022
2 months ago

ML / AI Engineer

$31k – $128k per year (Estimated)In office (Bengaluru, India)
Python
Python
pySpark
AI/ML
Computer Vision
CUDA Toolkit
Fine-tuning
NLP
PyTorch
RLHF
Spark
TPU
DeepSpeed
DPO
FSDP
DevOps
Platform Engineering
HPC
Apply
Report

DeepL

deepl.com
DeepL is a German artificial intelligence company founded in Cologne in 2017 that built a machine translation service widely judged more natural than the incumbents. It grew out of the Linguee bilingual dictionary and trained its own neural models on that curated corpus, releasing a product that handles nuance, register and idiom well enough to displace human first drafts in many professional settings. The company now serves more than thirty languages plus the DeepL Write editing assistant, DeepL Voice for live speech, and a developer API, selling largely to enterprises in regulated European industries that need translation without sending text to American providers.
deepl.com • HQ: Cologne, Germany • Language Learning • LLM & Generative AI • Professional Services • Artificial Intelligence • Natural Language Processing • Translation Services • 501-1000 employees • Est. 2017
HQ: Cologne, Germany • Language Learning • LLM & Generative AI • Professional Services • Artificial Intelligence • Natural Language Processing • Translation Services • 501-1000 employees • Est. 2017
Verified live · 3 hours ago 2 months ago

Senior Research Scientist | Model Steering

$98k – $204k per year (Estimated) • Senior • Remote/Hybrid (London, United Kingdom, Munich, Cologne, Germany) • Full-Time • Senior
Python
AI/ML
Fine-tuning
JAX
Knowledge Distillation
LLM
Multimodal AI
PyTorch
Reinforcement Learning
RLHF
Synthetic Data
TensorFlow
DPO
Edge AI
Human-in-the-Loop
Post-training
PPO
Pre-training
SFT
DeepSpeed
NLP
SGLang
TensorRT
TensorRT-LLM
vLLM
FSDP
Megatron-LM
Apply
Report

Zoox

zoox.com
Zoox is an American technology company headquartered in Foster City, California, that was founded in 2014. The company develops a purpose-built, fully autonomous electric vehicle designed specifically for ride-hailing services rather than individual ownership. Operating as an independent subsidiary of Amazon, the firm is currently testing its robotaxi fleet in several major United States cities including San Francisco and Las Vegas.
zoox.com • HQ: Foster City, United States • Robotics • Smart Mobility • Automotive Manufacturing • Manufacturing • Transportation & Logistics • Artificial Intelligence • Ride Hailing • Autonomous Driving • Est. 2014
HQ: Foster City, United States • Robotics • Smart Mobility • Automotive Manufacturing • Manufacturing • Transportation & Logistics • Artificial Intelligence • Ride Hailing • Autonomous Driving • Est. 2014
$182k – $257k per year • Equity • Junior • 2+ years exp • Remote (United States) • Visa sponsorship • Bachelor's Degree • Junior
DeepSpeed
JAX
PyTorch
Ray
Reinforcement Learning
Edge AI
DevOps
AWS
Robotics
Reinforcement Learning
Apply
Top 25% payVerified live · 3 hours ago 2 months ago

Software Engineer, ML Platform (ML Training)

$192k – $257k per year • Junior • 2+ years expRemote/Hybrid (Foster City, United States) • Full-Time • Junior
DeepSpeed
JAX
PyTorch
Ray
Reinforcement Learning
Edge AI
DevOps
AWS
Apply
Top 25% payVerified live · 3 hours ago 5 months ago

Senior AI Inference Engineer - Model Optimization & Deployment

$225k – $305k per year • Senior • Remote/Hybrid (Foster City, United States) • Full-Time • Senior
C++
Python
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
LLM
LoRA
Multimodal AI
ONNX
PyTorch
QLoRA
Quantization
Ray
TensorRT
TensorRT-LLM
VLM
PEFT
Megatron-LM
Robotics
Sensor Fusion
Apply
Report

Centific

centific.com
Centific is a global digital engineering and artificial intelligence services company headquartered in Redmond, Washington, and established in its current form in 2020. The company provides data-centric AI solutions, including data collection, curation, reinforcement learning from human feedback, and multilingual data services through its OneForma platform. It operates globally with a network of over one million experts and serves enterprise clients in sectors such as technology, healthcare, and retail to build and scale production-grade AI models.
centific.com • HQ: Redmond, United States • Natural Language Processing • Reinforcement Learning • Data Collection • Machine Learning • LLM & Generative AI • Artificial Intelligence • Data & Analytics • 1001-5000 employees • Est. 2020
HQ: Redmond, United States • Natural Language Processing • Reinforcement Learning • Data Collection • Machine Learning • LLM & Generative AI • Artificial Intelligence • Data & Analytics • 1001-5000 employees • Est. 2020
Verified live · 1 day ago 2 months ago

AI Research Engineer- Speech 1

$150k – $160k per year • Junior • 2+ years expIn office (Redmond, United States) • Master's Degree • Full-Time • Junior
Python
Python
FastAPI
AI/ML
Chain-of-Thought
Fine-tuning
LLM
LoRA
Multimodal AI
PyTorch
RAG
Tokenization
Transformers
PEFT
Edge AI
CUDA Toolkit
DeepSpeed
Gemini
Hallucination
ONNX
Quantization
Qwen
TensorRT
Weights & Biases
Whisper
FSDP
Hugging Face
LLM Evaluation
NVIDIA NeMo
DevOps
gRPC
Apply
Report

Grab

grab.com
Grab is a Southeast Asian technology company founded in 2012 by two Harvard Business School classmates as a taxi-hailing app for Malaysia, and now the region's dominant everyday services platform. It operates across eight countries in ride hailing, food and grocery delivery, courier services and digital payments, and famously absorbed Uber's Southeast Asian operations in 2018 in exchange for a stake in the company. Headquartered in Singapore and listed on Nasdaq since 2021, Grab has expanded into financial services through GXS Bank and its lending and insurance arms, and builds its own mapping stack through GrabMaps.
grab.com • HQ: Singapore • Navigation & Mapping • Commerce • Payments • Transportation & Logistics • FinTech • Delivery • Ride Hailing • 5000+ employees • Est. 2012
HQ: Singapore • Navigation & Mapping • Commerce • Payments • Transportation & Logistics • FinTech • Delivery • Ride Hailing • 5000+ employees • Est. 2012
Verified live · 1 hour ago 2 months ago

Lead Machine Learning Engineer (Foundation Models)

$115k – $276k per year (Estimated) • Lead • 8+ years expIn office (Singapore) • Full-Time • Senior
C++
Python
C++
PyTorch C++
AI/ML
DeepSeek
DeepSpeed
Fine-tuning
Knowledge Distillation
Llama
LLM
LoRA
Mistral
Multimodal AI
PyTorch
QLoRA
Qwen
Ray
Reinforcement Learning
RLHF
Tokenization
PEFT
DPO
Edge AI
FSDP
LLMOps
Megatron-LM
Post-training
Pre-training
Recommender Systems
SFT
Mixture of Experts
Apply
Report

Freenome

freenome.com
Freenome is a biotechnology company headquartered in Brisbane, California, and was founded in 2014. The company develops blood-based screening tests that utilize multiomics technology and artificial intelligence to detect cancer in its earliest and most treatable stages. It operates a clinical laboratory certified under CLIA and focuses on expanding access to cancer screening through its SimpleScreen portfolio and large-scale clinical validation studies.
freenome.com • Brisbane • Bioinformatics • Personalized Medicine • Medical AI • Artificial Intelligence • Diagnostics • Molecular Diagnostics • Biotechnology • Health Care • Est. 2014
Brisbane • Bioinformatics • Personalized Medicine • Medical AI • Artificial Intelligence • Diagnostics • Molecular Diagnostics • Biotechnology • Health Care • Est. 2014
Top 25% payVerified live · 1 hour ago 2 months ago

Senior Machine Learning Engineer

$162k – $227k per year • Senior • 5+ years expRemote/Hybrid • Senior
C++
Java
Julia
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
DeepSpeed
MLFlow
PyTorch
Ray
Scikit-learn
Spark
TensorBoard
TensorFlow
Weights & Biases
CUDA Toolkit
Multimodal AI
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
Kubernetes
GitHub
Apply
Report

Mistral AI

mistral.ai
Mistral AI is a French artificial intelligence company founded in Paris in 2023 by former researchers from Google DeepMind and Meta. It builds efficient open-weight and commercial large language models, including the Mistral, Mixtral, Magistral and Codestral families, and distributes them through its own developer platform as well as the major cloud marketplaces. Positioned as Europe's leading frontier model developer, the company also ships the Le Chat assistant, on-premise deployment options for regulated industries and a sovereign compute offering for customers who must keep data inside the region.
mistral.ai • HQ: Paris, France • Code Intelligence • Machine Learning • AI Agents • Artificial Intelligence • LLM & Generative AI • 1001-5000 employees • Est. 2023
HQ: Paris, France • Code Intelligence • Machine Learning • AI Agents • Artificial Intelligence • LLM & Generative AI • 1001-5000 employees • Est. 2023
Verified live · 2 hours ago 2 months ago

Research Engineer, Machine Learning

$148k – $277k per year (Estimated) • Middle • 4+ years expRemote/Hybrid (Palo Alto, San Francisco, United States) • Relocation • Master's Degree • Full-Time • Middle
Python
AI/ML
Accelerate
CUDA
CUDA Toolkit
DeepSpeed
JAX
Multimodal AI
NLP
PyTorch
TensorFlow
FSDP
Pre-training
DevOps
CI/CD
Kubernetes
SLURM
Apply
Verified live · 2 hours ago 2 months ago

Research Engineer, Machine Learning

$53k – $148k per year (Estimated) • Middle • 4+ years expIn office (Paris, France, Zurich, Switzerland, Warsaw, Poland, London, United Kingdom) • Relocation • Master's Degree • Full-Time • Middle
Python
AI/ML
Accelerate
CUDA
CUDA Toolkit
DeepSpeed
JAX
Multimodal AI
NLP
PyTorch
TensorFlow
FSDP
Pre-training
DevOps
CI/CD
Kubernetes
SLURM
Apply
Report

Gina's Tech Jobs

ginastechjobs.com
Gina's Tech Jobs is a technology recruitment agency headquartered in Charlotte, North Carolina, and founded in 1999. The agency places engineers into permanent roles at United States technology employers, with a heavy concentration in machine learning, artificial intelligence, and software engineering positions. It works largely on remote and hybrid searches and represents both established employers and smaller AI companies hiring technical staff.
ginastechjobs.com • HQ: Charlotte, United States • Machine Learning • Artificial Intelligence • IT Consulting • Information Technology • Professional Services • Human Resources • Est. 1999
HQ: Charlotte, United States • Machine Learning • Artificial Intelligence • IT Consulting • Information Technology • Professional Services • Human Resources • Est. 1999
$170k – $200k per year • Lead • San Francisco • Remote (United States) • Full-Time • Senior
DeepSpeed
Diffusion Models
Fine-tuning
JAX
LLM
Multimodal AI
PyTorch
Quantization
Ray
RLHF
Spark
TensorRT
TensorRT-LLM
vLLM
DPO
FSDP
Megatron-LM
PPO
Apply
Report

Kodiak

kodiak.ai
Our purpose-built autonomous solution, the Kodiak Driver, is the AV industry’s most advanced driverless technology, seamlessly integrating into any driving platform for reliable movement in any environment. Learn how we’re shaping the future of ground transportation.
kodiak.ai • HQ: Mountain View, United States • Autonomous Driving • Road Transport • Artificial Intelligence • Est. 2018
HQ: Mountain View, United States • Autonomous Driving • Road Transport • Artificial Intelligence • Est. 2018
Verified live · 5 hours ago 2 months ago

Senior AI Infrastructure Engineer - Model Training

$107k – $224k per year (Estimated) • Senior • 2+ years expIn office (Mountain View, United States) • Visa sponsorship • Bachelor's Degree • Senior
C++
Python
C++
PyTorch C++
AI/ML
CUDA Toolkit
DeepSpeed
Multimodal AI
PyTorch
FSDP
InfiniBand
Megatron-LM
NCCL
NVLink
Apply
Report
DeepSpeed Jobs - Remote & On-site
Frequently asked questions
DeepSpeed: How many jobs are available now?
There are 129 DeepSpeed job openings listed on Alion right now; listings are updated daily.
DeepSpeed: What is the typical salary for DeepSpeed roles?
The typical DeepSpeed role on Alion averages $278,569 USD annually.
DeepSpeed: Are remote, relocation, or visa sponsorship options available?
DeepSpeed listings feature remote, hybrid, and on-site roles, and some companies may offer relocation or visa sponsorship-check each listing for specifics.
DeepSpeed: How do I apply for DeepSpeed jobs on Alion?
Create an account on Alion, upload your resume, and apply directly via the DeepSpeed job listings; your application will be sent to the employer.