Updated Aug 31, 2026

SGLang Jobs - Remote & On-site

SGLang roles at vetted startups and product companies. Listings updated daily and include remote-friendly positions and jobs with transparent salary information. Browse and apply today.

Open positions
206
Companies
99
Salary range
$196K – $431K
Average salary
$296K
AI/ML
Data Science
Backend
Frontend
Mobile
DevOps
Web3
Games
Hardware
Robotics
Security
QA
Executive
Networking
Product
Design
Analytics
Support
Enterprise Apps
Quantum

JPMorganChase

jpmorganchase.com
JPMorgan Chase & Co. is a leading global financial services firm and the largest banking institution in the United States by assets. Headquartered in New York City, the company offers a comprehensive range of financial solutions, including investment banking, asset management, treasury services, and commercial banking. Through its widely recognized consumer division, Chase, it delivers retail banking, credit card, and mortgage services to tens of millions of households across the globe.
jpmorganchase.com • HQ: London, United Kingdom • Blockchain & Crypto • Clinical Research • Financial Services • 1001-5000 employees
HQ: London, United Kingdom • Blockchain & Crypto • Clinical Research • Financial Services • 1001-5000 employees
Verified live · 1 hour ago 1 day ago

Lead Software Engineer - Python / Go & AI/ML

Lead • In office • PhD • Staff
Python
AI/ML
AWQ
GPTQ
LLM
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
DevOps
AWS
Chaos Engineering
Kubernetes
Apply
Verified live · 1 hour ago 5 days ago

Principal Software Engineer - LLM Optimization

$160k – $343k per year (Estimated) • Lead • In office (Jersey City, United States) • PhD • Principal
AWQ
GPTQ
LLM
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
Human-in-the-Loop
LLM Guardrails
AI Agents
DevOps
Amazon EKS
AWS
Chaos Engineering
Kubernetes
Apply
Verified live · 1 hour ago 7 days ago

Lead Machine Learning Engineer

$178k – $381k per year (Estimated) • Lead • In office (New York, United States) • Bachelor's Degree • Staff
Python
AI/ML
LLM
PyTorch
Ray
SGLang
Spark
TensorFlow
CUDA
CUDA Toolkit
Kubeflow
DevOps
Kubernetes
Amazon ECS
Apply
Report

AZX

azx.io
AZX is an artificial intelligence engineering enterprise and Public Benefit Corporation specializing in AI transformation and custom software solutions for critical infrastructure sectors. Headquartered in Bellevue, Washington, the company designs, builds, and deploys production-grade AI systems, agentic workflows, and custom inference platforms tailored for asset-heavy industries - including energy, power utilities, commercial real estate, and supply chain logistics.
azx.io • HQ: Bellevue, United States • Transportation & Logistics • Supply Chain • Energy & Utilities • Artificial Intelligence
HQ: Bellevue, United States • Transportation & Logistics • Supply Chain • Energy & Utilities • Artificial Intelligence
Top 25% payVerified live · 1 hour ago 1 day ago

Senior Software Engineer (AI Inference & Runtime Platform)

$140k – $225k per year • Senior • 5+ years expSeattle • Remote (United States) • Bachelor's Degree • Full-Time • Senior
C++
Go
Python
Rust
Python
FastAPI
Databases
Neo4j
pgvector
Qdrant
PostgreSQL
AI/ML
LLM
SGLang
vLLM
Knowledge Graph
AI Agents
DevOps
AWS
Azure
Bicep
GCP
Karpenter
KEDA
Kubernetes
OpenTelemetry
OpenTofu
Terraform
Vector
Cybersecurity
Least Privilege
Apply
Report

Baseten

baseten.co
Baseten is an AI infrastructure platform designed to help developers and machine learning teams deploy, serve, and scale open-source and custom AI models. The platform provides performant, low-latency inference infrastructure alongside developer tools like Truss, an open-source model packaging framework. Headquartered in San Francisco, California, Baseten enables companies to run state-of-the-art models in production seamlessly without managing underlying cloud infrastructure.
baseten.co • HQ: San Francisco, United States • Hardware • Machine Learning • AI Infrastructure • Artificial Intelligence
HQ: San Francisco, United States • Hardware • Machine Learning • AI Infrastructure • Artificial Intelligence
Top 25% payVerified live · 4 hours ago 2 days ago

Technical Program Manager, Model Performance

$165k – $330k per year • Remote/Hybrid (San Francisco, New York, United States, Toronto, Montreal, Canada) • Full-Time
Cursor
DeepSeek
LLM
SGLang
TensorRT
TensorRT-LLM
vLLM
Management
Notion
Apply
Top 25% payVerified live · 4 hours ago 13 days ago

Forward Deployed Engineer (Training)

$200k – $400k per year • Junior • 1+ year expRemote/Hybrid (San Francisco, New York, United States) • Full-Time • Junior
Cursor
JAX
LLM
PyTorch
Ray
SGLang
TensorRT
TensorRT-LLM
vLLM
InfiniBand
Post-training
SFT
DevOps
Kubernetes
SLURM
Chips/EDA
PoC Library
Apply
28 days ago

Solutions Architect

$189k – $363k per year (Estimated) • Staff+ • San Francisco • Remote (United States) • Visa sponsorship • Bachelor's Degree • Architect
Cursor
Embeddings
SGLang
Management
Notion
Apply
Report

ElevenLabs

elevenlabs.io
ElevenLabs is transforming human-technology interaction with AI research and products. Powering the best enterprises, creators, and developers. From ElevenAgents for customer experience, ElevenCreative for content creation, to the leading AI voice generator.
elevenlabs.io • HQ: New York, United States • HQ: London, United Kingdom • AI Agents • LLM & Generative AI • Artificial Intelligence • Speech & Audio AI • 51-200 employees • Est. 2022
HQ: New York, United States • HQ: London, United Kingdom • AI Agents • LLM & Generative AI • Artificial Intelligence • Speech & Audio AI • 51-200 employees • Est. 2022
Verified live · 2 hours ago 4 days ago

Research Engineer - Inference

$105k – $220k per year (Estimated) • Remote (United States, United Kingdom, Bulgaria, Poland) • Full-Time
CUDA
CUDA Toolkit
ElevenLabs
Knowledge Distillation
Quantization
SGLang
TensorRT
Triton
vLLM
DevOps
GitHub
Apply
Report

eBay

ebay.com
eBay is a multinational e-commerce corporation that connects buyers and sellers in more than 190 markets worldwide thus enabling economic opportunity for individuals, entrepreneurs, businesses, and organizations.
ebay.com • HQ: San Jose, United States • E-commerce Platforms • Commerce • Marketplaces • 51-200 employees • Est. 1995
HQ: San Jose, United States • E-commerce Platforms • Commerce • Marketplaces • 51-200 employees • Est. 1995
$32k – $80k per year (Estimated) • Senior • 5+ years expIn office (Bengaluru, India) • Senior
Go
Python
Rust
AI/ML
Computer Vision
CUDA
CUDA Toolkit
LLM
PyTorch
Quantization
SGLang
TensorFlow
TensorRT
Triton
vLLM
DevOps
HPC
Apply
Report

Prime Intellect

primeintellect.com
Prime Intellect is an artificial intelligence infrastructure company headquartered in San Francisco, California, and founded in 2023. The company provides a decentralized platform for training, evaluating, and deploying large-scale AI models, featuring tools for reinforcement learning, agent development, and a global compute marketplace. It operates globally by aggregating computing resources from various providers to enable researchers and developers to build open-source models and autonomous agents.
primeintellect.com • HQ: San Francisco, United States • Cloud Computing • Information Technology • LLM & Generative AI • Reinforcement Learning • AI Agents • Artificial Intelligence • Hardware • AI Infrastructure • Est. 2023
HQ: San Francisco, United States • Cloud Computing • Information Technology • LLM & Generative AI • Reinforcement Learning • AI Agents • Artificial Intelligence • Hardware • AI Infrastructure • Est. 2023
Top 25% payVerified live · 2 days ago 1 month ago

Applied Research - RL & Agents

$150k – $300k per year • Equity • In office (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time
JavaScript
TypeScript
Databases
Databricks
AI/ML
AI Agents
DSPy
LangChain
LangGraph
OpenRouter
Perplexity
Ray
Reinforcement Learning
RLHF
SGLang
Together AI
vLLM
GRPO
LLM Evaluation
OpenAI
Post-training
SFT
Function Calling
Model Context Protocol
LLM
Synthetic Data
Frontend
Next.js
React.js
DevOps
Cloudflare
Datadog
Grafana
Prometheus
Docker
Kubernetes
Terraform
Management
Zapier
Apply
Top 25% payVerified live · 2 days ago 1 month ago

Member of Technical Staff - Training Platform

$150k – $300k per year • Staff+ • In office (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time • Staff
Python
TypeScript
JavaScript
Python
FastAPI
SQLAlchemy
Databases
Databricks
AI/ML
Fine-tuning
LangChain
LLM
LoRA
OpenRouter
Perplexity
QLoRA
RLHF
SGLang
TensorRT
TensorRT-LLM
Together AI
vLLM
PEFT
NCCL
NVLink
OpenAI
Post-training
SFT
Function Calling
Frontend
Next.js
React.js
shadcn/ui
Tailwind CSS
tRPC
Radix UI
DevOps
Ansible
Cloudflare
Datadog
GCP
GitOps
Google Cloud Run
Google GKE
Grafana
Helm
KEDA
kubectl
Kubernetes
Loki
OpenTelemetry
Prometheus
Rest API
Terraform
Management
Zapier
Apply
Top 25% payVerified live · 2 days ago 1 month ago

Member of Technical Staff - Inference

$150k – $300k per year • Equity • Staff+ • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • PhD • Full-Time • Staff
Python
C++
Rust
C++
PyTorch C++
Protobuf
Databases
Databricks
Apache Kafka
Redis
AI/ML
CUDA
CUDA Toolkit
LangChain
LLM
OpenRouter
Perplexity
PyTorch
Quantization
SGLang
TensorRT
TensorRT-LLM
Together AI
vLLM
InfiniBand
NCCL
OpenAI
Post-training
SFT
Function Calling
Triton
DevOps
AWS
CI/CD
Cloudflare
Datadog
GCP
Kubernetes
Service Mesh
SLI/SLO/SLA
Ansible
Grafana
gRPC
OpenTelemetry
Prometheus
Terraform
Management
Zapier
Apply
Report

xAI

x.ai
xAI is an American artificial intelligence company founded by Elon Musk in 2023 with the stated goal of building models that help humans understand the universe. It develops the Grok family of large language models, distributes them through a consumer assistant, a developer API and deep integration with the X social platform, and adds image and video generation through Grok Imagine. The company runs its own Colossus supercomputer clusters in Memphis, Tennessee, is headquartered in Palo Alto, California, and merged with X Corp in 2025 to combine model development with a large consumer distribution channel.
x.ai • HQ: Palo Alto, United States • Multimodal AI • AI Agents • Machine Learning • LLM & Generative AI • Artificial Intelligence • 1001-5000 employees • Est. 2023
HQ: Palo Alto, United States • Multimodal AI • AI Agents • Machine Learning • LLM & Generative AI • Artificial Intelligence • 1001-5000 employees • Est. 2023
Open 695 daysTop 25% pay 2 years ago

Software Engineer - Training/Inference (C++)

$440k per year • In office (Palo Alto, United States)
C++
Rust
AI/ML
Grok
Knowledge Distillation
LLM
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
Triton
DevOps
CI/CD
Apply
Top 25% payVerified live · 20 min ago 1 month ago

Member of Technical Staff - RL Inference

$440k per year • Staff+ • In office (Palo Alto, United States) • Staff
C++
Python
Rust
C++
PyTorch C++
AI/ML
CUDA Toolkit
LLM
PyTorch
Quantization
SGLang
vLLM
Apply
Top 25% payVerified live · 20 min ago 4 months ago

Backend Engineer - API

$355k per year • In office (London, United Kingdom)
C++
Go
Rust
Databases
ClickHouse
PostgreSQL
AI/ML
LLM
SGLang
TensorRT
vLLM
DevOps
Docker
gRPC
Kubernetes
Apply
Report

CME Group

cmegroup.com
CME Group is a global financial services corporation and the world's largest financial derivatives exchange, headquartered in Chicago, Illinois, and founded in 1848. The company operates a diverse derivatives marketplace offering futures and options products across various asset classes, including agriculture, energy, metals, interest rates, and equity indexes. It facilitates trading through its electronic platforms and clearinghouse services, serving institutional and individual investors across international markets.
cmegroup.com • HQ: Chicago, United States • Farming & Agriculture • Precious Metals • Oil & Gas • Natural Resources • Risk Management • Capital Markets • Financial Services • 501-1000 employees
HQ: Chicago, United States • Farming & Agriculture • Precious Metals • Oil & Gas • Natural Resources • Risk Management • Capital Markets • Financial Services • 501-1000 employees
Verified live · 1 day ago 5 days ago

Platform Engineer III

$18k – $52k per year (Estimated) • Middle • 5+ years expIn office (Bengaluru, India) • Full-Time • Middle
Go
Python
AI/ML
AI Agents
Gemini
LangChain
LangGraph
LLM
NVIDIA NIM
SGLang
Vertex AI
vLLM
Google ADK
DevOps
ArgoCD
CI/CD
GCP
GitOps
Google GKE
Istio
Kubernetes
OpenTelemetry
Platform Engineering
Service Mesh
Terraform
Apply
Report

Tenstorrent

tenstorrent.com
Tenstorrent is a Canadian semiconductor company founded in 2016 and led by veteran chip architect Jim Keller. It designs artificial intelligence accelerators and RISC-V processors, and sells intellectual property licences as well as hardware, with an open-source software stack. Investors and licensees include major automotive and technology groups seeking an alternative to closed accelerator ecosystems.
tenstorrent.com • HQ: Toronto, Canada • Embedded Systems • Hardware • Machine Learning • Est. 2016
HQ: Toronto, Canada • Embedded Systems • Hardware • Machine Learning • Est. 2016
Top 25% payVerified live · 11 hours ago 25 days ago

Staff Forward Deployed Engineer

$100k – $500k per year • Staff+ • 5+ years expIn office (Austin, United States) • Full-Time • Staff
LLM
SGLang
vLLM
AI Agents
Edge AI
DevOps
Grafana
Helm
Kubernetes
OpenTelemetry
Prometheus
HPC
Apply
Report

DoorDash

doordash.com
DoorDash is an American technology company headquartered in San Francisco that operates a leading global local commerce and on-demand delivery platform. The company connects consumers with local merchants, offering last-mile logistics for restaurant meals, groceries, convenience goods, and retail items delivered by independent couriers. Beyond its core consumer marketplace, DoorDash provides digital ordering software, merchant advertising tools, and white-label fulfillment services for businesses across multiple countries.
doordash.com • HQ: San Francisco, United States • Commerce • Marketplaces • Transportation & Logistics • Delivery
HQ: San Francisco, United States • Commerce • Marketplaces • Transportation & Logistics • Delivery
$4k – $14k per year • Equity • Senior • 6+ years expIn office (San Francisco, United States) • Bachelor's Degree • Senior
Python
AI/ML
Claude
Claude Code
Cursor
DeepSeek
Embeddings
Fine-tuning
LLM
LoRA
Qwen
Reinforcement Learning
RLHF
PEFT
DPO
LLM Guardrails
OpenAI Codex
Post-training
SFT
AI Agents
AWQ
GPTQ
Quantization
RAG
SGLang
TensorRT
TensorRT-LLM
vLLM
Model Context Protocol
DevOps
AWS
GCP
Kubernetes
Vector
Apply
Report

Near

near.ai
AI agents for the work your team has been afraid to automate. Private by default, verifiable by design.
near.ai • San Francisco • AI Agents • Artificial Intelligence • AI Infrastructure
San Francisco • AI Agents • Artificial Intelligence • AI Infrastructure
Verified live · 8 hours ago 1 month ago

LLM Inference Engineer

$138k – $301k per year (Estimated)San Francisco • Remote (United States)
CUDA
CUDA Toolkit
LLM
PyTorch
SGLang
TensorRT
Triton
vLLM
Apply
Report

Cloudflare

cloudflare.com
Cloudflare is a global web infrastructure and cybersecurity company that provides content delivery network (CDN) services, DDoS mitigation, and edge computing solutions. The platform acts as a protective shield and performance booster between website visitors and origin servers, safeguarding online applications from cyber threats while optimizing speed. Headquartered in San Francisco, it secures and accelerates millions of internet properties and handles a significant portion of all global web traffic.
cloudflare.com • HQ: San Francisco, United States • Information Technology • Cloud Computing • Email Security • Network Security • Cybersecurity • CDN • 1001-5000 employees • Est. 2009
HQ: San Francisco, United States • Information Technology • Cloud Computing • Email Security • Network Security • Cybersecurity • CDN • 1001-5000 employees • Est. 2009
Verified live · 1 hour ago 1 month ago

Senior Machine Learning Engineer

$170k – $329k per year (Estimated) • Senior • Remote/Hybrid • Internship • Senior
Python
AI/ML
Quantization
Embeddings
JAX
llama.cpp
LLM
Multimodal AI
ONNX
PyTorch
SGLang
TensorFlow
TensorRT
TensorRT-LLM
vLLM
Triton
Galileo
RAG
DevOps
Cloudflare
Apply
Report

RapidClaims

rapidclaims.ai
RapidClaims is a revenue cycle management company leveraging advanced AI and LLMs to enable autonomous medical coding and workflow automation for modernizing and scaling healthcare billing operations.
rapidclaims.ai • Bengaluru • Billing & Invoicing • Est. 2023
Bengaluru • Billing & Invoicing • Est. 2023
$31k – $78k per year (Estimated) • Senior • 5+ years expIn office (Bengaluru, India) • Senior
Python
Databases
ArangoDB
Neo4j
AI/ML
Braintrust
Fine-tuning
Function Calling
Hybrid Search
Langfuse
LangSmith
Llama
LLM
LoRA
Prompt Engineering
PyTorch
QLoRA
Qwen
RAG
Reranking
SGLang
TensorRT
TensorRT-LLM
vLLM
PEFT
DPO
Hugging Face
Human-in-the-Loop
Knowledge Graph
SFT
Structured Outputs
AI Agents
Arize Phoenix
Model Context Protocol
DevOps
Vector
Apply
Report

Crusoe

crusoeenergy.com
Crusoe (formerly Crusoe Energy Systems) is an energy-first AI infrastructure and cloud computing company headquartered in Denver, Colorado. The company specializes in building and operating high-performance AI data centers powered by stranded, wasted, or underutilized energy sources - such as flared natural gas from oil fields and surplus renewable power.
crusoeenergy.com • HQ: Denver, United States • Hardware • AI Infrastructure • Information Technology • Data Centers • 1001-5000 employees
HQ: Denver, United States • Hardware • AI Infrastructure • Information Technology • Data Centers • 1001-5000 employees
Verified live · 11 hours ago 7 days ago

Staff Applied AI Inference Engineer

$185k – $225k per year • Equity • Staff+ • In office (Denver, United States) • Bachelor's Degree • Full-Time • Staff
C++
Python
AI/ML
CUDA
CUDA Toolkit
LLM
SGLang
vLLM
DevOps
Docker
Kubernetes
Apply
Verified live · 11 hours ago 1 month ago

Staff Applied AI Inference Engineer

$215k – $260k per year • Equity • Staff+ • In office (San Francisco, Sunnyvale, United States) • Bachelor's Degree • Full-Time • Staff
C++
Python
AI/ML
CUDA Toolkit
LLM
SGLang
vLLM
CUDA
DevOps
Docker
Kubernetes
Apply
Report

OpenAI

openai.com
OpenAI is an American artificial intelligence research and deployment company founded in 2015 with the mission of ensuring that artificial general intelligence benefits all of humanity. It develops the GPT family of large language models and turns them into consumer and developer products, including the ChatGPT assistant, the Sora video model, the Codex coding agent and a commercial API used by millions of developers. Structured as a public benefit corporation controlled by a non-profit foundation, the company is backed by Microsoft and SoftBank and operates from San Francisco.
openai.com • HQ: San Francisco, United States • Multimodal AI • AI Agents • Artificial Intelligence • Machine Learning • LLM & Generative AI • 1001-5000 employees • Est. 2015
HQ: San Francisco, United States • Multimodal AI • AI Agents • Artificial Intelligence • Machine Learning • LLM & Generative AI • 1001-5000 employees • Est. 2015
Top 25% payVerified live · 7 min ago 7 days ago

Software Engineer, Model Runtime

$266k – $445k per year • Remote/Hybrid (San Francisco, United States) • Full-Time
C++
Python
Rust
AI/ML
LLM
SGLang
vLLM
OpenAI
Apply
Top 25% payVerified live · 7 min ago 1 month ago

Systems Generalist, GPT Infrastructure

$293k – $445k per year • Senior • 8+ years expRemote/Hybrid (San Francisco, Seattle, United States) • Full-Time • Staff
C++
Go
Python
Rust
C++
LLVM
AI/ML
OpenAI
CUDA
CUDA Toolkit
SGLang
Triton
Triton Inference Server
vLLM
ROCm
Apply
Open 125 daysTop 25% pay 4 months ago

Software Engineer, GPT Infrastructure

$293k – $385k per year • Remote/Hybrid (San Francisco, Seattle, United States) • Full-Time
C++
Go
Python
Rust
C++
LLVM
AI/ML
OpenAI
CUDA Toolkit
SGLang
Triton Inference Server
vLLM
LLM Evaluation
ROCm
Apply
Report

Furiosa

furiosa.ai
FuriosaAI designs high-performance, power-efficient AI accelerators (NPUs) used in data centers for computer vision, GenAI, LLMs, and demanding workloads.
furiosa.ai • Seoul • Hwaseong • Santa Clara • San Jose • Data Centers • LLM & Generative AI • Artificial Intelligence
Seoul • Hwaseong • Santa Clara • San Jose • Data Centers • LLM & Generative AI • Artificial Intelligence
Verified live · 5 min ago 10 days ago

Solutions Architect - US

$158k – $288k per year (Estimated) • Staff+ • In office (Santa Clara, United States) • Architect
Python
C++
Rust
C++
PyTorch C++
TensorFlow C++
AI/ML
AutoGen
LangChain
LangGraph
LlamaIndex
LLM
PyTorch
SGLang
TensorFlow
TensorRT
TensorRT-LLM
Triton
Triton Inference Server
vLLM
Model Context Protocol
Quantization
Apply
Verified live · 5 min ago 10 days ago

Senior Solution Architect

Staff+ • In office (Seoul, South Korea) • Architect
Python
C++
Rust
AI/ML
LLM
SGLang
TensorRT
TensorRT-LLM
vLLM
DevOps
Docker
Kubernetes
Chips/EDA
PoC Library
Apply
Senior • 3+ years expIn office (Seoul, South Korea) • Bachelor's Degree • Senior
C++
Rust
AI/ML
LLM
Multimodal AI
SGLang
vLLM
CUDA
CUDA Toolkit
TensorRT
TensorRT-LLM
Triton
Apply
Report

Innovaccer

innovaccer.com
Innovaccer is a healthcare technology company headquartered in San Francisco, California, and founded in 2014. The company provides a healthcare intelligence platform that unifies clinical, financial, and operational data to enable AI-powered workflows, population health management, and revenue cycle optimization. It serves health systems, payers, and life sciences organizations across the United States and internationally, utilizing autonomous AI agents to improve patient outcomes and administrative efficiency.
innovaccer.com • HQ: San Francisco, United States • Business Process Automation (BPA) • Software • Analytics Platforms • Medical AI • Artificial Intelligence • Health Care • Data & Analytics • Digital Health • 501-1000 employees • Est. 2014
HQ: San Francisco, United States • Business Process Automation (BPA) • Software • Analytics Platforms • Medical AI • Artificial Intelligence • Health Care • Data & Analytics • Digital Health • 501-1000 employees • Est. 2014
Verified live · 1 day ago 8 days ago

Senior Artificial Intelligence Researcher

$163k – $317k per year (Estimated) • Senior • In office (San Francisco, United States) • Bachelor's Degree • Full-Time • Senior
Python
AI/ML
AI Agents
DeepSpeed
Fine-tuning
LLM
PyTorch
RAG
SGLang
vLLM
FSDP
Hugging Face
Apply
Verified live · 1 day ago 8 days ago

Principal AI Researcher

$164k – $352k per year (Estimated) • Lead • In office (San Francisco, United States) • Bachelor's Degree • Full-Time • Principal
Python
AI/ML
AI Agents
DeepSpeed
Fine-tuning
LLM
PyTorch
RAG
SGLang
vLLM
FSDP
Hugging Face
Post-training
Apply
Verified live · 1 day ago 8 days ago

Artificial Intelligence Researcher

$158k – $345k per year (Estimated)In office (San Francisco, United States) • Bachelor's Degree • Full-Time
Python
AI/ML
AI Agents
DeepSpeed
Fine-tuning
LLM
PyTorch
RAG
SGLang
vLLM
FSDP
Hugging Face
Apply
Report

Thinking Machines Lab

thinkingmachines.ai
Thinking Machines Lab is an artificial intelligence research and product company based in San Francisco and founded in 2025. The company develops multimodal AI systems and open-weights models, such as Inkling, alongside developer tools like Tinker for model fine-tuning. It operates as a public benefit corporation focused on human-AI collaboration and open science, supported by significant venture capital investment.
thinkingmachines.ai • HQ: San Francisco, United States • MLOps • Machine Learning • Multimodal AI • LLM & Generative AI • Artificial Intelligence • 201-500 employees • Est. 2025
HQ: San Francisco, United States • MLOps • Machine Learning • Multimodal AI • LLM & Generative AI • Artificial Intelligence • 201-500 employees • Est. 2025
Top 25% payVerified live · 3 hours ago 10 days ago

Research, RL Scaling

$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • PhD • Full-Time
Clarity
Python
AI/ML
JAX
PyTorch
Quantization
Reinforcement Learning
TensorFlow
Mixture of Experts
LLM
SGLang
vLLM
AI Agents
Apply
Top 25% payVerified live · 3 hours ago 27 days ago

Research Engineer, Infrastructure, Inference

$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • Full-Time
JAX
PyTorch
Ray
SGLang
vLLM
Edge AI
DeepSpeed
DevOps
Kubernetes
SLURM
Apply
Top 25% payVerified live · 3 hours ago 27 days ago

Software Engineer, Developer Productivity, AI Tools

$350k – $475k per year • Remote/Hybrid (San Francisco, United States) • Relocation • Visa sponsorship • Bachelor's Degree • Full-Time
Python
Rust
AI/ML
Claude
Claude Code
Cursor
OpenAI Codex
SGLang
vLLM
TGI
DevOps
Buildkite
Docker
GitHub Actions
Kubernetes
GitHub
Apply
Report

NVIDIA

nvidia.com
NVIDIA is an American technology company founded in 1993 that invented the graphics processing unit and has become the dominant supplier of accelerated computing platforms for artificial intelligence. Its portfolio spans data centre GPUs and systems built on the Hopper and Blackwell architectures, GeForce consumer graphics, automotive and robotics platforms, high-speed networking acquired with Mellanox, and the CUDA software stack that binds the ecosystem together. Headquartered in Santa Clara, California, the company sells to cloud providers, enterprises, research institutions and gamers worldwide and is one of the most valuable listed businesses on the Nasdaq.
nvidia.com • HQ: Santa Clara, United States • AI Infrastructure • Semiconductors • AI Agents • Big Data • LLM & Generative AI • Machine Learning • Hardware • Data & Analytics • Information Technology • Artificial Intelligence • 5000+ employees • Est. 1993
HQ: Santa Clara, United States • AI Infrastructure • Semiconductors • AI Agents • Big Data • LLM & Generative AI • Machine Learning • Hardware • Data & Analytics • Information Technology • Artificial Intelligence • 5000+ employees • Est. 1993
$184k – $288k per year • Senior • 5+ years expIn office (Santa Clara, United States) • Master's Degree • Full-Time • Senior
C++
Python
C++
PyTorch C++
AI/ML
LLM
PyTorch
Ray
Reinforcement Learning
RLHF
AI Agents
DPO
FSDP
GRPO
Post-training
PPO
DeepSpeed
DeepSpeed-Chat
Multimodal AI
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
VLM
InfiniBand
Megatron-LM
Mixture of Experts
NCCL
NVIDIA NeMo
NVLink
Pre-training
DevOps
Kubernetes
Apply
$101k – $255k per year (Estimated) • Senior • 10+ years expIn office (Zurich, Switzerland) • Bachelor's Degree • Full-Time • Senior
Python
AI/ML
CUDA
CUDA Toolkit
LLM
PyTorch
SGLang
TensorRT
TensorRT-LLM
vLLM
NCCL
DevOps
Ansible
Argo Workflows
AWS
Azure
Chef
GCP
Grafana
Kubernetes
KubeVirt
OpenTelemetry
Prometheus
Puppet
Splunk
Terraform
AWS Step Functions
Apply
$152k – $242k per year • Senior • 4+ years expIn office (Redmond, Santa Clara, United States) • Master's Degree • Full-Time • Senior
C++
Python
C++
PyTorch C++
AI/ML
LLM
PyTorch
SGLang
Triton
vLLM
Hugging Face
Megatron-LM
Mixture of Experts
SFT
Apply
Report

TORC Robotics

torc.ai
Torc Robotics is an autonomous vehicle company and an independent subsidiary of Daimler Truck. Founded in 2005 by Virginia Tech engineers, the company specializes in developing Level 4 self-driving technology primarily tailored for long-haul commercial freight. Its core product ecosystem, TorcDrive, integrates AI-driven software, sensor fusion, and fleet command tools directly into production-intent vehicles like the Freightliner Cascadia to commercialize safe, scalable hub-to-hub autonomous trucking.
torc.ai • HQ: Blacksburg, United States • Artificial Intelligence • Road Transport • Autonomous Driving
HQ: Blacksburg, United States • Artificial Intelligence • Road Transport • Autonomous Driving
Verified live · 7 hours ago 3 months ago

Senior, ML Engineer - Auto Tagging

$153k – $274k per year (Estimated) • Equity • Senior • 6+ years expAnn Arbor • Remote (United States) • Relocation • Bachelor's Degree • Full-Time • Senior
Python
SQL
Databases
Databricks
AI/ML
Multimodal AI
Ray
Spark
Time Series Forecasting
SGLang
vLLM
DevOps
AWS
Azure
GCP
Vector
Robotics
Sensor Fusion
ROS
Apply
Verified live · 7 hours ago 21 day ago

Senior, Software Engineer - AutoTagging

$152k – $253k per year (Estimated) • Equity • Senior • 5+ years expAnn Arbor • Remote (United States) • Relocation • Bachelor's Degree • Full-Time • Senior
Python
Databases
Databricks
AI/ML
Multimodal AI
Time Series Forecasting
Ray
SGLang
Spark
vLLM
DevOps
AWS
CI/CD
CloudFormation
Datadog
GitHub Actions
Grafana
Terraform
Amazon CloudWatch
GitHub
Robotics
Sensor Fusion
ROS
Apply
Verified live · 7 hours ago 1 month ago

Senior, ML Engineer - VLM

$151k – $271k per year (Estimated) • Equity • Senior • 6+ years expAnn Arbor • Remote (United States) • Relocation • Bachelor's Degree • Full-Time • Senior
Python
C++
JavaScript
C++
OpenGL
Databases
Databricks
DynamoDB
LanceDB
AI/ML
Computer Vision
Embeddings
MLFlow
Multimodal AI
Pandas
PyTorch
Ray
Spark
VLM
LLM
Semantic Search
SGLang
vLLM
Semantic Search
Frontend
Three.JS
DevOps
Docker
GitHub Actions
GitHub
AWS
AWS Lambda
Terraform
Vector
Amazon ECS
Amazon S3
AWS Step Functions
Robotics
Sensor Fusion
Foxglove Studio
ROS
Chips/EDA
Cadence Pegasus
Apply
Report

Salary range

Seniority Jobs 25% Median 75%
Senior 18 $198K $242K $288K
Staff 22 $260K $285K $330K
All levels 79 $245K $288K $350K

Based only on the listings that state pay. Gross annual amounts, converted to USD so roles in different currencies stay comparable.

Fully remote 15%
31
Hybrid 22%
46
On site 63%
129
Visa sponsorship 10%
21
Relocation package 8%
17
Equity 14%
28

The market right now

Alion currently lists 206 open SGLang jobs from 99 companies. 17 of them were posted or refreshed in the last seven days. 82 of the employers have posted something in the last three months, which is the pool worth watching if you are starting a search now. Every listing links straight to the employer, so you apply on their own board rather than through an intermediary.

What these roles pay

79 of these SGLang jobs state pay directly. Across them the middle half of the market sits between $245K and $350K a year, with a median of $288K. By level, the median runs senior at $242K and staff at $285K. The step from senior to staff is worth about 1.2x on median pay. Figures are gross annual amounts converted to US dollars, so roles in different currencies stay comparable.

Where the work is

The largest concentrations of these roles are China (11), United States (8), Singapore (7), South Korea (6), and Switzerland (3). 15% of the listings are fully remote and a further 22% are hybrid, so a large part of this market is open to you regardless of where you live.

Who is hiring

The employers with the most open SGLang jobs right now are NVIDIA (15), Intel Corporation (11), Baseten (10), Together AI (7), Innovaccer (6), and Fireworks AI (6). Each company page on Alion carries its size, funding stage, tech stack and every other position it has open, so you can judge the employer before you spend an evening on the application.

What employers ask for

Reading across the current listings, the tools that come up most often are SGLang (206), vLLM (194), LLM (166), Python (141), TensorRT (105), PyTorch (99), TensorRT-LLM (91), and Kubernetes (84). The counts are how many of these openings name each one, which is a better guide to what is actually being hired for than a generic skills list.

Relocation, visas and equity

Of the current SGLang jobs, 21 state visa sponsorship, 17 offer a relocation package, and 28 include equity. These are filters on the list above, so you can narrow it to the ones that make a move possible for you.

SGLang Jobs - Remote & On-site
Frequently asked questions
How many SGLang jobs are open right now?
There are 206 open SGLang jobs on Alion from 99 companies. The list was last refreshed on August 31, 2026.
What do SGLang jobs pay?
The median is $288K a year, and the middle half of the market falls between $245K and $350K. This is based on the 79 listings that state pay.
Are any of these roles remote?
31 of the 206 listings (15%) are fully remote, and you can filter the list down to them in one click.
Which companies are hiring?
The most active employers right now are NVIDIA, Intel Corporation, Baseten, Together AI, Innovaccer, and Fireworks AI. Each has its own page on Alion with the rest of its open roles.
Can I get visa sponsorship or relocation?
Yes - of the current listings, 21 state visa sponsorship and 17 offer a relocation package. Both are filters on the jobs page.
How do I apply?
Open any listing and apply on the employer's own board through the link on the page. No account is required to browse or to apply.