SGLang Jobs - Remote & On-site

SGLang roles at vetted startups and product companies. Listings updated daily and include remote-friendly positions and jobs with transparent salary information. Browse and apply today.

AI/ML
Data Science
Backend
Frontend
Mobile
DevOps
Web3
Games
Hardware
Robotics
Security
QA
Executive
Networking
Product
Design
Analytics
Support
Enterprise Apps
Quantum

Cerebras Systems

cerebras.ai
Cerebras Systems is an American computer hardware company founded in 2016 and headquartered in Sunnyvale, California that builds accelerators for artificial intelligence at wafer scale. Instead of assembling clusters from many small chips, it manufactures a single processor the size of an entire silicon wafer, the Wafer Scale Engine, which removes most of the communication overhead in large model training and inference. The company sells CS-series systems to research laboratories and enterprises, operates its own inference cloud known for very high token throughput, and has built large supercomputers with partners including the Gulf technology group G42.
cerebras.ai • HQ: Sunnyvale, United States • Machine Learning • Artificial Intelligence • Hardware • Semiconductors • AI Infrastructure • 1001-5000 employees • Est. 2016
HQ: Sunnyvale, United States • Machine Learning • Artificial Intelligence • Hardware • Semiconductors • AI Infrastructure • 1001-5000 employees • Est. 2016
Verified live · 4 hours ago 2 months ago

Staff Software Engineer, GPU Inference

$230k – $398k per year (Estimated) • Staff+ • 8+ years expRemote/Hybrid (Toronto, Canada, Sunnyvale, United States) • Bachelor's Degree • Full-Time • Staff
C++
Python
C++
PyTorch C++
AI/ML
Cerebras
LLM
Multimodal AI
PyTorch
Quantization
SGLang
TensorRT
TensorRT-LLM
Triton Inference Server
vLLM
Edge AI
OpenAI
ROCm
AI Agents
CUDA Toolkit
Mixture of Experts
DevOps
CI/CD
Kubernetes
Apply
Verified live · 4 hours ago 8 months ago

Senior Product Manager, AI Models

$195k – $416k per year (Estimated) • Senior • 5+ years expRemote/Hybrid (Sunnyvale, United States) • Full-Time • Senior
Python
AI/ML
Cerebras
LLM
PyTorch
Quantization
SGLang
vLLM
AI Agents
Edge AI
Hugging Face
OpenAI
Apply
Verified live · 4 hours ago 10 months ago

Software Engineer, GPU Inference

$237k – $459k per year (Estimated) • Senior • 5+ years expIn office • Bachelor's Degree • Full-Time • Senior
C++
Python
C++
PyTorch C++
AI/ML
Cerebras
LLM
Multimodal AI
PyTorch
SGLang
TensorRT
TensorRT-LLM
vLLM
Quantization
Triton
Triton Inference Server
Edge AI
OpenAI
ROCm
AI Agents
CUDA
CUDA Toolkit
Mixture of Experts
DevOps
CI/CD
Kubernetes
Apply
Report

KIEFER

kiefer.gr
KIEFER is an energy and advanced technology group headquartered in Athens, Greece. The company builds and operates renewable energy projects, principally wind, and runs a technology arm developing Sophea AI, a family of Greek language models, along with robotics and sovereign GPU cloud infrastructure. It is establishing Greece's first sovereign AI data centre, powered by renewable generation, in partnership with NVIDIA.
kiefer.gr • HQ: Athens, Greece • Industrial Robotics • AI Infrastructure • Hardware • Wind Energy • LLM & Generative AI • Robotics • Artificial Intelligence • Energy & Utilities
HQ: Athens, Greece • Industrial Robotics • AI Infrastructure • Hardware • Wind Energy • LLM & Generative AI • Robotics • Artificial Intelligence • Energy & Utilities
Senior • Athens • Remote (Greece) • Relocation • Bachelor's Degree • Full-Time • Senior • Polish: C1 • Greek
Python
AI/ML
Computer Vision
Fine-tuning
LLM
PyTorch
Quantization
RAG
SGLang
TensorRT
vLLM
Edge AI
Pre-training
TGI
Speech Recognition
DevOps
Docker
Apply
Verified live · 12 hours ago 3 months ago

Senior ML Engineer (LLM)

$93k – $186k per year • Senior • Remote (Poland) • Relocation • Full-Time • Senior • Polish: C1 • Greek
Python
AI/ML
Computer Vision
Fine-tuning
LLM
PyTorch
Quantization
RAG
SGLang
TensorRT
Triton
vLLM
Edge AI
Pre-training
TGI
Speech Recognition
DevOps
Docker
Apply
Verified live · 12 hours ago 3 months ago

Senior ML Engineer (LLM)

$93k – $186k per year • Senior • Remote/Hybrid (Athens, Greece) • Relocation • Full-Time • Senior • Greek
Python
AI/ML
Computer Vision
Fine-tuning
LLM
PyTorch
Quantization
RAG
SGLang
TensorRT
Triton
vLLM
Edge AI
Pre-training
TGI
Speech Recognition
DevOps
Docker
Apply
Report

Adagrad AI

adagrad.ai
Adagrad AI is a leading AI solution provider that leverages hardware-accelerated algorithms, enhanced computer vision applications, and advanced tech to develop and deploy scalable solutions for businesses.
adagrad.ai • Mumbai • Bengaluru • Pune • Gurgaon • Hardware • Artificial Intelligence • Computer Vision • Est. 2018
Mumbai • Bengaluru • Pune • Gurgaon • Hardware • Artificial Intelligence • Computer Vision • Est. 2018
$34k – $81k per year (Estimated) • Lead • 7+ years expIn office (Pune, India) • Staff
Python
AI/ML
Computer Vision
CUDA Toolkit
PyTorch
TensorRT
Quantization
SGLang
vLLM
Robotics
GStreamer
Apply
Report

Luma AI

lumalabs.ai
Luma AI is an artificial intelligence company that specializes in building advanced 3D visual technology and generative AI tools. Founded in 2021, the company is best known for its realistic text-to-3D models, NeRF (Neural Radiance Fields) capture app, and its high-quality video generation model, Dream Machine. By enabling creators and developers to easily capture, generate, and edit photorealistic 3D assets and video, it aims to make next-generation digital media creation accessible to everyone.
lumalabs.ai • HQ: Palo Alto, United States • Artificial Intelligence • Generative Image • Generative Video • Est. 2021
HQ: Palo Alto, United States • Artificial Intelligence • Generative Image • Generative Video • Est. 2021
Verified live · 9 hours ago 2 months ago

Tech Lead Manager, Inference

$220k – $432k per year (Estimated) • Lead • Remote/Hybrid (Redwood City, United States) • Full-Time • Staff
Python
C
C++
Rust
C++
PyTorch C++
C
FFmpeg
AI/ML
LLM
PyTorch
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
CUDA
CUDA Toolkit
Multimodal AI
Ray
AWS Trainium
InfiniBand
NVLink
TPU
DevOps
Kubernetes
Apply
Verified live · 9 hours ago 2 months ago

Software Engineer, Inference

$150k – $374k per year (Estimated)Remote/Hybrid (Redwood City, United States) • Full-Time
Python
C
C
FFmpeg
Databases
Redis
AI/ML
LLM
PyTorch
SGLang
TensorRT
TensorRT-LLM
vLLM
Hugging Face
CUDA Toolkit
InfiniBand
NVLink
DevOps
CI/CD
Docker
Kubernetes
Amazon S3
Apply
$154k – $390k per year (Estimated)Remote/Hybrid (Redwood City, United States) • Full-Time
C
C
MPI
AI/ML
LLM
Multimodal AI
PyTorch
Ray
Reinforcement Learning
RLHF
SGLang
TRL
vLLM
Transformers
AI Agents
FSDP
Function Calling
GRPO
LLM Evaluation
NCCL
Post-training
PPO
DevOps
Kubernetes
Apply
Report

LG

lg.com
LG is a major global technology innovator and electronics brand recognized for manufacturing high-quality home appliances, consumer electronics, and display solutions. The multinational conglomerate produces a wide range of everyday products, including OLED televisions, smart refrigerators, washing machines, and commercial HVAC systems. Driven by its core philosophy of "Life's Good," the company continuously advances smart home integration and intuitive technology to elevate consumer living worldwide.
lg.com • Troy • Huntsville • Alpharetta • Englewood • Santa Clara • Professional Services • Hardware • Home Appliances • Manufacturing • Consumer Electronics • Est. 1997
Troy • Huntsville • Alpharetta • Englewood • Santa Clara • Professional Services • Hardware • Home Appliances • Manufacturing • Consumer Electronics • Est. 1997
$146k – $319k per year (Estimated)Remote/Hybrid (Santa Clara, United States) • Master's Degree • Contractor
Python
AI/ML
AI Agents
LLM
Multimodal AI
Quantization
Transformers
VLM
PEFT
Edge AI
Post-training
Mixture of Experts
Fine-tuning
GGUF
Knowledge Distillation
llama.cpp
LoRA
PyTorch
SGLang
TensorRT
TensorRT-LLM
vLLM
DPO
Analytics
ETL/ELT
Apply
Report

Scale AI

scale.com
Scale AI is an American company founded in San Francisco in 2016 that supplies the training data, evaluation and tooling behind large artificial intelligence systems. It began with human-labelled annotation for autonomous driving and computer vision, then expanded into reinforcement learning from human feedback, expert data generation, model evaluation and full-stack deployment platforms for enterprises and governments. In 2025 Meta acquired a large minority stake and hired co-founder Alexandr Wang, after which the company continued under new leadership serving defence, public sector and commercial customers.
scale.com • HQ: San Francisco, United States • Defense AI • Data & Analytics • Data Collection • LLM & Generative AI • Artificial Intelligence • 1001-5000 employees • Est. 2016
HQ: San Francisco, United States • Defense AI • Data & Analytics • Data Collection • LLM & Generative AI • Artificial Intelligence • 1001-5000 employees • Est. 2016
Verified live · 6 min ago 2 months ago

AI Infrastructure Engineer, Serving Platform

$77k – $230k per year (Estimated) • Middle • 4+ years expIn office (London, United Kingdom) • Middle
C++
Go
Python
Rust
AI/ML
LLM
Function Calling
SGLang
TensorRT
TensorRT-LLM
vLLM
DevOps
AWS
Docker
GCP
Kubernetes
Terraform
Apply
Report

Fireworks AI

fireworks.ai
Fireworks AI is an artificial intelligence infrastructure company headquartered in Redwood City, California, and founded in 2022. The company provides a high-performance inference and training platform that enables developers to deploy, fine-tune, and scale open-source generative models with optimized speed and cost. It operates globally as a cloud-based service provider, catering to technology firms and enterprises seeking to integrate specialized intelligence into their applications through a serverless or dedicated API.
fireworks.ai • HQ: Redwood City, United States • Machine Learning • Cloud Computing • Information Technology • MLOps • LLM & Generative AI • Artificial Intelligence • 501-1000 employees • Est. 2022
HQ: Redwood City, United States • Machine Learning • Cloud Computing • Information Technology • MLOps • LLM & Generative AI • Artificial Intelligence • 501-1000 employees • Est. 2022
Verified live · 7 hours ago 2 months ago

AI Field Engineer, Singapore

$157k – $276k per year • Staff+ • 10+ years expIn office (Singapore) • Full-Time • Staff
Python
AI/ML
Fine-tuning
LLM
Multimodal AI
PyTorch
Quantization
SGLang
vLLM
DPO
SFT
AWS Bedrock
Fireworks AI
TensorRT
TensorRT-LLM
Amazon SageMaker
AI Agents
DevOps
AWS
Azure
GCP
Kubernetes
Apply
Verified live · 7 hours ago 2 months ago

AI Field Engineer, EMEA

$128k – $280k per year (Estimated) • Staff+ • 10+ years expIn office (London, United Kingdom) • Full-Time • Staff
Python
AI/ML
Fine-tuning
LLM
Multimodal AI
PyTorch
Quantization
SGLang
vLLM
DPO
SFT
AWS Bedrock
Fireworks AI
TensorRT
TensorRT-LLM
Amazon SageMaker
AI Agents
DevOps
AWS
Azure
GCP
Kubernetes
Apply
Top 25% payVerified live · 7 hours ago 2 months ago

MTS, Research Engineer

$250k – $400k per year • In office (San Mateo, New York, United States) • Master's Degree • Full-Time
C++
Python
Rust
C
C++
PyTorch C++
TensorFlow C++
C
MPI
AI/ML
CUDA Toolkit
JAX
Multimodal AI
PyTorch
TensorFlow
NCCL
Fireworks AI
SGLang
vLLM
Apply
Report

Snowflake

snowflake.com
Snowflake is a cloud software company that helps organisations store, process, and analyse large volumes of data in a way that is designed for modern cloud infrastructure. Many businesses struggle with data spread across different systems, slow reporting cycles, and the cost and complexity of managing traditional data warehouses.
snowflake.com • HQ: Bozeman, United States • Database Management • Data & Analytics • Data Warehousing • 1001-5000 employees
HQ: Bozeman, United States • Database Management • Data & Analytics • Data Warehousing • 1001-5000 employees
Top 25% payVerified live · 1 hour ago 2 months ago

Senior/Staff Software Engineer - Machine Learning Platform (Inference)

$236k – $339k per year • Staff+ • 7+ years expIn office (Menlo Park, United States) • Bachelor's Degree • Full-Time • Staff
Snowflake
AI/ML
LLM
MLFlow
PEFT
PyTorch
SGLang
TensorFlow
TensorRT
TensorRT-LLM
vLLM
XGBoost
Transformers
AI Agents
Apply
Report

DeepL

deepl.com
DeepL is a German artificial intelligence company specializing in advanced neural machine translation and language technologies. The company utilizes proprietary deep learning models trained on vast multilingual datasets to deliver highly accurate, natural-sounding text translations and AI writing assistance. Operating through web interfaces, desktop and mobile applications, and enterprise APIs, DeepL provides secure communication tools for individuals and multinational organizations worldwide.
deepl.com • HQ: Cologne, Germany • Professional Services • Artificial Intelligence • Natural Language Processing • Translation Services • 501-1000 employees • Est. 2017
HQ: Cologne, Germany • Professional Services • Artificial Intelligence • Natural Language Processing • Translation Services • 501-1000 employees • Est. 2017
Verified live · 5 hours ago 2 months ago

Senior Research Scientist | Model Steering

$98k – $204k per year (Estimated) • Senior • Remote/Hybrid (London, United Kingdom, Munich, Cologne, Germany) • Full-Time • Senior
Python
AI/ML
Fine-tuning
JAX
Knowledge Distillation
LLM
Multimodal AI
PyTorch
Reinforcement Learning
RLHF
Synthetic Data
TensorFlow
DPO
Edge AI
Human-in-the-Loop
Post-training
PPO
Pre-training
SFT
DeepSpeed
NLP
SGLang
TensorRT
TensorRT-LLM
vLLM
FSDP
Megatron-LM
Apply
Report

Fuse Energy

fuseenergy.com
Fuse Energy is a British household energy supplier headquartered in London, founded by former Revolut executives. It was the first new domestic energy supplier in Great Britain since the 2021 energy crisis and operates its own renewable generation assets.
fuseenergy.com • HQ: London, United Kingdom • Solar Energy • Energy & Utilities • 501-1000 employees • Est. 2022
HQ: London, United Kingdom • Solar Energy • Energy & Utilities • 501-1000 employees • Est. 2022
2 months ago

AI Inference Engineer

Middle • 4+ years expIn office (Dubai, United Arab Emirates) • Full-Time • Middle
CUDA Toolkit
Knowledge Distillation
LLM
SGLang
TensorRT
TensorRT-LLM
Triton Inference Server
vLLM
DevOps
Kubernetes
SLI/SLO/SLA
SLURM
Web3
Solana
Apply
Verified live · 9 hours ago 2 months ago

AI Inference Engineer

$104k – $217k per year (Estimated)In office (London, United Kingdom) • Full-Time
CUDA
CUDA Toolkit
Knowledge Distillation
LLM
SGLang
TensorRT
TensorRT-LLM
Triton
Triton Inference Server
vLLM
DevOps
Kubernetes
SLI/SLO/SLA
SLURM
Apply
Report

Advantech

advantech.com
Advantech is a leading provider of industrial computing and internet of things platforms. Its products cover embedded boards, industrial servers and edge software. The company serves manufacturing, healthcare and smart city projects.
advantech.com • HQ: Taipei, Taiwan • Hardware • Internet of Things • Artificial Intelligence • 501-1000 employees • Est. 1983
HQ: Taipei, Taiwan • Hardware • Internet of Things • Artificial Intelligence • 501-1000 employees • Est. 1983
In office (Taipei, Taiwan)
C++
Go
AI/ML
AI Agents
LLM
SGLang
vLLM
Model Context Protocol
DevOps
CI/CD
Docker
Git
Gitflow
Kubernetes
Apply
Report

Oracle

oracle.com
Oracle Corporation is an American multinational computer technology corporation headquartered in Austin, Texas. Founded in 1977 by Larry Ellison, Bob Miner, and Ed Oates, Oracle is one of the world's largest enterprise software and cloud computing infrastructure providers.
oracle.com • HQ: Austin, United States • Software • LLM & Generative AI • Information Technology • Data & Analytics • AI Agents • Artificial Intelligence • Cloud Computing • Enterprise Resource Planning (ERP) • Database Management • 1001-5000 employees • Est. 1977
HQ: Austin, United States • Software • LLM & Generative AI • Information Technology • Data & Analytics • AI Agents • Artificial Intelligence • Cloud Computing • Enterprise Resource Planning (ERP) • Database Management • 1001-5000 employees • Est. 1977
Verified live · 2 hours ago 2 months ago

Principal Machine Learning Engineer

$138k – $295k per year (Estimated) • Lead • In office (Seattle, United States) • Principal
Fine-tuning
SGLang
vLLM
Apply
Report

Reducto

reducto.com
Click here to enter Click here to enter Click here to enter Click here to enter.
reducto.com • San Francisco • New York • Document Management • Data & Analytics • Artificial Intelligence
San Francisco • New York • Document Management • Data & Analytics • Artificial Intelligence
Top 25% payVerified live · 1 day ago 2 months ago

Machine Learning Infrastructure Tech Lead

$200k – $300k per year • Lead • 5+ years expIn office (San Francisco, United States) • Full-Time • Staff
Python
AI/ML
AI Agents
CUDA Toolkit
LLM
PyTorch
Ray
SGLang
TensorRT
TensorRT-LLM
vLLM
DevOps
Kubernetes
Apply
Report

SimpliSafe

simplisafe.com
SimpliSafe is a home security company headquartered in Boston, Massachusetts, and founded in 2006. The company designs and manufactures wireless security systems, including motion sensors, entry sensors, indoor and outdoor cameras, and smoke detectors, all of which are integrated with professional monitoring services. It serves over six million customers across the United States and the United Kingdom, offering DIY installation and no-contract subscription plans for residential and small business security.
simplisafe.com • HQ: Boston, United States • IoT Security • Cybersecurity • Security Services • Professional Services • Consumer Electronics • Smart Home Devices • Consumer Goods • Security Products • Est. 2006
HQ: Boston, United States • IoT Security • Cybersecurity • Security Services • Professional Services • Consumer Electronics • Smart Home Devices • Consumer Goods • Security Products • Est. 2006
Verified live · 12 hours ago 4 months ago

Staff Software Engineer, ML Infrastructure

$147k – $215k per year • Staff+ • 8+ years expRemote/Hybrid (Boston, United States) • Staff
C++
Python
Rust
Databases
Apache Kafka
AI/ML
Computer Vision
LLM
Flink
KServe
MLFlow
Quantization
Ray
SGLang
TensorRT
TensorRT-LLM
vLLM
LLM Guardrails
TGI
DevOps
Amazon EKS
AWS
CI/CD
Kubernetes
Amazon Kinesis
Amazon S3
IAM
Apply
Verified live · 12 hours ago 4 months ago

Staff Machine Learning Engineer, ML Infrastructure

$184k – $245k per year • Staff+ • 8+ years expRemote/Hybrid (Boston, United States) • Staff
C++
Python
Rust
Databases
Apache Kafka
AI/ML
Computer Vision
KServe
Kubeflow
LLM
Ray
vLLM
Triton
LLM Guardrails
Flink
MLFlow
Quantization
SGLang
TensorRT
TensorRT-LLM
TGI
DevOps
Amazon EKS
AWS
CI/CD
Kubernetes
Amazon S3
IAM
Amazon Kinesis
Apply
Report

SpaceX

spacex.com
SpaceX designs, manufactures and launches advanced rockets and spacecraft. The company was founded in 2002 to revolutionize space technology, with the ultimate goal of enabling people to live on other planets.
spacex.com • HQ: Brownsville, United States • Space & Aerospace • 1001-5000 employees
HQ: Brownsville, United States • Space & Aerospace • 1001-5000 employees
Top 25% payVerified live · 9 hours ago 3 months ago

Application Software Engineer, Inference

$135k – $160k per year • Equity • Junior • 2+ years expRemote/Hybrid (Palo Alto, United States) • Bachelor's Degree • Full-Time • Junior
C++
Rust
Go
Python
Databases
ClickHouse
PostgreSQL
AI/ML
LLM
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
DevOps
CI/CD
Docker
gRPC
Kubernetes
Apply
Report

SanDisk

sandisk.com
SanDisk is a globally recognized brand specializing in flash memory storage products, including memory cards, USB flash drives, and solid-state drives (SSDs). Originally founded in 1988 as a pioneer in digital storage technology, it became a subsidiary of Western Digital following its acquisition in 2016. The company serves consumers, creative professionals, and enterprise clients with reliable hardware designed to store, manage, and transfer data across a wide range of devices.
sandisk.com • HQ: Milpitas, United States • Hardware • Consumer Electronics • Storage • Digital Storage • 501-1000 employees • Est. 1988
HQ: Milpitas, United States • Hardware • Consumer Electronics • Storage • Digital Storage • 501-1000 employees • Est. 1988
Verified live · 15 min ago 2 months ago

AI SW Stack Deployment Architect (12+ years)

$47k – $104k per year (Estimated) • Staff+ • 12+ years expIn office (Bengaluru, India) • Full-Time • Senior
JAX
LLM
PyTorch
SGLang
TensorFlow
Transformers
vLLM
AI Agents
Frontend
Lighthouse
Apply
Report

RxSense

rxsense.com
RxSense was founded in 2015 with a mission to develop technology solutions that improve healthcare transparency and access to more affordable medications. The result was RxAgile, a cutting-edge cloud-based enterprise platform for pharmacy benefit ...
rxsense.com • Boston • Dublin • Software • Pharmaceuticals • Health Care
Boston • Dublin • Software • Pharmaceuticals • Health Care
Top 25% payVerified live · 1 day ago 2 months ago

Principal AI Ops Engineer

$190k – $225k per year • Lead • 2+ years exp • Remote (United States) • Bachelor's Degree • Contractor • Principal
Python
TypeScript
C#
C#
.NET
AI/ML
Claude
LangGraph
LLM
SGLang
vLLM
LangChain
Anthropic
OpenAI
OpenAI Agents SDK
TGI
DevOps
AIOps
AWS
CI/CD
FinOps
Kubernetes
IAM
Apply
Report

Parallel Wireless

parallelwireless.com
Parallel Wireless is a telecommunications software and hardware company headquartered in Nashua, New Hampshire, and founded in 2012. The company develops cloud-native Open RAN (Radio Access Network) solutions that support 2G, 3G, 4G, and 5G technologies through software-defined radios and network automation tools. It serves mobile network operators and private industries globally, providing interoperable infrastructure designed to reduce deployment costs and increase operational efficiency.
parallelwireless.com • HQ: Nashua, United States • Network Equipment • Mobile Networks • Telecom Infrastructure • Telecommunications • Hardware • Wireless • Est. 2012
HQ: Nashua, United States • Network Equipment • Mobile Networks • Telecom Infrastructure • Telecommunications • Hardware • Wireless • Est. 2012
Verified live · 1 day ago 2 months ago

Senior/Principal Local LLM & Generative AI Platform Engineer

$118k – $282k per year (Estimated) • Lead • 7+ years expRemote/Hybrid (Kfar Saba, Israel) • Bachelor's Degree • Full-Time • Principal
C++
Go
Java
Python
AI/ML
Embeddings
Fine-tuning
Knowledge Distillation
LLM
Multimodal AI
Quantization
RAG
Reranking
Tokenization
PEFT
Structured Outputs
Function Calling
CUDA
CUDA Toolkit
KServe
llama.cpp
LoRA
QLoRA
Ray
Ray Serve
SGLang
TensorRT
TensorRT-LLM
Triton
vLLM
Post-training
ROCm
Synthetic Data
DevOps
CI/CD
Docker
Git
Kubernetes
Vector
Cybersecurity
Least Privilege
Apply
Report

Modal

modal.com
Modal is an AI infrastructure company headquartered in New York City and founded in 2021. It provides a serverless cloud platform featuring sub-second cold starts and instant autoscaling that enables developers to run GPU-accelerated workloads, including model inference and fine-tuning, using a Python-native SDK. The company operates a globally distributed compute network designed for AI applications and serves a diverse range of industries such as generative AI and biotechnology.
modal.com • HQ: New York, United States • MLOps • LLM & Generative AI • Machine Learning • Information Technology • Hardware • Cloud Computing • AI Infrastructure • Artificial Intelligence • Est. 2021
HQ: New York, United States • MLOps • LLM & Generative AI • Machine Learning • Information Technology • Hardware • Cloud Computing • AI Infrastructure • Artificial Intelligence • Est. 2021
Top 25% payVerified live · 1 day ago 2 months ago

Member of Technical Staff - Research, Inference

$150k – $350k per year • Staff+ • In office (New York, San Francisco, United States) • Full-Time • Staff
Flash Attention
LLM
Multimodal AI
Quantization
SGLang
Analytics
Seaborn
Matplotlib
Apply
Open 180 daysVerified live · 1 day ago 6 months ago

Forward Deployed Engineer - ML

$52k – $154k per year (Estimated) • Junior • 2+ years expIn office (Stockholm, Sweden) • Full-Time • Junior
LLM
RLHF
SGLang
TRL
vLLM
Transformers
SFT
Analytics
Seaborn
Matplotlib
Apply
Open 188 daysTop 25% pay 7 months ago

Forward Deployed Engineer - ML

$180k – $250k per year • Junior • 2+ years expIn office (New York, San Francisco, United States) • Full-Time • Junior
LLM
RLHF
SGLang
TRL
vLLM
Transformers
SFT
Analytics
Seaborn
Matplotlib
Apply
Report

Amgen

amgen.com
Amgen is a global biopharmaceutical pioneer headquartered in Thousand Oaks, California, that specializes in discovering, developing, and manufacturing innovative biologic therapies. The company focuses on treating serious illnesses with high unmet medical needs across key areas including oncology, cardiovascular disease, inflammation, rare diseases, and nephrology. Leveraging advanced human genetics, molecular engineering, and biosimilar development, it serves millions of patients worldwide through established blockbuster treatments and cutting-edge pipelines.
amgen.com • HQ: Thousand Oaks, United States • Biotechnology • Health Care • Pharmaceuticals • 1001-5000 employees • Est. 1980
HQ: Thousand Oaks, United States • Biotechnology • Health Care • Pharmaceuticals • 1001-5000 employees • Est. 1980
Verified live · 1 day ago 3 months ago

Principal Machine Learning Engineer - Forecasting

$32k – $76k per year (Estimated) • Lead • 12+ years expIn office (Hyderabad, India) • Full-Time • Principal
Python
SQL
Python
FastAPI
AI/ML
AI Agents
Dagster
JAX
LLM
MLFlow
NLP
Prefect
PyTorch
Ray
Scikit-learn
Spark
TensorFlow
Human-in-the-Loop
LLM Guardrails
Function Calling
SGLang
TensorRT
TensorRT-LLM
vLLM
Model Context Protocol
RAG
DevOps
CI/CD
Analytics
A/B Testing
Apply
Report
SGLang Jobs - Remote & On-site
Frequently asked questions
SGLang: How many jobs are available now?
There are 219 SGLang jobs on Alion right now.
SGLang: What is the average salary for these roles?
The average salary for SGLang roles on Alion is $295,494 per year.
SGLang: Are remote, relocation, or visa options available?
Many SGLang listings are remote-friendly or hybrid, and several employers offer relocation or visa sponsorship.
SGLang: How do I apply on Alion?
Create an Alion account, view the SGLang job posting, and click Apply to submit your resume and profile.