vLLM Jobs - Remote & On-site

vLLM jobs curated from vetted startups and product companies, updated daily. Many vLLM roles offer remote or hybrid schedules and include transparent compensation to help you decide. Browse and apply today.

AI/ML
Data Science
Backend
Frontend
Mobile
DevOps
Web3
Games
Hardware
Robotics
Security
QA
Executive
Networking
Product
Design
Analytics
Support
Enterprise Apps
Quantum

Ciq

ciq.com
Ciq is a company that provides a platform for managing cybersecurity risks. They offer a variety of services, including threat intelligence, vulnerability management, and compliance. The company's mission is to help businesses reduce their cybersecurity risks.
ciq.com • Artificial Intelligence • Information Technology • Operating Systems • Est. 2020
Artificial Intelligence • Information Technology • Operating Systems • Est. 2020
Open 119 daysVerified live · 2 hours ago 4 months ago

Sr./Principal Performance Engineer

$112k – $243k per year (Estimated) • Equity • Lead • 15+ years exp • Remote • PhD • Principal
C
C
MPI
AI/ML
LLM
ONNX
OpenMP
Quantization
TensorRT
vLLM
AI Agents
InfiniBand
DevOps
CI/CD
eBPF
SLURM
HPC
Apply
Report

Handshake

joinhandshake.com
Handshake is a professional career network for students and early-career professionals headquartered in San Francisco, California, and founded in 2014. The company provides a digital platform that connects university students with internships and entry-level jobs, including specialized roles in the artificial intelligence economy. It serves over 15 million students and alumni from more than 1,500 colleges and universities, partnering with over 900,000 employers across the United States, United Kingdom, and Europe.
joinhandshake.com • HQ: San Francisco, United States • Machine Learning • Artificial Intelligence • Online Directories & Classifieds • Internet Services • Higher Education • Education • Professional Services • Human Resources • Est. 2014
HQ: San Francisco, United States • Machine Learning • Artificial Intelligence • Online Directories & Classifieds • Internet Services • Higher Education • Education • Professional Services • Human Resources • Est. 2014
Top 25% payVerified live · 1 hour ago 2 months ago

Senior Software Engineer, Machine Learning Infrastructure

$176k – $220k per year • Senior • 5+ years expIn office (San Francisco, United States) • Full-Time • Senior
Go
Python
TypeScript
Databases
Google BigQuery
Google Bigtable
Redis
AI/ML
Embeddings
Fine-tuning
LLM
Reinforcement Learning
Spark
Feature Store
LLM Evaluation
Post-training
Scale AI
PyTorch
Ray
Ray Serve
RLHF
Vertex AI
vLLM
AI Agents
Function Calling
Model Context Protocol
DevOps
AWS
CI/CD
Docker
GCP
Kubernetes
Terraform
Apply
Report

Confido

confidotech.com
Confido is the AI infrastructure for CPG brands, providing a single source of truth for accounting, finance, sales, and operations teams.
confidotech.com • New York • Consumer Goods • Data & Analytics • Artificial Intelligence • 51-200 employees
New York • Consumer Goods • Data & Analytics • Artificial Intelligence • 51-200 employees
2 months ago

Senior ML Ops Engineer

$185k – $358k per year (Estimated) • Senior • In office (New York, United States) • Relocation • Full-Time • Senior
Java
Python
Ruby
Databases
Amazon Aurora
Apache Kafka
Redis
Snowflake
AI/ML
LLM
VLM
Human-in-the-Loop
AI Agents
AWS Bedrock
BentoML
MLFlow
Multimodal AI
ONNX
Ray
TensorRT
Vertex AI
vLLM
Amazon SageMaker
LLM Evaluation
LLMOps
DevOps
CI/CD
Platform Engineering
AWS
GitHub Actions
Kubernetes
Terraform
Vector
GitHub
Apply
Report

Gina's Tech Jobs

ginastechjobs.com
Gina's Tech Jobs is a technology recruitment agency headquartered in Charlotte, North Carolina, and founded in 1999. The agency places engineers into permanent roles at United States technology employers, with a heavy concentration in machine learning, artificial intelligence, and software engineering positions. It works largely on remote and hybrid searches and represents both established employers and smaller AI companies hiring technical staff.
ginastechjobs.com • HQ: Charlotte, United States • Machine Learning • Artificial Intelligence • IT Consulting • Information Technology • Professional Services • Human Resources • Est. 1999
HQ: Charlotte, United States • Machine Learning • Artificial Intelligence • IT Consulting • Information Technology • Professional Services • Human Resources • Est. 1999
$170k – $200k per year • Lead • San Francisco • Remote (United States) • Full-Time • Senior
DeepSpeed
Diffusion Models
Fine-tuning
JAX
LLM
Multimodal AI
PyTorch
Quantization
Ray
RLHF
Spark
TensorRT
TensorRT-LLM
vLLM
DPO
FSDP
Megatron-LM
PPO
Apply
Report

Radiant

radiant.com
AbcdefghijklmnopqrstuvwxyzABCDEFGHIJKLMNOPQRSTUVWXYZ ABOUT PHOTO APP What you write here is totally up to you, there is no right or wrong way to complete your 'About Us' page. We advise making your language friendly and approachable, while being informative and professional.
radiant.com • HQ: London, United Kingdom
HQ: London, United Kingdom
Verified live · 9 hours ago 2 months ago

Senior SDET

$59k – $146k per year (Estimated) • Senior • Remote/Hybrid (London, United Kingdom) • Full-Time • Senior
Go
Databases
PostgreSQL
AI/ML
KServe
LLM
vLLM
DevOps
CI/CD
gRPC
Helm
Kubernetes
Apply
Open 122 daysVerified live · 9 hours ago 4 months ago

Principal Software Engineer: Cloud Platform

$97k – $194k per year (Estimated) • Lead • Remote/Hybrid (London, United Kingdom) • Full-Time • Principal
Go
Python
Databases
PostgreSQL
AI/ML
KServe
Ray
vLLM
Triton
DevOps
CI/CD
gRPC
Kubernetes
Apply
Open 158 daysVerified live · 9 hours ago 6 months ago

Senior Backend Software Engineer - Core Services

$68k – $129k per year • Senior • Remote/Hybrid (London, United Kingdom) • Full-Time • Senior
Go
Databases
PostgreSQL
AI/ML
KServe
vLLM
DevOps
gRPC
Helm
Kubernetes
CI/CD
Apply
Report

Tekion

tekion.com
Tekion is an end-to-end, AI-native automotive retail platform that unifies dealership operations - DMS, CRM, digital retail, service, payments and analytics - on a single cloud-based operating system. Tekion powers 3,000+ dealerships and has processed $43B+ in transactions.
tekion.com • HQ: Pleasanton, United States • Software • Artificial Intelligence • Commerce • Customer Relationship Management (CRM) • 1001-5000 employees • Est. 2016
HQ: Pleasanton, United States • Software • Artificial Intelligence • Commerce • Customer Relationship Management (CRM) • 1001-5000 employees • Est. 2016
Verified live · 4 hours ago 2 months ago

Staff Software Development Test Engineer

$33k – $72k per year (Estimated) • Staff+ • 8+ years expIn office (Bengaluru, India) • Full-Time • Staff
Go
Java
Python
Databases
Milvus
Pinecone
AI/ML
AI Agents
LLM
Synthetic Data
Feature Store
Feast
KServe
Kubeflow
MLFlow
Ray
Vertex AI
vLLM
Amazon SageMaker
Seldon Core
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
SLI/SLO/SLA
Vector
Datadog
Grafana
Prometheus
Apply
Report

Deepgram

deepgram.com
Deepgram is an AI speech platform that provides real-time and asynchronous automated speech recognition (ASR) and text-to-speech (TTS) APIs. Built on custom deep learning models, its technology delivers fast, accurate, and cost-effective voice transcription, language understanding, and audio synthesis for developers and enterprise businesses. Headquartered in San Francisco, California, the company enables organizations to integrate advanced voice capabilities and intelligence directly into their applications and conversational AI workflows.
deepgram.com • San Francisco • London • Washington • New York • Sydney • Conversational AI • Artificial Intelligence • Speech & Audio AI • Est. 2015
San Francisco • London • Washington • New York • Sydney • Conversational AI • Artificial Intelligence • Speech & Audio AI • Est. 2015
Top 25% payVerified live · 1 day ago 2 months ago

Senior Technical Program Manager (Engineering) - AI Tooling & Systems

$152k – $208k per year • Senior • 5+ years exp • Remote (United States) • Full-Time • Senior
CUDA
CUDA Toolkit
Knowledge Distillation
LLM
MLFlow
Multimodal AI
Prompt Engineering
PyTorch
Quantization
TensorRT
Vertex AI
vLLM
Weights & Biases
Amazon SageMaker
Deepgram
Feature Store
Hugging Face
Text-to-Speech
TorchServe
Vapi
Mobile
Twilio
DevOps
AWS
Azure
Cloudflare
GCP
Analytics
A/B Testing
Apply
Report

Commonai

commonai.org
Home About Programmes News Careers About Us CommonAI provides a new backbone for AI infrastructure. We believe that collaborative engineering is the means to accelerate AI innovation.
commonai.org • Cambridge • Clinical Research
Cambridge • Clinical Research
Verified live · 1 day ago 2 months ago

Project Technical Lead - AI Systems Simulation

$108k – $228k per year (Estimated) • Lead • In office (Cambridge, United States) • PhD • Full-Time • Staff
vLLM
Apply
Verified live · 1 day ago 3 months ago

Senior Software Engineer (vLLM)

$99k – $179k per year (Estimated) • Senior • In office (Cambridge, United States) • Bachelor's Degree • Full-Time • Senior
C++
Python
Rust
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
LLM
PyTorch
Ray
TensorRT
TensorRT-LLM
vLLM
Hugging Face
TGI
TPU
DevOps
CI/CD
Apply
Open 196 daysVerified live · 1 day ago 7 months ago

Senior Performance Engineer

$98k – $177k per year (Estimated) • Senior • In office (Cambridge, United States) • Bachelor's Degree • Full-Time • Senior
Python
AI/ML
CUDA
CUDA Toolkit
NumPy
Pandas
PyTorch
vLLM
DevOps
Grafana
Prometheus
Apply
Report

Parallel Wireless

parallelwireless.com
Parallel Wireless is a telecommunications software and hardware company headquartered in Nashua, New Hampshire, and founded in 2012. The company develops cloud-native Open RAN (Radio Access Network) solutions that support 2G, 3G, 4G, and 5G technologies through software-defined radios and network automation tools. It serves mobile network operators and private industries globally, providing interoperable infrastructure designed to reduce deployment costs and increase operational efficiency.
parallelwireless.com • HQ: Nashua, United States • Network Equipment • Mobile Networks • Telecom Infrastructure • Telecommunications • Hardware • Wireless • Est. 2012
HQ: Nashua, United States • Network Equipment • Mobile Networks • Telecom Infrastructure • Telecommunications • Hardware • Wireless • Est. 2012
Verified live · 1 day ago 2 months ago

Senior/Principal Local LLM & Generative AI Platform Engineer

$118k – $282k per year (Estimated) • Lead • 7+ years expRemote/Hybrid (Kfar Saba, Israel) • Bachelor's Degree • Full-Time • Principal
C++
Go
Java
Python
AI/ML
Embeddings
Fine-tuning
Knowledge Distillation
LLM
Multimodal AI
Quantization
RAG
Reranking
Tokenization
PEFT
Structured Outputs
Function Calling
CUDA
CUDA Toolkit
KServe
llama.cpp
LoRA
QLoRA
Ray
Ray Serve
SGLang
TensorRT
TensorRT-LLM
Triton
vLLM
Post-training
ROCm
Synthetic Data
DevOps
CI/CD
Docker
Git
Kubernetes
Vector
Cybersecurity
Least Privilege
Apply
Report

Radical Numerics

radicalnumerics.ai
Radical Numerics is an artificial intelligence research lab dedicated to developing general biological intelligence and generative genomics models. Headquartered in San Francisco, California, the company builds multimodal platforms capable of reading, writing, and engineering biological sequences across DNA, RNA, and proteins. Its technology aims to accelerate biopharmaceutical research, enhance early disease diagnostics, and establish robust biodefense capabilities.
radicalnumerics.ai • HQ: San Francisco, United States
HQ: San Francisco, United States
Verified live · 1 day ago 2 months ago

Member of Technical Staff, ML Engineer

$182k – $369k per year (Estimated) • Staff+ • In office (San Francisco, United States) • Full-Time • Staff
Python
AI/ML
LLM
TensorRT
TensorRT-LLM
Triton
vLLM
Apply
Verified live · 1 day ago 3 months ago

Member of Technical Staff, Inference

$193k – $391k per year (Estimated) • Staff+ • In office (San Francisco, United States) • Full-Time • Staff
Python
AI/ML
CUDA
CUDA Toolkit
Multimodal AI
PyTorch
Triton
Mixture of Experts
DeepSpeed
LLM
SGLang
TensorRT
TensorRT-LLM
vLLM
Apply
Report

Code Metal

codemetal.com
Code Metal is an artificial intelligence company headquartered in Boston, Massachusetts, and founded in 2024. The company builds tooling that translates and optimises software for edge and embedded hardware, automating the porting work needed when code must run on a different processor architecture. It targets defence, aerospace, and industrial customers maintaining large legacy codebases tied to specific chips.
codemetal.com • HQ: Boston, United States • Edge AI • Embedded Software • Software • Hardware • Embedded Systems • Artificial Intelligence • Code Intelligence • 201-500 employees • Est. 2024
HQ: Boston, United States • Edge AI • Embedded Software • Software • Hardware • Embedded Systems • Artificial Intelligence • Code Intelligence • 201-500 employees • Est. 2024
Verified live · 1 day ago 2 months ago

Applied AI Research Engineer

$132k – $290k per year (Estimated)Remote/Hybrid (Boston, United States) • Relocation • Bachelor's Degree • Full-Time
Python
AI/ML
Embeddings
Fine-tuning
Knowledge Distillation
LLM
PyTorch
Reinforcement Learning
Synthetic Data
Transformers
vLLM
Hugging Face
Human-in-the-Loop
LLM Guardrails
Post-training
Structured Outputs
AI Agents
Function Calling
RAG
Apply
Report

Pure Storage

purestorage.com
Pure Storage is an enterprise technology company that specializes in all-flash data storage systems, data management software, and integrated cloud services. The firm focuses on replacing legacy disk-based infrastructure with modern, high-performance flash technology to help organizations simplify their operations and enhance data accessibility. Through a software-led platform and subscription-based service model, the company provides resilient, scalable solutions designed to support diverse workloads ranging from artificial intelligence to cloud-native applications.
purestorage.com • HQ: Santa Clara, United States • Information Technology • Data Centers • Hardware • Storage • 1001-5000 employees • Est. 2009
HQ: Santa Clara, United States • Information Technology • Data Centers • Hardware • Storage • 1001-5000 employees • Est. 2009
Verified live · 2 hours ago 3 months ago

Senior LLM AI Engineer

Senior • In office (Prague, Czech Republic) • Senior
Python
SQL
AI/ML
Claude
Fine-tuning
Gemini
LLM
ONNX
PyTorch
RLHF
Semantic Search
TensorRT
TensorRT-LLM
Transformers
vLLM
DPO
Hugging Face
Semantic Search
AI Agents
Model Context Protocol
Apply
Report

Mirantis

mirantis.com
Mirantis is a B2B open-source cloud computing, container management, and AI infrastructure company headquartered in Campbell, California. Originally known as a core contributor to OpenStack, Mirantis now provides open-cloud software and managed services centered on Kubernetes, multi-cloud platforms, and enterprise AI workloads.
mirantis.com • Calgary • Warsaw • Brussels • Prague • Poznań • DevOps • Information Technology • Cloud Computing • 1001-5000 employees
Calgary • Warsaw • Brussels • Prague • Poznań • DevOps • Information Technology • Cloud Computing • 1001-5000 employees
Verified live · 2 hours ago 3 months ago

Senior Kubernetes DevOps Engineer - K0rdent AI Apps/Core Services

$135k – $262k per year (Estimated) • Equity • Senior • Remote/Hybrid (San Jose, United States) • Master's Degree • Full-Time • Senior
KServe
Kubeflow
vLLM
DevOps
ArgoCD
Grafana
Harbor
Kubernetes
KubeVirt
Platform Engineering
AWS
OpenStack
Apply
Open 96 daysVerified live · 2 hours ago 4 months ago

Product Manager - AI Inference & Model Serving

$145k – $248k per year (Estimated) • Equity • Senior • 7+ years expAustin • Remote (United States) • Full-Time • Senior
LLM
SGLang
TensorRT
TensorRT-LLM
vLLM
Triton
DevOps
AWS
Kubernetes
Platform Engineering
Chips/EDA
PoC Library
Apply
Report

Anyscale

anyscale.com
Anyscale is an artificial intelligence infrastructure company headquartered in San Francisco, California, and founded in 2019 by the creators of Ray at the UC Berkeley RISELab. The company offers a managed platform for running Ray, the open source framework used to scale model training, batch inference, and reinforcement learning across clusters. It sells to machine learning engineering teams that need to move distributed AI workloads from laptops to production without rebuilding their stack.
anyscale.com • HQ: San Francisco, United States • Hardware • Information Technology • MLOps • Artificial Intelligence • AI Infrastructure • Cloud Computing • Machine Learning • 201-500 employees • Est. 2019
HQ: San Francisco, United States • Hardware • Information Technology • MLOps • Artificial Intelligence • AI Infrastructure • Cloud Computing • Machine Learning • 201-500 employees • Est. 2019
Verified live · 10 hours ago 3 months ago

Machine Learning Engineer, Customer Engineering

$170k – $199k per year • Senior • 7+ years expRemote/Hybrid (San Francisco, United States) • Full-Time • Senior
Fine-tuning
LLM
Ray
vLLM
OpenAI
DevOps
Amazon EKS
AWS
Azure
Azure AKS
CI/CD
GCP
GitHub Actions
Google GKE
Kubernetes
Terraform
GitHub
Apply
Top 25% payVerified live · 10 hours ago 4 months ago

Distributed LLM Inference Engineer

$170k – $245k per year • Equity • In office (San Francisco, Palo Alto, United States) • Full-Time
LLM
PyTorch
Ray
vLLM
OpenAI
CUDA Toolkit
TensorFlow
TensorRT
TensorRT-LLM
CUDA
Triton
Apply
Report

Banco Santander

santander.com
Banco Santander is a prominent multinational financial services company that provides retail, commercial, and investment banking products to millions of customers globally. The institution is recognized as one of the largest banking groups in the world, maintaining a dominant market presence throughout Europe and the Americas. Headquartered in Madrid, Spain, the global bank operates an extensive network of branches and subsidiaries to deliver tailored financial solutions to its diverse international client base.
santander.com • HQ: Madrid, Spain • Events & Ticketing • Corporate Events • Artificial Intelligence • Credit Cards • Financial Services • Banking • 1001-5000 employees
HQ: Madrid, Spain • Events & Ticketing • Corporate Events • Artificial Intelligence • Credit Cards • Financial Services • Banking • 1001-5000 employees
Verified live · 4 hours ago 3 months ago

Frontier Agentic AI Engineer — Santander AI Lab

$41k – $115k per year (Estimated) • Middle • 4+ years expRemote/Hybrid (Spain) • Bachelor's Degree • Full-Time • Middle • Spanish: B2 • English: B2
Python
Python
FastAPI
AI/ML
AI Agents
AutoGen
AWS Bedrock
Claude
CrewAI
Fine-tuning
Langfuse
LangGraph
Llama
LLM
LoRA
Mistral
Model Context Protocol
Ollama
Promptfoo
RLHF
Semantic Kernel
Unsloth
vLLM
Windsurf
LangChain
PEFT
A2A
Amazon SageMaker
Anthropic
Devin
GraphRAG
Knowledge Graph
OpenAI
SFT
Function Calling
DevOps
AWS
AWS Lambda
CI/CD
Docker
Git
Kubernetes
Rest API
Terraform
Amazon S3
Apply
Report

Amgen

amgen.com
Amgen is a global biopharmaceutical pioneer headquartered in Thousand Oaks, California, that specializes in discovering, developing, and manufacturing innovative biologic therapies. The company focuses on treating serious illnesses with high unmet medical needs across key areas including oncology, cardiovascular disease, inflammation, rare diseases, and nephrology. Leveraging advanced human genetics, molecular engineering, and biosimilar development, it serves millions of patients worldwide through established blockbuster treatments and cutting-edge pipelines.
amgen.com • HQ: Thousand Oaks, United States • Biotechnology • Health Care • Pharmaceuticals • 1001-5000 employees • Est. 1980
HQ: Thousand Oaks, United States • Biotechnology • Health Care • Pharmaceuticals • 1001-5000 employees • Est. 1980
Verified live · 1 day ago 3 months ago

Principal Machine Learning Engineer - Forecasting

$32k – $76k per year (Estimated) • Lead • 12+ years expIn office (Hyderabad, India) • Full-Time • Principal
Python
SQL
Python
FastAPI
AI/ML
AI Agents
Dagster
JAX
LLM
MLFlow
NLP
Prefect
PyTorch
Ray
Scikit-learn
Spark
TensorFlow
Human-in-the-Loop
LLM Guardrails
Function Calling
SGLang
TensorRT
TensorRT-LLM
vLLM
Model Context Protocol
RAG
DevOps
CI/CD
Analytics
A/B Testing
Apply
Report

Accellor

accellor.com
Accellor is an AI-first digital transformation and technology services firm headquartered in Frisco, Texas, and founded in 2018. The company develops purpose-built artificial intelligence solutions, including agentic AI for supply chain optimization, AI-assisted software engineering, and intelligent automation for sectors like healthcare and financial services. It operates globally with a focus on scaling AI from proof of concept to enterprise-wide production through strategic partnerships with major technology providers like Salesforce and Microsoft.
accellor.com • HQ: Frisco, United States • Health Care • Business Process Automation (BPA) • Software • AI Agents • LLM & Generative AI • Artificial Intelligence • Information Technology • IT Consulting • Est. 2018
HQ: Frisco, United States • Health Care • Business Process Automation (BPA) • Software • AI Agents • LLM & Generative AI • Artificial Intelligence • Information Technology • IT Consulting • Est. 2018
Verified live · 1 day ago 3 months ago

AI Principal Engineer

$162k – $346k per year (Estimated) • Lead • 10+ years expIn office (San Francisco, United States) • Principal
C++
Go
Java
Python
Rust
TypeScript
C++
PyTorch C++
TensorFlow C++
AI/ML
AI Agents
ChatGPT
CUDA Toolkit
Function Calling
JAX
LLM
Multimodal AI
PyTorch
Ray
Reinforcement Learning
TensorFlow
vLLM
CUDA
Triton
Context Engineering
LLM Evaluation
NCCL
OpenAI
OpenAI Codex
Post-training
Pre-training
Hallucination
RAG
DevOps
Kubernetes
Platform Engineering
CI/CD
Docker
Grafana
OpenTelemetry
Prometheus
Terraform
Apply
Report

Multiverse Computing

multiversecomputing.com
Multiverse Computing is a technology company that develops software solutions based on quantum and quantum inspired computing. It provides tools and platforms that apply these methods to areas such as optimization, machine learning and financial analysis.
multiversecomputing.com • HQ: Donostia, Spain • Science & Engineering • Artificial Intelligence • Quantum Science & Computing • 51-200 employees • Est. 2019
HQ: Donostia, Spain • Science & Engineering • Artificial Intelligence • Quantum Science & Computing • 51-200 employees • Est. 2019
$107k – $201k per year (Estimated) • Staff+ • In office (London, United Kingdom) • Relocation • Bachelor's Degree • Full-Time • Architect • English
Python
SQL
AI/ML
Kubeflow
llama.cpp
LLM
MLFlow
ONNX
OpenVINO
PyTorch
Safetensors
TensorFlow
Vertex AI
vLLM
Amazon SageMaker
Edge AI
Flyte
Hugging Face
Computer Vision
Quantization
RAG
AI Agents
EU AI Act
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Cybersecurity
GDPR
Apply
$76k – $183k per year (Estimated) • Staff+ • In office (Munich, Germany) • Relocation • Bachelor's Degree • Full-Time • Architect • English • German: C1
Python
SQL
AI/ML
Kubeflow
llama.cpp
LLM
MLFlow
ONNX
OpenVINO
PyTorch
Safetensors
TensorFlow
Vertex AI
vLLM
Amazon SageMaker
Edge AI
Flyte
Hugging Face
Computer Vision
Quantization
RAG
AI Agents
EU AI Act
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Cybersecurity
GDPR
Apply
$60k – $144k per year (Estimated) • Lead • 5+ years expIn office (Donostia, Spain) • Relocation • Visa sponsorship • Master's Degree • Full-Time • Senior • Spanish
Python
AI/ML
Computer Vision
Fine-tuning
LLM
RAG
AI Agents
TensorRT
vLLM
DevOps
Docker
Git
GitHub
GitLab
Apply
Report

Phrase

phrase.com
Phrase is the AI-powered language intelligence platform for enterprise. Automate workflows, connect any tool or engine, and scale multilingual content globally.
phrase.com • Artificial Intelligence • Professional Services • Translation Services • Est. 2012
Artificial Intelligence • Professional Services • Translation Services • Est. 2012
Verified live · 1 day ago 3 months ago

Principal AI Researcher

$98k – $234k per year (Estimated) • Lead • In office • PhD • Full-Time • Principal
Python
AI/ML
Claude
Fine-tuning
Gemini
LangGraph
LLM
LoRA
NLP
Prefect
Pydantic AI
PyTorch
Reinforcement Learning
RLHF
Vertex AI
vLLM
Weights & Biases
LangChain
PEFT
Amazon SageMaker
DPO
Flyte
Hugging Face
Human-in-the-Loop
OpenAI
Pre-training
Structured Outputs
AI Agents
Function Calling
TGI
DevOps
AWS
GCP
Apply
Report

Crosby

crosby.ai
Crosby operates as an artificial intelligence native law firm, using its own software to turn around contracts in hours. Founded in 2024 in San Francisco, it employs lawyers who work alongside the models rather than selling software alone. Startups use it as an alternative to traditional outside counsel.
crosby.ai • HQ: San Francisco, United States • Professional Services • Artificial Intelligence • Legal Services • Est. 2024
HQ: San Francisco, United States • Professional Services • Artificial Intelligence • Legal Services • Est. 2024
Top 25% payVerified live · 2 days ago 3 months ago

Member of Technical Staff, Research Engineer

$200k – $300k per year • Staff+ • 3+ years expIn office (New York, United States) • Full-Time • Staff
Python
AI/ML
Cursor
Fine-tuning
LLM
RAG
vLLM
Anthropic
GRPO
Human-in-the-Loop
OpenAI
Post-training
PPO
SFT
AI Agents
Apply
Report
vLLM Jobs - Remote & On-site
Frequently asked questions
vLLM: How many jobs are available now?
There are 562 vLLM job openings listed on Alion right now; listings are updated daily.
vLLM: What is the typical salary for vLLM roles?
The typical vLLM role on Alion averages $241,236 USD annually.
vLLM: Are remote, relocation, or visa sponsorship options available?
vLLM listings include remote, hybrid, and on-site roles, and some companies may provide relocation or visa sponsorship-check individual postings.
vLLM: How do I apply for vLLM jobs on Alion?
Sign in to Alion, upload your resume, and apply directly from the vLLM job pages; applications are sent to the hiring company.