CUDA Toolkit Jobs - Remote & On-site

CUDA Toolkit - curated CUDA Toolkit jobs and roles at vetted startups and product companies, updated daily. Find remote and hybrid positions on Alion and narrow results by domain, seniority, and GPU expertise. Browse and apply today.

AI/ML
Data Science
Backend
Frontend
Mobile
DevOps
Web3
Games
Hardware
Robotics
Security
QA
Executive
Networking
Product
Design
Analytics
Support
Enterprise Apps
Quantum

Wintermute

wintermute.com
Wintermute is one of the largest algorithmic trading companies in digital assets. We provide liquidity algorithmically across all major cryptocurrency exchanges and trading platforms, a broad range of OTC trading solutions as well as support high profile blockchain projects and traditional financial institutions moving into crypto. Wintermute is not just a trading company, it is one the most prominent and influential players in the digital asset markets: we are connected and partnering with all major players in the industry, we actively participate in the development of the blockchain ecosystem through investments, partnerships, and incubation of projects.
wintermute.com • HQ: London, United Kingdom • DeFi • Cryptocurrencies • Blockchain & Crypto
HQ: London, United Kingdom • DeFi • Cryptocurrencies • Blockchain & Crypto
Open 285 daysVerified live · 1 day ago 10 months ago

Machine Learning Researcher

$100k – $210k per year (Estimated)In office (London, United Kingdom) • Relocation • Bachelor's Degree • Full-Time
Python
C++
AI/ML
Time Series Forecasting
CUDA
CUDA Toolkit
Web3
DeFi
Apply
Report

Baseten

baseten.co
Baseten is an AI infrastructure platform designed to help developers and machine learning teams deploy, serve, and scale open-source and custom AI models. The platform provides performant, low-latency inference infrastructure alongside developer tools like Truss, an open-source model packaging framework. Headquartered in San Francisco, California, Baseten enables companies to run state-of-the-art models in production seamlessly without managing underlying cloud infrastructure.
baseten.co • HQ: San Francisco, United States • Hardware • Machine Learning • AI Infrastructure • Artificial Intelligence
HQ: San Francisco, United States • Hardware • Machine Learning • AI Infrastructure • Artificial Intelligence
Open 324 daysTop 25% pay 11 months ago

Software Engineer - Model Products

$180k – $360k per year • Middle • 3+ years expRemote/Hybrid (San Francisco, New York, United States, Toronto, Montreal, Canada) • Full-Time • Middle
CUDA Toolkit
Cursor
Function Calling
LLM
Multimodal AI
Quantization
TensorRT
TensorRT-LLM
CUDA
Structured Outputs
SGLang
vLLM
TGI
DevOps
SLI/SLO/SLA
Kubernetes
Management
Notion
Apply
Open 409 daysTop 25% pay 1 year ago

Software Engineer - GPU Kernels

$180k – $360k per year • Remote/Hybrid (San Francisco, New York, United States, Toronto, Montreal, Canada) • Full-Time
C++
AI/ML
CUDA Toolkit
Cursor
Embeddings
Quantization
CUDA
Mixture of Experts
Flash Attention
Transformers
Triton
DevOps
AWS
Management
Notion
Apply
Open 885 daysTop 25% pay 2 years ago

Software Engineer - Model Performance

$180k – $360k per year • Remote/Hybrid (San Francisco, New York, United States, Toronto, Montreal, Canada) • Bachelor's Degree • Full-Time
C++
Python
C++
PyTorch C++
AI/ML
CUDA Toolkit
Cursor
Embeddings
LLM
LoRA
PyTorch
Quantization
SGLang
Spark
TensorRT
TensorRT-LLM
PEFT
CUDA
DevOps
Docker
Kubernetes
Management
Notion
Apply
Report

Specter

specter.com
Specter tracks growth signals across millions of private companies to help investors find opportunities early. Founded in 2021, it combines web, hiring, app and social data into company momentum scores. Venture funds and corporate development teams use it for sourcing.
specter.com • HQ: Singapore • Est. 2021
HQ: Singapore • Est. 2021
Open 331 dayVerified live · 9 hours ago 11 months ago

ML Research Engineer

$165k – $319k per year (Estimated) • Senior • 5+ years expIn office (San Francisco, United States) • Full-Time • Senior
C++
Rust
C++
PyTorch C++
AI/ML
AI Agents
Computer Vision
CUDA
CUDA Toolkit
Fine-tuning
LLM
Multimodal AI
ONNX
PyTorch
Quantization
RAG
TensorRT
TensorRT-LLM
Apply
Report

Abridge

abridge.com
Abridge is an AI-powered clinical documentation platform that transforms patient-clinician conversations into structured, real-time medical notes. Designed to reduce administrative burden and clinician burnout, the platform integrates seamlessly into major electronic health record (EHR) systems like Epic. By leveraging specialized healthcare language models, Abridge enables medical providers to spend less time on paperwork and focus more on delivering quality patient care.
abridge.com • HQ: San Francisco, United States • Health Care • Medical AI • Artificial Intelligence • 201-500 employees
HQ: San Francisco, United States • Health Care • Medical AI • Artificial Intelligence • 201-500 employees
Open 370 daysTop 25% pay 1 year ago

Machine Learning Infrastructure Engineer, Model Inference

$221k – $260k per year • Equity • Senior • 5+ years expRemote/Hybrid • Full-Time • Senior
CUDA Toolkit
LLM
PyTorch
TensorFlow
CUDA
Triton
DevOps
Ansible
GitOps
Kubernetes
Terraform
Apply
Report
Genesis Molecular AI is a drug discovery company headquartered in Burlingame, California, and founded in 2019. The company combines deep learning models of molecular structure and binding with in-house chemistry and biology to design small molecule drug candidates against difficult targets. It advances its own pipeline while running partnered discovery programmes with large pharmaceutical companies, and previously operated under the name Genesis Therapeutics.
genesis.ml • HQ: Burlingame, United States • Pharmaceuticals • Bioinformatics • Biopharma • Health Care • Drug Discovery • Medical AI • Artificial Intelligence • Biotechnology • Est. 2019
HQ: Burlingame, United States • Pharmaceuticals • Bioinformatics • Biopharma • Health Care • Drug Discovery • Medical AI • Artificial Intelligence • Biotechnology • Est. 2019
$117k – $266k per year (Estimated) • Lead • 2+ years expRemote/Hybrid (San Mateo, New York, United States) • Master's Degree • Full-Time • Principal
Python
AI/ML
CUDA
CUDA Toolkit
PyTorch
PyTorch Lightning
Ray
Reinforcement Learning
Post-training
Diffusion Models
LLM
Quantization
RLHF
Synthetic Data
TensorRT
Triton
SFT
Apply
Report

Liquid AI

liquid.ai
Liquid AI is an artificial intelligence company headquartered in Boston, Massachusetts, and founded in 2023 as a spin-off from the MIT Computer Science and Artificial Intelligence Laboratory. The company builds Liquid Foundation Models, an architecture derived from liquid neural networks that aims to match transformer quality at a fraction of the memory and compute. It targets on-device and edge deployment where models must run on phones, vehicles, and embedded hardware rather than in a data center.
liquid.ai • HQ: Boston, United States • AI Infrastructure • Hardware • Machine Learning • Edge AI • LLM & Generative AI • Artificial Intelligence • Est. 2023
HQ: Boston, United States • AI Infrastructure • Hardware • Machine Learning • Edge AI • LLM & Generative AI • Artificial Intelligence • Est. 2023
Open 397 daysVerified live · 2 hours ago 1 year ago

Member of Technical Staff - GPU Performance Engineer

$193k – $390k per year (Estimated) • Staff+ • Remote/Hybrid (San Francisco, Boston, United States) • Full-Time • Staff
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
PyTorch
CUDA
cuDNN
LLM Guardrails
Post-training
Triton
Apply
Report

CADDi

caddi.com
CADDi is an AI data platform for manufacturing that analyzes and connects manufacturing data fragmented across departments, locations, and systems, turning drawing data into enterprise assets. It offers modular products including a discovery engine (CADDi Explorer), an AI agent (CADDi Agent), design reuse and composer tools, design and cost review platforms, production readiness and asset lifecycle management, and RFQ/quote automation to optimize procurement and manufacturing workflows. CADDi serves industrial manufacturing sectors such as automotive, construction equipment, plant and chemical, heavy industry and electrical equipment, and machinery and devices, enabling design reuse, cost review, and production readiness across supplier networks.
caddi.com • Tokyo • Chicago • Bangkok • Ho Chi Minh City • Hanoi • Software • Artificial Intelligence • Manufacturing
Tokyo • Chicago • Bangkok • Ho Chi Minh City • Hanoi • Software • Artificial Intelligence • Manufacturing
Open 409 daysVerified live · 1 day ago 1 year ago

VN Technology Software Engineer for CAD/3D Data

In office (Ho Chi Minh City, Vietnam) • Relocation • Full-Time • English
C++
C++
OpenGL
AI/ML
CUDA Toolkit
OpenCL
CUDA
DevOps
CI/CD
Docker
Git
GitHub Actions
GitHub
AWS
GCP
Design
SolidWorks
Apply
Report

Genmo

genmo.ai
Genmo is a San Francisco company founded in 2021 that develops open video generation models. Its Mochi model was released with open weights and became one of the stronger freely available text-to-video systems, used widely by researchers and creative developers. The company positions open models as the foundation for a broader generative video ecosystem.
genmo.ai • HQ: San Francisco, United States • LLM & Generative AI • Artificial Intelligence • Generative Video • Est. 2021
HQ: San Francisco, United States • LLM & Generative AI • Artificial Intelligence • Generative Video • Est. 2021
Open 409 daysVerified live · 2 days ago 1 year ago

GPU Performance Engineer

$159k – $309k per year (Estimated) • Senior • 5+ years expIn office (San Francisco, United States) • Bachelor's Degree • Full-Time • Senior
C++
Python
AI/ML
CUDA
CUDA Toolkit
Triton
InfiniBand
Frontend
Sass
Apply
Report

Xylem

xylem.com
Xylem is a global water technology provider headquartered in Washington, D.C., and established in 2011 as a spinoff from ITT Corporation. The company designs and manufactures a wide range of products including water pumps, treatment systems, smart metering technology, and analytical instruments for the entire water cycle. Operating in over 150 countries through numerous specialized brands, the firm serves municipal, industrial, and commercial customers with a focus on sustainable water management and infrastructure.
xylem.com • HQ: Washington, United States • Irrigation • Farming & Agriculture • Industrial Machinery • Manufacturing • Water Resources • Natural Resources • Energy & Utilities • Water Utilities • 5000+ employees • Est. 2011
HQ: Washington, United States • Irrigation • Farming & Agriculture • Industrial Machinery • Manufacturing • Water Resources • Natural Resources • Energy & Utilities • Water Utilities • 5000+ employees • Est. 2011
Open 480 daysVerified live · 1 day ago 1 year ago

Junior Software Developer

$17k – $46k per year (Estimated) • Junior • 3+ years expRemote/Hybrid (Bengaluru, India) • Bachelor's Degree • Full-Time • Junior
C++
C++
PyTorch C++
TensorFlow C++
AI/ML
Computer Vision
CUDA
CUDA Toolkit
ONNX
OpenCV
PyTorch
TensorFlow
DevOps
Git
Apply
Report

WeRide

weride.ai
To Transform Urban Living with Autonomous Driving WeRide - A Leading Global Autonomous Driving Technology Company WeRide (NASDAQ: WRD, HKEX: 0800) is a leading global commercial-stage company that develops autonomous driving technologies from Level 2 to Level 4.
weride.ai • San Jose • Dubai • Kuala Lumpur • Tokyo • Riyadh • Transportation & Logistics • Autonomous Driving • Artificial Intelligence • 51-200 employees • Est. 2017
San Jose • Dubai • Kuala Lumpur • Tokyo • Riyadh • Transportation & Logistics • Autonomous Driving • Artificial Intelligence • 51-200 employees • Est. 2017
Open 573 daysVerified live · 1 day ago 1 year ago

Machine Learning Engineer - Singapore

$81k – $227k per year (Estimated) • Middle • 3+ years expIn office • Bachelor's Degree • Full-Time • Middle
C++
MATLAB
Python
AI/ML
Computer Vision
CUDA
CUDA Toolkit
Robotics
Motion Planning
Sensor Fusion
Apply
Report

Cartesia

cartesia.ai
Cartesia is an artificial intelligence research and technology company that builds real-time, low-latency generative voice models for interactive applications. Founded by Stanford researchers, the company utilizes State Space Models (SSMs) to power its flagship text-to-speech platform, Sonic, which enables sub-second voice synthesis, instant voice cloning, and emotional expression across dozens of languages. By providing real-time speech generation and voice agent infrastructure, Cartesia allows developers and enterprises to power conversational AI systems, virtual assistants, and customer service platforms.
cartesia.ai • HQ: San Francisco, United States • LLM & Generative AI • Artificial Intelligence • Speech & Audio AI • 11-50 employees • Est. 2022
HQ: San Francisco, United States • LLM & Generative AI • Artificial Intelligence • Speech & Audio AI • 11-50 employees • Est. 2022
Open 626 daysTop 25% pay 2 years ago

Inference Engineer

$180k – $250k per year • Equity • In office (San Francisco, United States) • Visa sponsorship • Full-Time
Databricks
AI/ML
CUDA Toolkit
Multimodal AI
SGLang
Transformers
vLLM
CUDA
Triton
Edge AI
Apply
Report

Chai Discovery

chaidiscovery.com
Chai Discovery is a San Francisco research company founded in 2024 that builds foundation models for molecular structure and interaction prediction. Its Chai-1 model predicts complexes of proteins, small molecules, nucleic acids and antibodies, and was released for free non-commercial use alongside a commercial offering. The company is backed by OpenAI and Thrive Capital and targets the design of antibodies and other biologics.
chaidiscovery.com • HQ: San Francisco, United States • Artificial Intelligence • Est. 2024
HQ: San Francisco, United States • Artificial Intelligence • Est. 2024
Open 725 daysVerified live · 12 min ago 2 years ago

Research Engineer

$159k – $348k per year (Estimated)In office (San Francisco, United States) • Bachelor's Degree • Full-Time
Python
AI/ML
CUDA
CUDA Toolkit
PyTorch
DevOps
Kubernetes
SLURM
HPC
Analytics
ETL/ELT
Apply
Report

Pony.ai

pony.ai
Pony.ai is a global leader in autonomous driving technology that develops full-stack software and hardware platforms for self-driving vehicles. The company operates driverless Robotaxi services and autonomous freight trucking networks across major international markets. Leveraging partnerships with top automakers and industrial partners, Pony.ai aims to make transportation safer, more efficient, and commercially scalable worldwide.
pony.ai • HQ: Fremont, United States • Artificial Intelligence • Autonomous Driving • Est. 2016
HQ: Fremont, United States • Artificial Intelligence • Autonomous Driving • Est. 2016
Open 831 dayVerified live · 1 day ago 2 years ago

(Senior) Software Engineer, Deep Learning

$140k per year • Equity • Senior • 2+ years expIn office (Fremont, United States) • Full-Time • Senior
C++
Python
AI/ML
Reinforcement Learning
CUDA Toolkit
TensorRT
CUDA
Robotics
Imitation Learning
Reinforcement Learning
Motion Planning
Apply
Report

Flexcompute

flexcompute.com
Flexcompute is a computer-aided engineering (CAE) and deep-tech simulation software company. Headquartered in Watertown, Massachusetts (in the Greater Boston area), and spun out by researchers from MIT and Stanford, the company develops high-performance physics simulation engines that accelerate hardware R&D across aerospace, defense, automotive, photonics, and semiconductor sectors.
flexcompute.com • HQ: Watertown, United States • Media & Entertainment • Publishing • Artificial Intelligence • Hardware
HQ: Watertown, United States • Media & Entertainment • Publishing • Artificial Intelligence • Hardware
Open 911 daysVerified live · 1 day ago 2 years ago

CFD Solver Developer

$103k – $203k per year (Estimated) • Junior • In office • Master's Degree • Full-Time
C++
Python
C
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenMP
Management
Linear
Apply
Report

Mat3ra

mat3ra.com
Mat3ra, formerly Exabyte.io, provides a cloud platform for materials modelling that combines physics simulation with machine learning. Founded in 2014 in San Francisco, it lets semiconductor and chemicals teams run high-throughput calculations without managing clusters. The platform stores results in a searchable materials data bank.
mat3ra.com • HQ: San Francisco, United States • Materials Science & Nanotechnology • Machine Learning • Est. 2014
HQ: San Francisco, United States • Materials Science & Nanotechnology • Machine Learning • Est. 2014
Open 1367 daysTop 25% pay 4 years ago

☁ Cloud HPC Engineer ☁

$130k – $200k per year • Equity • Remote/Hybrid • Bachelor's Degree • Full-Time
Bash
Python
AI/ML
CUDA
CUDA Toolkit
OpenMP
DevOps
Ansible
AWS
Azure
Docker
GCP
Red Hat
SaltStack
GitHub
HPC
Apply
Report

Runway

runway.team
Runway is a mobile release management company headquartered in San Francisco, California, and founded in 2020. The company builds a platform that coordinates app store releases, automating build tracking, regression checks, staged rollouts, and the cross-team communication that surrounds a mobile launch. It integrates with the app stores, continuous integration systems, issue trackers, and chat tools that mobile teams already use, and went through Y Combinator.
runway.team • HQ: San Francisco, United States • Business Process Automation (BPA) • Project Management • Information Technology • DevOps • Software • Mobile App Development • Est. 2020
HQ: San Francisco, United States • Business Process Automation (BPA) • Project Management • Information Technology • DevOps • Software • Mobile App Development • Est. 2020
Open 1584 daysTop 25% pay 4 years ago

Member of Technical Staff, Applied Research Scientist

$260k – $370k per year • Staff+ • 4+ years exp • Remote (United States) • Full-Time • Staff
C++
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
Computer Vision
CUDA Toolkit
PyTorch
Runway
TensorFlow
CUDA
Human-in-the-Loop
Apply
Report

Elicit

elicit.com
Elicit is a research company that builds an artificial intelligence assistant for systematic literature review. Its product searches millions of papers, extracts methods, populations and findings into structured tables and drafts evidence summaries with citations. It grew out of the non-profit research organisation Ought, founded in 2018, and is used by academic, medical and policy researchers.
elicit.com • HQ: Oakland, United States • 11-50 employees • Est. 2018
HQ: Oakland, United States • 11-50 employees • Est. 2018
Open 2046 daysVerified live · 2 days ago 5 years ago

Machine Learning Engineer

$185k – $220k per year • In office (Oakland, United States) • Relocation • Full-Time
Python
AI/ML
CUDA
CUDA Toolkit
Fine-tuning
Tokenization
AI Agents
Apply
Report

Synopsys

synopsys.com
Synopsys (NASDAQ: SNPS) is a premier American multinational technology corporation and the global leader in Electronic Design Automation (EDA), semiconductor intellectual property (IP), and software security and quality solutions. Headquartered in Sunnyvale, California, and led by CEO Sassine Ghazi, Synopsys operates on a B2B enterprise software, proprietary semiconductor IP licensing, and silicon engineering services model.
synopsys.com • HQ: Sunnyvale, United States • Design & Creative • Hardware • Artificial Intelligence • 501-1000 employees
HQ: Sunnyvale, United States • Design & Creative • Hardware • Artificial Intelligence • 501-1000 employees
Top 25% payVerified live · 5 hours ago 4 days ago

Synopsys

$101k – $151k per year • Junior • 2+ years expIn office • Bachelor's Degree • Junior
C++
Fortran
C
C++
CMake
Conan
C
MPI
AI/ML
CUDA
CUDA Toolkit
OpenMP
DevOps
Azure
Azure DevOps
CI/CD
Docker
Git
HPC
Apply
Report
CUDA Toolkit Jobs - Remote & On-site
Frequently asked questions
CUDA Toolkit: How many jobs are available now?
There are 740 active CUDA Toolkit jobs listed on Alion right now.
CUDA Toolkit: What is the typical salary for CUDA Toolkit roles?
The average listed salary for CUDA Toolkit roles on Alion is $270,354 USD.
CUDA Toolkit: Are remote, relocation, or visa options available?
Many CUDA Toolkit listings include remote or hybrid options; some employers may offer relocation or visa sponsorship-review each posting for specifics.
CUDA Toolkit: How do I apply to CUDA Toolkit jobs on Alion?
Open the CUDA Toolkit listing on Alion and click the application link or follow the employer's instructions; sign in to save roles and get alerts.