Updated Aug 31, 2026

Triton Jobs - Remote & On-site

Triton jobs hiring now at vetted startups and product companies on Alion, updated daily. Average salary around $267,877/year. Compare remote, hybrid and on-site positions with relocation and visa sponsorship. Browse and apply today.

Open positions
268
Companies
161
Salary range
$185K – $380K
Average salary
$272K
AI/ML
Data Science
Backend
Frontend
Mobile
DevOps
Web3
Games
Hardware
Robotics
Security
QA
Executive
Networking
Product
Design
Analytics
Support
Enterprise Apps
Quantum

Micron Technology

micron.com
Micron Technology is an American semiconductor company founded in Boise, Idaho in 1978 and one of only a handful of manufacturers of both DRAM and NAND flash memory. It designs and fabricates memory and storage products used in data centres, personal computers, smartphones, cars and industrial equipment, and sells them under the Micron and consumer-facing Crucial brands. High-bandwidth memory for AI accelerators has become the fastest-growing part of the business, and the company operates fabrication and assembly sites across the United States and Asia.
micron.com • HQ: Boise, United States • Data Centers • Semiconductor Manufacturing • Consumer Electronics • Hardware • Storage • Digital Storage • Semiconductors • 5000+ employees • Est. 1978
HQ: Boise, United States • Data Centers • Semiconductor Manufacturing • Consumer Electronics • Hardware • Storage • Digital Storage • Semiconductors • 5000+ employees • Est. 1978
Verified live · 18 min ago 3 hours ago

Intern - AI Systems and Infrastructure Engineering

Intern • In office (Austin, United States) • Master's Degree • Internship • Intern
C++
Python
C++
PyTorch C++
AI/ML
AI Agents
LLM
PyTorch
TensorRT
TensorRT-LLM
vLLM
CUDA
CUDA Toolkit
NCCL
Triton
Apply
Verified live · 18 min ago 27 days ago

MTS, AI Engineering, SMAI

Senior • 5+ years expIn office (Taichung, Taoyuan, Taiwan) • PhD • Full-Time • Senior
C++
Java
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
AI Agents
AutoGen
Chain-of-Thought
Computer Vision
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
Function Calling
Kubeflow
LangChain
LangGraph
LlamaIndex
LLM
LoRA
OpenCL
PEFT
Prompt Engineering
PyTorch
QLoRA
Ray
RLHF
Scikit-learn
TensorFlow
TensorRT
TensorRT-LLM
Triton
vLLM
Transformers
FSDP
Megatron-LM
NVLink
SFT
Frontend
Sass
Mobile
Metal
DevOps
CI/CD
Docker
Git
Jenkins
Kubernetes
SLURM
HPC
Apply
Verified live · 18 min ago 2 months ago

Member of Technical Staff (MTS), Machine Learning, SMAI

$136k – $298k per year (Estimated) • Staff+ • 10+ years expIn office (Singapore) • Full-Time • Staff
C++
Java
Python
C++
PyTorch C++
TensorFlow C++
Databases
Snowflake
AI/ML
AI Agents
AutoGen
Chain-of-Thought
Computer Vision
CrewAI
CUDA
CUDA Toolkit
DeepSpeed
Fine-tuning
Function Calling
Kubeflow
LangChain
LangGraph
LlamaIndex
LLM
LoRA
PEFT
Prompt Engineering
PyTorch
QLoRA
Ray
RLHF
Scikit-learn
TensorFlow
TensorRT
TensorRT-LLM
Triton
vLLM
Transformers
FSDP
Megatron-LM
NVLink
SFT
DevOps
CI/CD
Docker
GCP
Git
Jenkins
Kubernetes
SLURM
HPC
Analytics
ETL/ELT
Apply
Report

Keenfinity Group

keenfinity.com
Keenfinity Group is a global technology and manufacturing enterprise specializing in high-performance audio, video security, and intrusion and access control solutions. Formed from the spin-off and rebranding of Bosch's security and communications division, the company operates well-known brands such as Electro-Voice, Dynacord, RTS, Telex, Avonic, Radionix, and IQSIGHT. Operating across more than 40 countries, the firm delivers integrated hardware, artificial intelligence, and electronic manufacturing services (EMS) for critical public, commercial, and enterprise infrastructure worldwide.
keenfinity.com • Aveiro • Mississauga • Riyadh • Zhuhai • Nuremberg • Professional Services • Education • Higher Education • 1001-5000 employees • Est. 1953
Aveiro • Mississauga • Riyadh • Zhuhai • Nuremberg • Professional Services • Education • Higher Education • 1001-5000 employees • Est. 1953
Verified live · 1 hour ago 1 day ago

Embedded Software Tester (Audio_PA) (w/m/div.)

Remote/Hybrid (Aveiro, Portugal) • Full-Time • Senior • English • Portuguese
Python
AI/ML
Triton
DevOps
Git
Apply
$51k – $103k per year (Estimated) • Staff+ • 10+ years expIn office (Aveiro, Portugal) • Bachelor's Degree • Full-Time • Senior
Triton
Apply
Open 94 daysVerified live · 1 hour ago 4 months ago

High-Speed Digital PCB Layout Engineer (m/f/div.)

$39k – $95k per year (Estimated) • Senior • 5+ years expRemote/Hybrid (Ovar, Portugal) • Master's Degree • Full-Time • Senior
Triton
Chips/EDA
LTspice
Mentor Xpedition
Apply
Report

HUMBLE BEE AI

humblebee.ai
HumbleBeeAI. We specialize in Computer Vision, Generative AI, and Business Analytics. An AI studio with tools, talent, and purpose. We build from our own AI ecosystem to deliver scalable, ethical, and lasting solutions for clients worldwide.
humblebee.ai • Tashkent • 5000+ employees
Tashkent • 5000+ employees
1 day ago

NLP / LLM Engineer

$18k – $24k per year (net) • Middle • 3+ years expIn office (Tashkent, Uzbekistan) • Full-Time • Middle • English
Python
Python
FastAPI
Databases
ElasticSearch
Milvus
Pinecone
Qdrant
Weaviate
AI/ML
AI Agents
ChatGPT
DeepEval
Embeddings
Gemini
Hybrid Search
LangChain
Langfuse
LangGraph
LlamaIndex
LLM
LoRA
NLP
PEFT
Prompt Engineering
PyTorch
QLoRA
RAG
Reranking
Semantic Search
Synthetic Data
Tokenization
Triton
vLLM
Transformers
Anthropic
DPO
GraphRAG
Hugging Face
OCR
OpenAI
Semantic Search
SFT
Structured Outputs
Function Calling
TGI
DevOps
CI/CD
Docker
Git
GitHub
Analytics
A/B Testing
Apply
Report

Mirantis

mirantis.com
Mirantis is a B2B open-source cloud computing, container management, and AI infrastructure company headquartered in Campbell, California. Originally known as a core contributor to OpenStack, Mirantis now provides open-cloud software and managed services centered on Kubernetes, multi-cloud platforms, and enterprise AI workloads.
mirantis.com • Calgary • Warsaw • Brussels • Prague • Poznań • DevOps • Information Technology • Cloud Computing • 1001-5000 employees
Calgary • Warsaw • Brussels • Prague • Poznań • DevOps • Information Technology • Cloud Computing • 1001-5000 employees
Verified live · 3 hours ago 3 days ago

Director, Presales Solution Architecture - NeoCloud

$185k – $352k per year (Estimated) • Executive • Remote/Hybrid (United States) • Full-Time • Senior • English
CUDA
CUDA Toolkit
Fine-tuning
LLM
PyTorch
TensorRT
TensorRT-LLM
Triton
DevOps
Kubernetes
KubeVirt
Platform Engineering
SLURM
AWS
Apply
Verified live · 3 hours ago 10 days ago

Senior Manager, Sales Engineering — AI / GPU Cloud (NeoCloud)

$162k – $314k per year (Estimated) • Senior • Remote/Hybrid (San Jose, United States) • Full-Time • Senior
CUDA
CUDA Toolkit
Fine-tuning
LLM
PyTorch
TensorRT
TensorRT-LLM
Triton
InfiniBand
NCCL
NVIDIA NeMo
NVLink
DevOps
Kubernetes
KubeVirt
Platform Engineering
SLURM
HPC
AWS
Apply
Open 95 daysVerified live · 3 hours ago 4 months ago

Product Manager - AI Inference & Model Serving

$145k – $248k per year (Estimated) • Equity • Senior • 7+ years expAustin • Remote (United States) • Full-Time • Senior
LLM
SGLang
TensorRT
TensorRT-LLM
vLLM
Triton
DevOps
AWS
Kubernetes
Platform Engineering
Chips/EDA
PoC Library
Apply
Report

DevRev

devrev.ai
DevRev is a startup that is building AI-based CRM solutions that aim to bring developers closer to the customer enabling them to build, operate, support, and grow their products, founded by IIT alumnus.
devrev.ai • Bengaluru • Buenos Aires • Austin • Palo Alto • Mumbai • Artificial Intelligence • Software • Customer Relationship Management (CRM) • Est. 2020
Bengaluru • Buenos Aires • Austin • Palo Alto • Mumbai • Artificial Intelligence • Software • Customer Relationship Management (CRM) • Est. 2020
Verified live · 3 hours ago 1 day ago

Software Engineer - AI Performance

$28k – $59k per year (Estimated) • Staff+ • 10+ years exp • Remote (Philippines) • Staff
C#
Go
Java
Node JS
Python
JavaScript
AI/ML
AI Agents
DPO
LLM
LoRA
Ray
Triton
vLLM
PEFT
DevOps
AWS
Azure
GCP
Grafana
Kubernetes
Prometheus
Vector
Apply
$28k – $59k per year (Estimated) • Staff+ • 10+ years expPhilippines • Remote (Philippines) • Full-Time • Staff
C#
Go
Java
Node JS
Python
JavaScript
AI/ML
AI Agents
DPO
LLM
LoRA
Ray
Triton
vLLM
PEFT
DevOps
AWS
Azure
GCP
Grafana
Kubernetes
Prometheus
Vector
Apply
Report

Graphcore

graphcore.ai
Graphcore is a British semiconductor company founded in Bristol in 2016 that designs the Intelligence Processing Unit, a processor architected specifically for machine learning rather than adapted from graphics. Its chips place large amounts of memory directly on the die and expose fine-grained parallelism, an approach aimed at sparse and irregular models that map poorly onto conventional accelerators. The company sells IPU systems and the Poplar software stack to research and enterprise customers, and has operated as a wholly owned subsidiary of SoftBank Group since its acquisition in 2024.
graphcore.ai • HQ: Bristol, United Kingdom • Hardware • Artificial Intelligence • Machine Learning • AI Infrastructure • Semiconductors • 1001-5000 employees • Est. 2016
HQ: Bristol, United Kingdom • Hardware • Artificial Intelligence • Machine Learning • AI Infrastructure • Semiconductors • 1001-5000 employees • Est. 2016
Verified live · 2 hours ago 1 day ago

Engineering Manager

$91k – $153k per year (Estimated) • Lead • In office (Gdańsk, Poland) • Visa sponsorship • Staff • Polish: B2 • English: B2
C++
Python
C++
LLVM
PyTorch C++
AI/ML
PyTorch
Triton
Apply
Verified live · 2 hours ago 5 days ago

Software Engineer - Triton

$57k – $103k per year (Estimated)In office (Bristol, United Kingdom)
C++
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
JAX
PyTorch
TensorFlow
Triton
Apply
Verified live · 2 hours ago 5 months ago

PyTorch Engineer

$100k – $210k per year (Estimated)In office (Bristol, United Kingdom)
C++
Python
C++
PyTorch C++
AI/ML
PyTorch
Triton
DevOps
Git
Apply
Report
Tilde Research is a moonshot AI lab advancing mechanistic interpretability, new architectures, and pretraining science.
tilderesearch.com • San Francisco
San Francisco
Open 411 daysVerified live · 2 days ago 1 year ago

Kernel Engineer (Internship and Full-time)

Intern • In office (San Francisco, United States) • Internship • Intern
CUDA
CUDA Toolkit
PyTorch
Triton
Apply
Open 412 daysVerified live · 2 days ago 1 year ago

ML Engineer (Internship and Full-time)

Intern • In office (San Francisco, United States) • Internship • Intern
PyTorch
Triton
Post-training
Apply
Report

Ouster

ouster.com
Ouster is a San Francisco lidar company founded in 2015 that builds digital sensors based on a single-photon avalanche diode architecture. Its sensors and perception software are used in industrial automation, smart infrastructure, robotics and automotive projects. The company is listed on the New York Stock Exchange and merged with rival Velodyne in 2023.
ouster.com • HQ: San Francisco, United States • Lasers & Photonics • Sensors • Hardware • 51-200 employees • Est. 2015
HQ: San Francisco, United States • Lasers & Photonics • Sensors • Hardware • 51-200 employees • Est. 2015
Top 25% payVerified live · 1 day ago 2 days ago

Sr. Data Infrastructure & Quality Engineer - Ouster

$140k – $200k per year • Senior • 8+ years expIn office • Senior
Multimodal AI
Feature Store
Triton
Analytics
ETL/ELT
Apply
Report

Furiosa

furiosa.ai
FuriosaAI designs high-performance, power-efficient AI accelerators (NPUs) used in data centers for computer vision, GenAI, LLMs, and demanding workloads.
furiosa.ai • Seoul • Hwaseong • Santa Clara • San Jose • Data Centers • LLM & Generative AI • Artificial Intelligence
Seoul • Hwaseong • Santa Clara • San Jose • Data Centers • LLM & Generative AI • Artificial Intelligence
Verified live · 1 hour ago 13 days ago

Software Engineer, Compiler (Front-end)

In office (Seoul, South Korea) • Bachelor's Degree
C++
Python
Rust
C++
PyTorch C++
TensorFlow C++
LLVM
AI/ML
LLM
Multimodal AI
ONNX
PyTorch
TensorFlow
TPU
Triton
Apply
Verified live · 1 hour ago 10 days ago

Solutions Architect - US

$158k – $288k per year (Estimated) • Staff+ • In office (Santa Clara, United States) • Architect
Python
C++
Rust
C++
PyTorch C++
TensorFlow C++
AI/ML
AutoGen
LangChain
LangGraph
LlamaIndex
LLM
PyTorch
SGLang
TensorFlow
TensorRT
TensorRT-LLM
Triton
Triton Inference Server
vLLM
Model Context Protocol
Quantization
Apply
Verified live · 1 hour ago 13 days ago

Senior Software Engineer, Inference Engine (Platform Software)

Senior • 3+ years expIn office (Seoul, South Korea) • Bachelor's Degree • Senior
C++
Rust
AI/ML
LLM
Multimodal AI
SGLang
vLLM
CUDA
CUDA Toolkit
TensorRT
TensorRT-LLM
Triton
Apply
Report

Orcrist Technologies

orcrist.org
Orcrist Technologies offers pioneering AI and data analytics solutions in the private and public sectors, turning sensors into strategy. We are a Berlin-based data defense technology company building AI-powered software for real-time situational awareness and sensor fusion. Our mission is to give decision-makers the clarity they need-when it matters most.
orcrist.org • Berlin • Dresden • Defense AI • Data & Analytics • Artificial Intelligence
Berlin • Dresden • Defense AI • Data & Analytics • Artificial Intelligence
Verified live · 10 hours ago 3 months ago

Infrastructure/Systems Engineer

$69k – $139k per year (Estimated) • Senior • 5+ years expRemote/Hybrid (Berlin, Germany) • Internship • Senior • German: B1
Bash
Python
AI/ML
KServe
vLLM
Triton
InfiniBand
CUDA Toolkit
LLM
Quantization
TensorRT
TensorRT-LLM
CUDA
NCCL
NVLink
DevOps
Ansible
Kubernetes
Terraform
Ubuntu
CI/CD
Git
Red Hat
Cybersecurity
ISO 27001
Apply
Report

ElevenLabs

elevenlabs.io
ElevenLabs is transforming human-technology interaction with AI research and products. Powering the best enterprises, creators, and developers. From ElevenAgents for customer experience, ElevenCreative for content creation, to the leading AI voice generator.
elevenlabs.io • HQ: New York, United States • HQ: London, United Kingdom • AI Agents • LLM & Generative AI • Artificial Intelligence • Speech & Audio AI • 51-200 employees • Est. 2022
HQ: New York, United States • HQ: London, United Kingdom • AI Agents • LLM & Generative AI • Artificial Intelligence • Speech & Audio AI • 51-200 employees • Est. 2022
Verified live · 2 hours ago 4 days ago

Research Engineer - Inference

$105k – $220k per year (Estimated) • Remote (United States, United Kingdom, Bulgaria, Poland) • Full-Time
CUDA
CUDA Toolkit
ElevenLabs
Knowledge Distillation
Quantization
SGLang
TensorRT
Triton
vLLM
DevOps
GitHub
Apply
Report

Zensors

zensors.com
zensors.com • San Francisco • Artificial Intelligence • Computer Vision • 11-50 employees
San Francisco • Artificial Intelligence • Computer Vision • 11-50 employees
Top 25% payVerified live · 2 hours ago 4 days ago

AI/ML Infrastructure Engineer

$150k – $240k per year • Middle • 3+ years expIn office (San Francisco, United States) • Bachelor's Degree • Full-Time • Middle
C++
Python
SQL
C++
PyTorch C++
Databases
Neo4j
AI/ML
Computer Vision
CUDA
CUDA Toolkit
Knowledge Distillation
PyTorch
Quantization
TensorRT
Transformers
Vision Transformer
ONNX
Triton
Triton Inference Server
DevOps
Platform Engineering
Apply
Report

Beacon AI

beaconai.com
Beacon AI is an aviation technology company that builds AI-powered pilot assistance systems and flight safety software for commercial, private, and defense aviation. Often described as an "R2-D2 for pilots", the platform acts as an intelligent onboard teammate - integrating directly with cockpit infrastructure to assist aviators with real-time checklists, route optimization, fleet monitoring, and situational awareness to reduce human error and workload.
beaconai.com • HQ: San Francisco, United States • Artificial Intelligence • Space & Aerospace • Avionics & Flight Control
HQ: San Francisco, United States • Artificial Intelligence • Space & Aerospace • Avionics & Flight Control
Top 25% payVerified live · 1 hour ago 4 days ago

Software Engineer, Cloud Infrastructure (Multiple Seniority Levels)

$135k – $260k per year • Middle • 4+ years expIn office (San Carlos, United States) • Visa sponsorship • Full-Time • Middle
Python
Databases
Amazon Aurora
DynamoDB
OpenSearch
pgvector
Pinecone
PostgreSQL
Redis
AI/ML
LLM
AWS Bedrock
Embeddings
Function Calling
LangChain
Quantization
RAG
TensorRT
TensorRT-LLM
Time Series Forecasting
Triton
Triton Inference Server
Amazon SageMaker
Human-in-the-Loop
LLM Evaluation
LLM Guardrails
DevOps
AWS
Kubernetes
Amazon EKS
AWS CDK
AWS Lambda
CI/CD
GitHub Actions
OpenTelemetry
Terraform
Vector
Amazon CloudWatch
Amazon ECS
Amazon EventBridge
Amazon S3
AWS Step Functions
GitHub
IAM
Cybersecurity
Least Privilege
Analytics
A/B Testing
Apply
Top 25% payVerified live · 1 hour ago 2 months ago

Software Engineer, Artificial Intelligence/LLM (Multiple Seniority Levels)

$135k – $260k per year • In office (San Carlos, United States) • Visa sponsorship • Full-Time
Python
TypeScript
Databases
Amazon Aurora
DynamoDB
OpenSearch
pgvector
Pinecone
PostgreSQL
Weaviate
AI/ML
AWS Bedrock
Embeddings
Function Calling
Hallucination
LangChain
LLM
RAG
Time Series Forecasting
Anthropic
Human-in-the-Loop
LLM Guardrails
OpenAI
Multimodal AI
TensorRT
TensorRT-LLM
Triton
DevOps
AWS
Vector
Amazon S3
CI/CD
Analytics
A/B Testing
Apply
Report

Anthropic

anthropic.com
Anthropic is an American artificial intelligence safety and research company founded in 2021 by former OpenAI researchers, among them the siblings Dario and Daniela Amodei. It develops the Claude family of large language models and ships them through a consumer assistant, an enterprise developer platform and the Claude Code agentic coding tool, alongside open standards such as the Model Context Protocol. Incorporated as a public benefit corporation and headquartered in San Francisco, the company concentrates on interpretability, alignment and reliability research and counts Google and Amazon among its largest investors.
anthropic.com • HQ: San Francisco, United States • Code Intelligence • AI Agents • Machine Learning • LLM & Generative AI • Artificial Intelligence • 1001-5000 employees • Est. 2021
HQ: San Francisco, United States • Code Intelligence • AI Agents • Machine Learning • LLM & Generative AI • Artificial Intelligence • 1001-5000 employees • Est. 2021
Top 25% payVerified live · 1 hour ago 3 months ago

Staff+ Software Engineer, Inference Velocity

$405k per year • Staff+ • 8+ years expIn office (San Francisco, United States) • Visa sponsorship • Bachelor's Degree • Staff
Claude
CUDA Toolkit
CUDA
Anthropic
AWS Trainium
TPU
Multimodal AI
Triton
DevOps
AWS
CI/CD
HPC
Kubernetes
Platform Engineering
Apply
Open 342 daysTop 25% pay 12 months ago

Performance Engineer, GPU

$280k per year • In office (San Francisco, United States) • Visa sponsorship • Bachelor's Degree
Claude
CUDA Toolkit
Flash Attention
JAX
Multimodal AI
PyTorch
Quantization
CUDA
Triton
Anthropic
NCCL
NVLink
Apply
Report

Checkout.com

checkout.com
Checkout.com is a London payments company founded in 2012 that processes card and alternative payments for large online merchants. It built a full stack acquiring and processing platform rather than assembling third party components, which lets it optimise authorisation rates directly. Its customers include marketplaces, streaming services and digital asset platforms across Europe, the Middle East and Asia.
checkout.com • HQ: London, United Kingdom • Blockchain & Crypto • Financial Services • Artificial Intelligence • 51-200 employees • Est. 2012
HQ: London, United Kingdom • Blockchain & Crypto • Financial Services • Artificial Intelligence • 51-200 employees • Est. 2012
Verified live · 1 hour ago 4 days ago

Senior Machine Learning Engineer

$92k – $190k per year (Estimated) • Senior • 5+ years expIn office (London, United Kingdom) • Full-Time • Senior
Python
Databases
Databricks
AI/ML
Kubeflow
PyTorch
Scikit-learn
Spark
TensorFlow
Triton
Vertex AI
XGBoost
Amazon SageMaker
Feature Store
Seldon Core
DevOps
AWS
Azure
Apply
Report

Cognism

cognism.com
Cognism provides sales intelligence data with phone-verified contact details and compliance controls for European markets. Founded in 2015 in London, it built machine learning pipelines for entity resolution across company and person records. Its Diamond Data verification distinguishes it in a crowded data market.
cognism.com • HQ: London, United Kingdom • Information Security • Sales & Marketing • Data Quality • Est. 2015
HQ: London, United Kingdom • Information Security • Sales & Marketing • Data Quality • Est. 2015
Verified live · 2 hours ago 26 days ago

MLOps Engineer

$56k – $109k per year (Estimated) • Middle • 3+ years expIn office (Warsaw, Poland) • Middle • English: C1
Python
SQL
Python
FastAPI
Databases
Databricks
ElasticSearch
AI/ML
Keras
Kubeflow
MLFlow
PyTorch
Scikit-learn
TensorFlow
Triton
Vertex AI
Amazon SageMaker
Metaflow
DevOps
AWS
Azure
CI/CD
CircleCI
Docker
GCP
GitHub Actions
Grafana
Kibana
Kubernetes
Logstash
Terraform
Amazon ECS
GitHub
Apply
Report

eBay

ebay.com
eBay is a multinational e-commerce corporation that connects buyers and sellers in more than 190 markets worldwide thus enabling economic opportunity for individuals, entrepreneurs, businesses, and organizations.
ebay.com • HQ: San Jose, United States • E-commerce Platforms • Commerce • Marketplaces • 51-200 employees • Est. 1995
HQ: San Jose, United States • E-commerce Platforms • Commerce • Marketplaces • 51-200 employees • Est. 1995
$32k – $80k per year (Estimated) • Senior • 5+ years expIn office (Bengaluru, India) • Senior
Go
Python
Rust
AI/ML
Computer Vision
CUDA
CUDA Toolkit
LLM
PyTorch
Quantization
SGLang
TensorFlow
TensorRT
Triton
vLLM
DevOps
HPC
Apply
Report

Reducto

reducto.ai
Reducto is a San Francisco company founded in 2023 that converts complex documents into structured input for language models. Its pipeline handles tables, charts, forms and scanned pages with vision models, producing accurate chunks for retrieval systems where generic parsers fail. It is used by enterprises building document-grounded AI in finance, healthcare and insurance.
reducto.ai • HQ: San Francisco, United States • Information Security • LLM & Generative AI • Cybersecurity • 51-200 employees • Est. 2023
HQ: San Francisco, United States • Information Security • LLM & Generative AI • Cybersecurity • 51-200 employees • Est. 2023
Top 25% payVerified live · 2 hours ago 5 days ago

LLM/ML Engineer (Inference)

$200k – $300k per year • Equity 0.1–1% • Middle • 3+ years expIn office (San Francisco, United States) • Full-Time • Middle
Python
AI/ML
LLM
PyTorch
TensorRT
TensorRT-LLM
vLLM
TGI
CUDA
CUDA Toolkit
Triton
Apply
Report

Tzafon

tzafon.ai
Tzafon is advancing machine intelligence. Our team develops models and systems for autonomous agents interacting with the real world.
tzafon.ai • San Francisco • AI Agents
San Francisco • AI Agents
Open 325 daysTop 25% pay 11 months ago

Member of Technical Staff - Foundations

$200k – $500k per year • Equity • Staff+ • Remote/Hybrid (San Francisco, United States) • Visa sponsorship • Full-Time • Staff
Python
AI/ML
CUDA
CUDA Toolkit
Fine-tuning
JAX
PyTorch
Triton
FSDP
Post-training
Pre-training
TPU
Synthetic Data
vLLM
Anthropic
OpenAI
Apply
Report

Pear VC

pear.vc
Pear VC is a venture capital firm specializing in pre-seed and seed-stage investments in startups. Founded by Pejman Nozad and Mar Hershenson, the firm provides hands-on support in areas such as product development, recruiting, fundraising, and go-to-market strategy. Pear has backed companies including DoorDash, Gusto, Aurora Solar, Vanta, and Guardant Health.
pear.vc • HQ: San Francisco, United States • Financial Services • Venture Capital
HQ: San Francisco, United States • Financial Services • Venture Capital
Open 306 daysVerified live · 8 hours ago 11 months ago

Member of Technical Staff, Backend - NomadicML

$238k – $482k per year (Estimated) • Staff+ • In office (San Francisco, United States) • Full-Time • Staff
Go
Python
TypeScript
JavaScript
Databases
Snowflake
AI/ML
Dagster
DeepSpeed
ONNX
Ray
Multimodal AI
Ray Serve
Triton
Frontend
Next.js
React.js
DevOps
AWS
Azure
GCP
gRPC
Kubernetes
Amazon S3
IAM
Apply
Report

Salary range

Seniority Jobs 25% Median 75%
Middle 9 $120K $225K $240K
Senior 32 $242K $265K $288K
Staff 24 $234K $300K $359K
Lead 10 $232K $257K $270K
All levels 99 $230K $265K $300K

Based only on the listings that state pay. Gross annual amounts, converted to USD so roles in different currencies stay comparable.

Fully remote 13%
34
Hybrid 24%
64
On site 63%
170
Visa sponsorship 6%
16
Relocation package 6%
17
Equity 12%
32

The market right now

Alion currently lists 268 open Triton jobs from 161 companies. 32 of them were posted or refreshed in the last seven days. 113 of the employers have posted something in the last three months, which is the pool worth watching if you are starting a search now. Every listing links straight to the employer, so you apply on their own board rather than through an intermediary.

What these roles pay

99 of these Triton jobs state pay directly. Across them the middle half of the market sits between $230K and $300K a year, with a median of $265K. By level, the median runs middle at $225K, senior at $265K, staff at $300K, and lead at $257K. Figures are gross annual amounts converted to US dollars, so roles in different currencies stay comparable.

Where the work is

The largest concentrations of these roles are United States (12), Taiwan (7), Singapore (7), Germany (7), and France (4). 13% of the listings are fully remote and a further 24% are hybrid, so a large part of this market is open to you regardless of where you live.

Who is hiring

The employers with the most open Triton jobs right now are NVIDIA (20), Keenfinity Group (7), Adobe (7), Pragmatike (5), Together AI (5), and Graphcore (4). Each company page on Alion carries its size, funding stage, tech stack and every other position it has open, so you can judge the employer before you spend an evening on the application.

What employers ask for

Reading across the current listings, the tools that come up most often are Triton (268), Python (179), PyTorch (145), Kubernetes (137), CUDA Toolkit (125), CUDA (125), LLM (125), and vLLM (106). The counts are how many of these openings name each one, which is a better guide to what is actually being hired for than a generic skills list.

Relocation, visas and equity

Of the current Triton jobs, 16 state visa sponsorship, 17 offer a relocation package, and 32 include equity. These are filters on the list above, so you can narrow it to the ones that make a move possible for you.

Triton Jobs - Remote & On-site
Frequently asked questions
How many Triton jobs are open right now?
There are 268 open Triton jobs on Alion from 161 companies. The list was last refreshed on August 31, 2026.
What do Triton jobs pay?
The median is $265K a year, and the middle half of the market falls between $230K and $300K. This is based on the 99 listings that state pay.
Are any of these roles remote?
34 of the 268 listings (13%) are fully remote, and you can filter the list down to them in one click.
Which companies are hiring?
The most active employers right now are NVIDIA, Keenfinity Group, Adobe, Pragmatike, Together AI, and Graphcore. Each has its own page on Alion with the rest of its open roles.
Can I get visa sponsorship or relocation?
Yes - of the current listings, 16 state visa sponsorship and 17 offer a relocation package. Both are filters on the jobs page.
How do I apply?
Open any listing and apply on the employer's own board through the link on the page. No account is required to browse or to apply.