368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$220k – $260k per year
Location
In office (San Francisco)
Seniority
Staff · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
E-invoicing, B2B payments and digitalization without friction. Voxel, the digital transformation in the travel and Food & Hospitality sectors.

Who We Are

Voxel is building the future of Computer Vision and Machine Learning for operations, risk, and safety. We use computer vision and AI to enable existing security cameras to automatically detect hazards and high-risk activities, keep people safe and drive operational efficiencies. Our technology addresses the key cost drivers for workers’ compensation, general liability, and property damage, which cost US employers over $500 billion annually. Our customers include Fortune 500 companies across grocery, retail, manufacturing, food and beverage, logistics, and pharmaceutical distribution. We’ve passed $10M ARR with strong expansion revenue. Based in SF, backed by industry-leading VCs.

About the Role

Voxel’s perception system is the technical core of everything we ship. Our models detect human activity, equipment interactions, environmental hazards, and operational state in real time across thousands of cameras in manufacturing, logistics, retail, and pharmaceutical environments. Safety was our wedge; it proved our platform works. Now customers are pulling us into operations: equipment utilization, workflow compliance, process efficiency. Every new use case runs through the perception team.

We're hiring a Staff Software Engineer to own ML Infrastructure at Voxel. Our applied ML team is shipping vision models into production every week, across thousands of cameras at Fortune 500 customers, and the infrastructure underneath determines how fast we can move. You'll set the technical direction for how we train, track, and ship vision models, build the foundational systems that the applied ML team relies on, and shape the architectural decisions that will define our ML stack for the next several years.

This is a hands-on role. You'll write code, make architecture calls, and own outcomes end to end. You'll partner closely with applied CV engineers, the ML Data team, and the Platform team, and you'll be the technical voice in the room when ML infrastructure tradeoffs come up.

What You'll Do

  • Set the technical direction for ML infrastructure at Voxel: what we build, what we buy, and how the pieces fit together as the team and model portfolio scale

  • Architect and build the training infrastructure that lets the applied ML team run multiple experiments concurrently and iterate quickly on new architectures (PyTorch, AWS)

  • Own the train-to-deploy handoff: export trained models to optimized inference formats (TensorRT, ONNX), quantify accuracy and latency impact, and partner with Platform on production deployment

  • Pick and roll out the experiment tracking and lifecycle stack (Weights & Biases, MLflow, ClearML, or similar) so researchers can run, compare, and reproduce experiments efficiently

  • Establish DevOps-for-ML best practices (IaC, CI/CD, observability, cost monitoring) so researchers can iterate quickly and safely

  • Mentor engineers across Vision & AI on ML infrastructure best practices, raising the bar for how the org thinks about training, evaluation, and deployment

  • Anticipate where the infrastructure needs to be in 12 to 18 months, including the upcoming move to on-device inference, and architect for that future

What We're Looking For

  • 7+ years building and shipping large-scale software systems, with at least 3 years focused on ML infrastructure or large-scale data infrastructure

  • A track record of being the person who decides the architecture, not just the person who implements it. You've owned tool selection, framework choices, and build-vs-buy calls for systems other engineers depend on

  • Deep fluency in PyTorch and the modern ML training stack. You know what good experiment tracking looks like, what makes a training pipeline reliable at scale, and where the failure modes live

  • Strong Python. Performant, maintainable code that holds up in production

  • A pragmatic shipping orientation. You can tell the difference between architectural decisions that need to be right and ones that can be revisited later, and you don't over-engineer the latter

  • Strong communication skills. You can explain complex tradeoffs clearly to ML researchers, infra peers, and leadership

Nice to Have

  • Production experience on AWS (S3, EC2, EKS, or similar) for ML workloads

  • Hands-on experience with model export and inference optimization (TensorRT, ONNX, or similar), including measuring accuracy and latency tradeoffs against training-time baselines

  • Experience with modern ML orchestration tools (Ray, Sematic, Flyte, Metaflow, Prefect, or similar)

  • Familiarity with GPU performance profiling and optimization (Nsight, PyTorch profiler, or similar)

  • Background in computer vision model training

Compensation & Benefits

  • Equity through Voxel’s Equity Incentive Plan

  • Total compensation includes base salary, annual bonus, and equity

  • Comprehensive health, dental, and vision insurance

  • Competitive paid parental leave

  • Unlimited PTO and flexible work arrangements

  • Daily meals in-office, team events, annual company onsite

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$20k – $49k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
Ruby
SQL
Databases
Amazon Neptune
Neo4j
AI/ML
Hallucination
LangChain
LangGraph
LLM
Model Context Protocol
Spark
AI Agents
LLM Guardrails
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
GitHub Actions
GitLab CI
Jenkins
Kubernetes
Rest API
Terraform
GitHub
GitLab
QA
Playwright
Postman
Selenium
Swagger
Apply
Remote • Full-Time • 3+ years exp • Cairo
C#
JavaScript
TypeScript
C#
.NET
AI/ML
Copilot
DevOps
Azure
Azure DevOps
CI/CD
Rest API
Analytics
ETL/ELT
Management
Power Apps
Power Automate
Apply
Remote • Full-Time • 3+ years exp • Cairo
C#
JavaScript
TypeScript
C#
.NET
AI/ML
Copilot
DevOps
Azure
Azure DevOps
CI/CD
Rest API
Analytics
ETL/ELT
Management
Power Apps
Power Automate
Apply
Remote • Full-Time • 1+ year exp • Cairo
C#
JavaScript
TypeScript
C#
.NET
AI/ML
Copilot
DevOps
Azure
Azure DevOps
CI/CD
Rest API
Analytics
ETL/ELT
Management
Power Apps
Power Automate
Apply
SOC Senior Analyst 2 days ago
In office • Full-Time • Bachelor's Degree • Riyadh
DevOps
AWS
Azure
GCP
Cybersecurity
MITRE ATT&CK
Apply
$130k – $180k per year • Equity • In office • Full-Time • 1+ year exp • Bachelor's Degree • San Francisco
C++
Python
AI/ML
Anomaly Detection
Computer Vision
DevOps
Amazon EKS
AWS
Docker
Git
Kubernetes
Amazon ECS
Amazon S3
Robotics
Localization
Apply
$125k – $150k per year • Equity • In office • Full-Time • 2+ years exp
Python
SQL
TypeScript
AI/ML
Computer Vision
DevOps
AWS
Azure
Datadog
GCP
Grafana
SRE
Apply
$200k – $240k per year • Equity • In office • Full-Time • 4+ years exp • San Francisco
Python
AI/ML
ClearML
Computer Vision
MLFlow
ONNX
Prefect
PyTorch
Ray
TensorRT
Weights & Biases
Flyte
Metaflow
DevOps
Amazon EC2
Amazon EKS
AWS
CI/CD
Kubernetes
Amazon S3
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$185k – $385k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
JavaScript
Python
Databases
MySQL
PostgreSQL
AI/ML
OpenAI
Frontend
React.js
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.