368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$150k – $240k per year
Location
In office (San Francisco)
Seniority
Middle · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

Zensors is the spatial intelligence platform for the physical world. Our AI platform provides real-time insights-from airport queue times to office utilization-helping organizations make smarter operational decisions.

Zensors is processing massive streams of video data 24/7 with human-level accuracy. To do this at scale, we rely on cutting-edge optimization to ensure our vision transformer and detection models run efficiently on both cloud and edge compute resources.

Responsibility

The AI Infrastructure team at Zensors builds the engine that powers our visual sensing platform. We provide the tools to automate the lifecycle of our AI workflow, including model development, evaluation, optimization, deployment, and monitoring across thousands of video streams.

As a Machine Learning Engineer in ML Runtime & Optimization, you will develop technologies to accelerate the training and inference of computer vision models that power smart spaces and cities.

Your responsibilities will include:

  • Optimizing Core ML Pipelines: Identifying key bottlenecks in our current video analytics pipeline and performing in-depth analysis to ensure the best possible performance on current server and edge compute architectures.

  • Cross-Stack Collaboration: Collaborating closely with AI research and platform engineering teams to optimize core parallel algorithms and influence the design of our next-generation inference infrastructure.

  • Model Acceleration: Applying advanced model optimization techniques-such as quantization (Int8/FP16), pruning, and layer fusion-to our Vision Transformers (ViTs) and CNNs to maximize throughput and minimize latency.

  • Building Efficient Operators: Working across the entire ML framework/compiler stack (e.g., PyTorch, CUDA, TensorRT, and NVIDIA DeepStream) to write custom optimized ML operator libraries.

  • Resource Efficiency: Reducing the compute cost per video stream to enable massive scalability of our SaaS product.

  • Data Management: Building, improving, maintaining, and operating systems to facilitate the collection, labeling, and use of visual data for ML training.

Requirements

  • BS/MS or Ph.D. in Computer Science, Electrical Engineering, or a related discipline.
  • Strong programming skills in C/C++ and Python.
  • Experience with model optimization, quantization, and efficient deep learning techniques (e.g., knowledge distillation, pruning).
  • Deep understanding of GPU hardware performance, including execution models, thread hierarchy, memory/cache management, and the cost/performance trade-offs of video processing.
  • Experience with profiling and benchmarking tools (e.g., Nsight Systems, Nsight Compute) to validate performance on complex architectures.
  • Experience identifying and resolving compute and data flow bottlenecks, particularly in high-bandwidth video processing pipelines.
  • Strong communication skills and the ability to work cross-functionally between research and infrastructure teams.

Preferred Qualifications

  • Familiarity with database systems (e.g., SQL, Neo4j).
  • Work in Computer Vision, Deep Learning, and Vision Transformers.
  • Experience with video processing frameworks such as NVIDIA DeepStream, DALI, or FFmpeg.
  • Familiarity with ML compilers (e.g., TVM, MLIR) or inference engines like TensorRT or ONNX Runtime.
  • Knowledge of distributed training systems or cloud-scale inference serving (e.g., Triton Inference Server).
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$67k – $160k per year (Estimated) • In office • Full-Time • France
C++
C++
PyTorch C++
TensorFlow C++
AI/ML
Computer Vision
Multimodal AI
OpenCV
PyTorch
TensorFlow
Edge AI
Apply
$72k – $168k per year (Estimated) • Equity • Remote • Full-Time
C++
Go
Java
Python
Rust
Scala
Apply
$58k – $171k per year (Estimated) • In office • Full-Time • 1+ year exp • Munich
Python
AI/ML
Computer Vision
ONNX
PyTorch
DevOps
Docker
SLURM
Apply
$19k – $51k per year (Estimated) • Remote/Hybrid • Full-Time • Moscow
C++
C
C++
STL
C
Valgrind
DevOps
RTOS
IoT
FreeRTOS
Apply
$120k – $220k per year • Equity 0.2–0.8% • Remote/Hybrid • Full-Time • 6+ years exp • San Francisco
C++
Go
JavaScript
Python
TypeScript
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$185k – $385k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
JavaScript
Python
Databases
MySQL
PostgreSQL
AI/ML
OpenAI
Frontend
React.js
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.