Salary
≈ $155k – $297k per year (Estimated)
Location
In office (San Mateo)
Seniority
Middle · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Zaimler gives enterprise AI agents a live understanding of your business, answered from your system of record and resolved at the moment they act on it.
About the Role
You'll own our inference and model-serving infrastructure end to end. This isn't a research role. It's a build role: you're setting up and scaling the systems that let our agents actually run in production, fast and reliably, at increasing concurrency.
You report to Sofus and work closely with our ML and infra teams.
What You’ll Own
- Set up and scale inference/Ray Serve for ML and LLM model serving, integrated with our data analysis and agent workflows
- Scale agent GPU infrastructure for concurrency and efficiency across multiple agent workloads
- Optimize and improve the engine builder and model server that power scalable agent orchestration
What You Need
- Proven ability to build scalable ML/AI platforms from scratch, end-to-end, for production use cases. You've owned a zero-to-one build before, or can show you're capable of it
- Deep understanding of the inference stack: vLLM, KV cache, and the optimization layers underneath model serving
- Experience building distributed systems for AI/ML workloads at scale, connecting them to real product or vertical integrations
- 3+ years of relevant experience. We care about capability, not tenure
Nice to Have
- Ray / Ray Serve experience
- Familiarity with AIBrix
Why Join
- A rare chance to shape both company and product direction as an early team engineer
- Work alongside engineers and researchers from LinkedIn, Visa, Meta, and Branch
- Onsite culture in San Mateo, built for deep collaboration and high-velocity building
- Full benefits (medical, dental, vision, 401k)
- We sponsor H-1B visas and assist with immigration
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
686,059 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
San Mateo
Senior AI Engineer
1 hour ago
≈ $38k – $96k per year (Estimated) • In office • 3+ years exp • Bengaluru
Python
Python
FastAPI
Django
Databases
Redis
RabbitMQ
Apache Kafka
AI/ML
LangGraph
LangChain
Claude
vLLM
AI Agents
AWS Bedrock
Gemini
LLM
OpenAI
Multi-Agent Systems
DevOps
Datadog
Prometheus
WebSockets
AWS
Docker
Kubernetes
Grafana
Apply
Senior NLP Data Scientist, Quant
1 hour ago
≈ $38k – $73k per year (Estimated) • In office • Moscow
Python
AI/ML
Fine-tuning
NLP
NER
LLM
RAG
Apply
Ведущий программист (.NET C#)
1 hour ago
≈ $112k – $246k per year (Estimated) • Remote • Full-Time • 5+ years exp
C#
C#
.NET
Databases
RabbitMQ
MS SQL
AI/ML
Copilot
Cursor
ChatGPT
Claude Code
Model Context Protocol
LLM
RAG
DevOps
CI/CD
Git
Apply
Ведущий программист (.NET C#)
1 hour ago
≈ $82k – $207k per year (Estimated) • Remote • Full-Time • 5+ years exp • Tbilisi
C#
C#
.NET
Databases
RabbitMQ
MS SQL
AI/ML
Copilot
Cursor
ChatGPT
Claude Code
Model Context Protocol
LLM
RAG
DevOps
CI/CD
Git
Apply
Data Analyst (NLP / LLM)
1 hour ago
$17k per year (net) • In office • Bachelor's Degree • Saint Petersburg
Python
SQL
AI/ML
Polars
LangChain
Scikit-learn
NLP
Pandas
NumPy
LLM
Hugging Face
Analytics
Seaborn
Matplotlib
Plotly
Apply
Forward Deployed Engineer
6 hours ago
≈ $153k – $276k per year (Estimated) • In office • Full-Time • 5+ years exp • New York
Python
SQL
AI/ML
AI Agents
LLM
Knowledge Graph
Apply
Forward Deployed Engineer
1 month ago
≈ $96k – $198k per year (Estimated) • In office • Full-Time • 5+ years exp • London
Python
SQL
AI/ML
AI Agents
LLM
Knowledge Graph
Apply
Data Infrastructure Engineer (Query Engine)
1 month ago
≈ $123k – $271k per year (Estimated) • In office • Full-Time • San Mateo
Databases
Snowflake
Databricks
Apply
Director of Marketing
6 months ago
≈ $174k – $330k per year (Estimated) • In office • Full-Time • 6+ years exp • San Mateo
AI/ML
AI Agents
RAG
Apply
Head of Product
7 months ago
≈ $185k – $351k per year (Estimated) • In office • Full-Time • 10+ years exp • San Mateo
AI/ML
Knowledge Graph
Apply
Director, Finance - Connected Products
12 hours ago
$140k – $220k per year • Remote/Hybrid • Full-Time • 10+ years exp • Master's Degree • San Mateo
Apply
$118k – $148k per year • Equity • Remote/Hybrid • San Mateo
Apply
$243k – $295k per year • Equity • Remote/Hybrid • 6+ years exp • San Mateo
Rust
C++
Databases
MySQL
PostgreSQL
RocksDB
Amazon Aurora
Google Cloud Spanner
DevOps
Nomad
CI/CD
Kubernetes
Apply
Patient Services Specialist
1 day ago
≈ $47k – $112k per year (Estimated) • In office • Part-Time • High School Diploma • San Mateo
Apply
$121k – $148k per year • In office • 8+ years exp • Bachelor's Degree • San Mateo
Marketing
Salesforce
Apply
This is one of many
686,059 more open roles from verified company boards, updated every day.

