429,391open jobs
14,506companies
64,190added this week
Browse all
Salary
$177k – $383k per year (Estimated)
Location
In office (San Mateo)
Employment
Full-Time
Overview
Company
Impact
Profile match
Clera is a San Francisco company that runs an AI recruiting platform positioned as a talent agent rather than a job board, matching engineers and other technical candidates to roles at venture-backed startups. Its system reads a candidate's background and preferences, then represents them to hiring teams at companies funded by firms such as Andreessen Horowitz, Y Combinator, Index Ventures and General Catalyst, compressing the introduction step that traditional headhunting handles manually. The model is aimed at the segment where recruiter fees are highest and candidate supply is thinnest, and the platform handles screening, scheduling and pipeline tracking for both sides of the match.

About the Role

This is a hands-on infrastructure engineering role at an early-stage enterprise AI company building a context and data governance layer for AI agents in highly regulated industries. You will own the inference and model-serving infrastructure end to end, ensuring AI agents run reliably, accurately, and at scale in production environments where performance is non-negotiable.

What You'll Do

  • Design, build, and operate inference and model-serving infrastructure from development through production deployment.

  • Scale systems to support AI agents running reliably under increasing concurrency and production load.

  • Identify and resolve infrastructure bottlenecks in close collaboration with ML and platform engineering teams.

  • Optimize systems for latency, throughput, and reliability at scale.

What We're Looking For

  • 5 or more years building and operating machine learning inference systems, model-serving platforms, or ML infrastructure in production environments.

  • Hands-on experience designing and scaling inference serving infrastructure using tools such as TensorFlow Serving, TorchServe, Triton, KServe, or equivalent custom systems.

  • Strong systems engineering fundamentals with expertise in distributed systems, containerization, and orchestration (Docker, Kubernetes).

  • Demonstrated ability to optimize production ML systems for latency, throughput, and reliability under high concurrency.

  • Experience with cloud infrastructure platforms such as AWS, GCP, or Azure for deploying and managing ML workloads.

  • Proficiency with monitoring, observability, and debugging tools such as Prometheus, Grafana, ELK, or distributed tracing frameworks.

  • Proficiency in at least one systems programming or backend language: Python, Go, Rust, C++, or Java.

  • Experience with knowledge graphs, semantic search, or graph databases (e.g., Neo4j, Amazon Neptune) is a plus.

  • Familiarity with agentic AI systems, autonomous agents, or multi-step reasoning pipelines is a plus.

  • Experience with enterprise data infrastructure, data pipelines, or data integration platforms is a plus.

Location

This role is on-site in San Mateo, California. Visa sponsorship is not available.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
429,391 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Mateo
$51k – $131k per year (Estimated) • In office • Les Mureaux
Python
Java
C++
VHDL
Perl
Apply
Software Engineer III 3 hours ago
$103k – $180k per year • In office • Full-Time • 8+ years exp • Plano • Jersey City • Charlotte
Python
Go
AI/ML
Copilot
ChatGPT
DevOps
CI/CD
Git
Configuration Management
GitHub
Apply
$260k – $302k per year • Remote/Hybrid • Full-Time • 7+ years exp • San Francisco
Python
AI/ML
LLM
OpenAI
Apply
$48k – $60k per year • Remote/Hybrid • Full-Time • Sherbrooke
Python
Apply
Remote
Python
SQL
Python
pySpark
Databases
Apache Iceberg
AI/ML
Spark
Scikit-learn
AI Agents
TensorFlow
PyTorch
RAG
DevOps
Rest API
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Apply
$76k – $99k per year • Equity • In office • Full-Time • 1+ year exp • Berlin
Apply
$95k per year • In office • Full-Time • 3+ years exp • London
Design
Figma
Apply
$120k – $150k per year • In office • Full-Time • Miami
JavaScript
TypeScript
SQL
Node JS
Frontend
React.js
DevOps
AWS
Apply
$81k – $116k per year • Equity • In office • Full-Time • Berlin
Apply
$150k – $250k per year • In office • Full-Time • 5+ years exp
AI/ML
AI Agents
LLM Guardrails
Agentic Workflows
Analytics
Fivetran
Apply
$115k – $150k per year • Equity • In office • 4+ years exp • Bachelor's Degree • San Mateo
AI/ML
AI Agents
Physical AI
Apply
$80k – $120k per year • Equity • In office • 2+ years exp • Bachelor's Degree • San Mateo
AI/ML
AI Agents
Physical AI
Marketing
LinkedIn
Apply
$135k – $180k per year • Equity • In office • 4+ years exp • Bachelor's Degree • San Mateo
AI/ML
AI Agents
Physical AI
Apply
$100k – $120k per year • In office • Full-Time • 1+ year exp • San Mateo
Apply
$243k – $295k per year • Equity • Remote/Hybrid • 8+ years exp • San Mateo
Python
C++
AI/ML
Spark
MLFlow
Multimodal AI
Kubeflow
Ray
Synthetic Data
Human-in-the-Loop
DevOps
Amazon S3
Game Dev
Unity
Unreal Engine
Design
Blender
Apply
See all jobs
This is one of many
429,391 more open roles from verified company boards, updated every day.