657,709open jobs
38,284companies
94,121added this week
Browse all
Salary
$173k – $379k per year (Estimated)
Location
In office (San Mateo)
Employment
Full-Time
Overview
Company
Impact
Profile match
Clera is a San Francisco company that runs an AI recruiting platform positioned as a talent agent rather than a job board, matching engineers and other technical candidates to roles at venture-backed startups. Its system reads a candidate's background and preferences, then represents them to hiring teams at companies funded by firms such as Andreessen Horowitz, Y Combinator, Index Ventures and General Catalyst, compressing the introduction step that traditional headhunting handles manually. The model is aimed at the segment where recruiter fees are highest and candidate supply is thinnest, and the platform handles screening, scheduling and pipeline tracking for both sides of the match.

About the Role

This is a hands-on infrastructure engineering role at an early-stage enterprise AI company building a context and data governance layer for AI agents deployed in highly regulated industries. You will own the inference and model-serving infrastructure end to end, making production AI agents fast, reliable, and scalable as concurrency grows.

What You'll Do

  • Design, build, and own inference and model-serving infrastructure from initial architecture through production deployment.

  • Scale systems that enable AI agents to run reliably and efficiently under increasing concurrent load.

  • Identify and resolve infrastructure bottlenecks in collaboration with ML and platform engineering teams.

  • Drive performance optimization across latency, throughput, and reliability for production workloads.

What We're Looking For

  • 5+ years building and operating ML inference systems, model-serving platforms, or ML infrastructure in production environments.

  • Hands-on experience designing and scaling inference-serving systems using frameworks such as TensorFlow Serving, TorchServe, Triton, KServe, or equivalent custom solutions.

  • Strong distributed systems fundamentals, including experience managing concurrent requests and resource allocation under load.

  • Proficiency with containerization and orchestration technologies, particularly Docker and Kubernetes, for ML workloads.

  • Experience with cloud infrastructure platforms (AWS, GCP, or Azure) for deploying and managing ML systems.

  • Solid monitoring and observability skills using tools such as Prometheus, Grafana, ELK, or distributed tracing solutions.

  • Proficiency in at least one systems or backend language: Python, Go, Rust, C++, or Java.

  • Familiarity with knowledge graphs, semantic search, or graph databases is a plus.

  • Background in agentic or autonomous AI systems, real-time inference, or enterprise data infrastructure is a plus.

Location

On-site in San Mateo, California, United States. Visa sponsorship is not available for this role.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
657,709 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Mateo
$42k – $66k per year (Estimated) • In office • Internship • Bachelor's Degree • Auburn Hills
Python
C++
AI/ML
Computer Vision
Cybersecurity
GDPR
Robotics
Sensor Fusion
Apply
Senior AI Engineer 5 hours ago
$47k – $115k per year (Estimated) • In office • Hyderabad
Python
SQL
C#
Python
FastAPI
C#
.NET
Databases
PostgreSQL
Redis
Apache Kafka
AI/ML
LangGraph
AutoGen
LangChain
Model Context Protocol
AI Agents
LiteLLM
Portkey
AWS Bedrock
LLM
RAG
Amazon SageMaker
DevOps
Rest API
OpenTelemetry
Kong
Azure
AWS
Kubernetes
Amazon EKS
Vector
FinOps
Amazon ECS
Management
SharePoint
Apply
$120k – $180k per year • In office • Full-Time • San Francisco
Python
AI/ML
Computer Vision
Apply
$130k – $160k per year • In office • Full-Time • 5+ years exp
Python
JavaScript
TypeScript
SQL
Node JS
Databases
RabbitMQ
Apache Kafka
DevOps
gRPC
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Analytics
ETL/ELT
Apply
$53k – $130k per year (Estimated) • In office • Internship • Master's Degree • Catania
Python
Verilog
SystemVerilog
Perl
Chips/EDA
Formal Verification
Apply
$55k – $80k per year • In office • Full-Time • Stockholm
Apply
$100k – $150k per year • In office • Full-Time • 2+ years exp • San Francisco
Marketing
X (Twitter)
LinkedIn
Apply
Account Executive 5 hours ago
$81k – $150k per year • In office • Full-Time • 1+ year exp • Bachelor's Degree • Munich
Marketing
Salesforce
Apply
$39k – $75k per year (Estimated) • In office • Full-Time • 1+ year exp • Berlin
Marketing
Meta Ads
LinkedIn Ads
Google Ads
LinkedIn
Apply
$63k – $92k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Munich
Apply
$170k – $185k per year • In office • 5+ years exp • San Mateo
Apply
$80k – $100k per year • Equity • In office • Full-Time • 1+ year exp • San Mateo
Apply
$149k – $287k per year (Estimated) • In office • Full-Time • 5+ years exp • San Mateo
AI/ML
AI Agents
Knowledge Graph
DevOps
Terraform
GCP
CloudFormation
Pulumi
Azure
CI/CD
AWS
Kubernetes
Apply
$175k – $200k per year • In office • Full-Time • New York • London • San Mateo
AI/ML
Multimodal AI
AI Agents
PyTorch
LLM
RAG
Fireworks AI
Apply
$70k – $113k per year • In office • Full-Time • San Mateo
Apply
See all jobs
This is one of many
657,709 more open roles from verified company boards, updated every day.