368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$160k – $250k per year
Location
In office (San Francisco)
Seniority
Staff · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Clera is a San Francisco-based AI recruiting and talent-matching platform designed as an AI talent agent for candidates and hiring teams. Acting as a tech-driven alternative to traditional headhunting, Clera directly connects job seekers to open roles at top startups backed by venture firms like Andreessen Horowitz (a16z), Y Combinator, Index Ventures, and General Catalyst.

About the Role

A well-funded, early-stage AI startup in the mechanical engineering software space is looking for a Staff Engineer - Agentic AI to own the core agent intelligence layer that turns engineers' intent into reliable, cost-efficient multi-step workflows across complex desktop engineering tools. This is a high-impact, senior technical leadership role reporting directly to the CTO, sitting at the intersection of applied agentic AI, user research, and product delivery.

The company serves Fortune 100 hardware engineering customers and is backed by notable investors. You'll join a small, senior team and have a direct line to executive leadership. The role is on-site in San Francisco, CA.

What You'll Do

  • Lead development of the core agent intelligence layer executing multi-step workflows across complex desktop engineering software (CAD, CAE, PLM).

  • Report to the CTO and serve as technical lead for a small team of AI engineers, a user researcher, and domain expert contractors.

  • Own the full product loop: define agent capabilities from user stories, build implementations, and benchmark against real workflows.

  • Drive agent task success rate - define the eval framework, establish baselines, and systematically improve completion metrics.

  • Set and enforce per-task token budgets; track cost per completed workflow to ensure commercial viability.

  • Build rigorous, reproducible evaluation infrastructure grounded in validated user stories (SWE-bench-level rigor applied to engineering workflows).

  • Lead user story mapping and validation through direct interviews and collaboration with domain experts.

  • Translate validated user stories into testable evals, closing the loop between research and benchmarking.

  • Own agent architecture decisions: tool-calling strategies, state management, error recovery, model routing, and context management.

  • Act as a player-coach: write production code, review designs, unblock the team, and raise engineering standards.

  • Collaborate cross-functionally with integrations, product, and customers during POCs to align agent behavior with real-world usage.

What We're Looking For

Required (Dealbreakers):

  • 7+ years in software engineering, including at least 2 years building agentic LLM-based agents that act in the real world (tool-calling, multi-step workflows, failure handling, cost constraints).

  • Deep experience designing LLM application architectures: model selection, context/window management, retrieval strategies, tool-calling frameworks, and orchestration patterns.

  • Strong evaluation and benchmarking instincts for agentic systems - task completion, cost efficiency, failure mode analysis; familiarity with SWE-bench, GAIA, or τ-bench.

  • Proven track record shipping AI systems with measurable outcomes (e.g., agent task success rate, cost efficiency) - not just demos.

  • Strong Python skills and hands-on experience with LLM tooling (function calling, tool use APIs, tracing/observability tools such as Logfire or LangSmith, evaluation frameworks).

  • Experience leading a small technical team (3-6 engineers): setting direction, performing code reviews, driving architecture decisions.

Strongly Preferred:

  • Experience with desktop automation, COM, or programmatic control of applications (beyond web APIs).

  • Background in mechanical engineering, CAD/CAE, PLM, or adjacent industries.

  • Familiarity with enterprise deployment constraints - agent behavior on locked-down corporate workstations.

  • Published work or open-source contributions in agentic AI systems.

  • Experience building or contributing to public benchmarks for AI agents.

Compensation & Benefits

  • Salary: $160,000 - $250,000 USD annually, depending on experience.

  • Equity participation in an early-stage, Series A company.

  • Note: Visa sponsorship is not available for this role.

Location

  • On-site in San Francisco, CA, United States.

  • This is not a remote role.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
Founding Engineer 2 days ago
$137k – $190k per year • In office • Full-Time • 3+ years exp • New York
Python
TypeScript
AI/ML
AI Agents
LLM
Apply
Founding Engineer 2 days ago
$150k – $220k per year • In office • Full-Time • 4+ years exp • New York
Python
AI/ML
AI Agents
Fine-tuning
LLM
Reinforcement Learning
Robotics
Reinforcement Learning
Apply
Full Stack Engineer 2 days ago
$70k – $93k per year • In office • Full-Time • 3+ years exp • Munich
Python
TypeScript
JavaScript
Databases
PostgreSQL
AI/ML
LLM
Frontend
tRPC
DevOps
CI/CD
Docker
Kubernetes
Apply
$71k – $150k per year (Estimated) • Remote/Hybrid • Full-Time • 12+ years exp • Master's Degree • Warsaw
Python
SQL
AI/ML
ChatGPT
Claude
NumPy
Pandas
Scikit-learn
Analytics
Tableau
Management
Confluence
Apply
$130k – $170k per year • In office • Full-Time • 3+ years exp • San Francisco
TypeScript
Node JS
JavaScript
Node JS
Prisma
Databases
Supabase
Typesense
AI/ML
LLM
Recommender Systems
Apply
Founding Engineer 2 days ago
$76k – $99k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Berlin
TypeScript
JavaScript
Databases
PostgreSQL
Frontend
React.js
Apply
$130k – $170k per year • In office • Full-Time • 3+ years exp • San Francisco
TypeScript
Node JS
JavaScript
Node JS
Prisma
Databases
Supabase
Typesense
AI/ML
LLM
Recommender Systems
Apply
$115k – $263k per year (Estimated) • In office • Full-Time • 6+ years exp • Munich
AI/ML
AI Agents
Apply
$81k – $105k per year • In office • Full-Time • 7+ years exp • Munich
TypeScript
JavaScript
Databases
PostgreSQL
AI/ML
LLM
Frontend
tRPC
DevOps
CI/CD
Docker
Kubernetes
Apply
Full Stack Engineer 2 days ago
$70k – $93k per year • In office • Full-Time • 3+ years exp • Munich
Python
TypeScript
JavaScript
Databases
PostgreSQL
AI/ML
LLM
Frontend
tRPC
DevOps
CI/CD
Docker
Kubernetes
Apply
$222k – $277k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco
DevOps
CI/CD
Immutable Infrastructure
Apply
$70k – $196k per year • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
Databases
Databricks
Google BigQuery
SAP HANA
Snowflake
AI/ML
Knowledge Graph
DevOps
Azure
Apply
$70k – $196k per year • Remote/Hybrid • Full-Time • 5+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
DevOps
SLI/SLO/SLA
Apply
$70k – $206k per year • In office • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
AI/ML
AI Agents
Apply
$293k – $385k per year • In office • Full-Time • San Francisco
AI/ML
OpenAI
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.