370,952open jobs
9,558companies
49,284added this week
Browse all
Salary
$180k – $240k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
Your career needs an agent.. Pluto is an AI voice agent that learns your story in 10 minutes and makes you discoverable to the right people and AI agents.

Location: San Francisco, CA

Work Model: Onsite, 5 days per week. Relocation support provided.

Industry: AI infrastructure / model inference

Compensation: $180,000 - $240,000 base, plus equity

About the Company

Our partner is a venture-backed AI infrastructure company delivering the fastest inference available on open models, serving both enterprise and serverless customers. They run their own compute, and customer demand currently outpaces the capacity they can bring online. The team is small, flat, and deliberately staying that way as they scale.

The Opportunity

This is a forward-deployed engineering role that owns the customer, not just the code. Sales takes the first meeting. From there you are both the customer's engineer and their point of contact: you decide what to prove, you build it, you keep it running in production, and you carry the relationship.

Because you sit closer to real production load than anyone else on the team, what you learn goes straight into the roadmap. You will report directly into engineering leadership, and the longer-term plan is for engineers in this function to own their own pods as the team grows. These are among the first hires into the function, so its shape is still yours to influence.

Responsibilities

  • Own assigned customer accounts end to end once the first sales meeting is complete
  • Win the technical evaluation by proving performance on the customer's own workload rather than a synthetic benchmark, and explain the result to their engineering team
  • Stand up dedicated deployments and optimize the full stack for each customer's workload
  • Own production for your accounts: investigate latency and error-rate regressions, resolve or route them, and communicate directly with the customer
  • Carry on-call responsibility for customer issues on your accounts
  • Feed real-world performance findings back to the engineering team to shape the product roadmap

Requirements

  • Strong ML systems and inference knowledge, including profiling and benchmarking models, working with traces, and building to technical specifications
  • Demonstrated ability to own a customer relationship end to end, with the communication skills to work directly with technical stakeholders
  • Comfort operating from incomplete specifications and driving work to completion independently
  • Experience in a startup or other fast-moving engineering environment
  • Based in, or willing to relocate to, San Francisco for a fully onsite role
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
370,952 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$220k – $500k per year • In office • Full-Time • New York
AI/ML
AI Agents
Apply
$70k – $100k per year • Remote • Internship
TypeScript
JavaScript
Frontend
Next.js
React.js
Apply
$150k – $220k per year • In office • Full-Time • 3+ years exp • New York
TypeScript
JavaScript
Databases
PostgreSQL
AI/ML
AI Agents
LLM
Frontend
Next.js
React.js
Apply
$40k – $70k per year • Remote • Internship
JavaScript
Python
TypeScript
Apply
$150k – $220k per year • In office • Full-Time • 3+ years exp • New York
TypeScript
JavaScript
Databases
PostgreSQL
AI/ML
AI Agents
LLM
Frontend
Next.js
React.js
Apply
$140k – $200k per year • Equity 0.5–1.5% • In office • Full-Time • 3+ years exp • San Francisco
SQL
TypeScript
Node JS
JavaScript
Databases
Neo4j
PostgreSQL
Supabase
AI/ML
AI Agents
Model Context Protocol
Frontend
Next.js
React.js
DevOps
Git
Vercel
Management
Slack
Apply
$148k – $282k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • San Francisco • San Jose
AI/ML
LLM
Prompt Engineering
Apply
$150k – $250k per year • Equity 0.3–1.5% • In office • Full-Time • San Francisco
Python
TypeScript
AI/ML
AI Agents
LLM
NLP
Transformers
DevOps
AWS
GitHub
Terraform
Management
Notion
Apply
$155k – $338k per year (Estimated) • In office • Full-Time • San Francisco
Python
Rust
TypeScript
Node JS
JavaScript
Node JS
Electron
Databases
PostgreSQL
Supabase
AI/ML
AI Agents
LLM
Frontend
Next.js
React.js
DevOps
AWS
GitHub
Vercel
Management
Linear
Apply
$106k – $200k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Milwaukee • San Francisco • Cincinnati • Kansas City • Chicago
Apply
See all jobs
This is one of many
370,952 more open roles from verified company boards, updated every day.