435,278open jobs
15,177companies
61,432added this week
Browse all
Salary
$140k – $165k per year
Location
Remote/Hybrid (Palo Alto, United States)
Seniority
Middle · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Instrumental builds camera stations and software that detect defects on electronics assembly lines. Its machine learning models spot anomalies engineers did not anticipate. The company works with consumer electronics manufacturers in Asia.

Instrumental builds the manufacturing acceleration platform behind the world’s most complex electronics. We capture digital exhaust and engineering context from assembly lines-images, test logs, BOM data, performance, repair cycles-and our AI engines identify insights that are difficult or impossible for human engineers to find. We accelerate the companies building the AI era by improving manufacturing yield, throughput, and ramp. NVIDIA, Meta, Cisco, and their manufacturing partners rely on Instrumental to accelerate new product introduction and production.

The Instrumental platform collects, intelligently transforms, and contextually presents manufacturing data to technical end-users, enabling them to optimize their manufacturing process in real-time. Our core technology is proprietary ML algorithms, packaged in an accessible, user-centric user interface-we believe we must have both the best technology and the best access to that technology to win.

As a Site Reliability Engineer, you’ll operate, improve, and scale our AWS-based SaaS platform. You’ll combine hands-on production operations with engineering, focusing on reliability, automation, observability, and operational excellence. You’ll participate in a bi-weekly on-call rotation, but the goal isn’t simply to keep systems running-it’s to continuously engineer away the operational complexity that comes with scaling our platform and customer base.

Requirements:

  • 3-4 years of experience in Site Reliability Engineering, DevOps, Cloud Operations, Platform Engineering, or Systems Engineering supporting production SaaS environments.
  • Strong hands-on experience with AWS, including EC2, VPC, IAM, RDS, ECS, and S3.
  • Experience managing infrastructure using Terraform or other Infrastructure as Code technologies.
  • Experience designing and supporting CI/CD pipelines using GitHub Actions, Jenkins, GitLab CI/CD, or similar platforms.
  • Strong experience with monitoring and observability tools, preferably Datadog, including dashboards, alerting, logging, and APM.
  • Experience with Docker and Kubernetes.
  • Scripting experience with Python and/or Bash.
  • Experience supporting production environments through an on-call rotation, including incident response and root cause analysis.
  • Proven ability to take ownership of production issues and drive them through investigation, remediation, and long-term resolution.

Who You Are:

  • Dead serious about performance, scalability, and reliability (PSR): You care deeply about how systems behave in the real world and continuously look for ways to make them more reliable, scalable, observable, and supportable.
  • Automation, automation, automation: If something is repetitive, manual, or error-prone, your first instinct is to automate it and make it disappear.
  • An engineer at heart: You don’t want to repeatedly fight the same fires. You look for the underlying cause and build durable engineering solutions that reduce operational toil and technical debt.
  • Strong systems thinker: You understand how infrastructure, applications, networks, deployments, monitoring, and people interact-and can troubleshoot complex production issues across those boundaries.
  • Collaborative and reliable: You partner closely with software engineers to make services production-ready, improve operational workflows, and build reliability into systems before they become problems.
  • Comfortable with growth and ambiguity: You’re comfortable making good decisions without perfect information and adapting as the platform, customer base, and company scale quickly.

Nice to Have:

  • Experience working in a high-growth B2B SaaS environment.
  • Experience implementing SRE practices such as SLIs, SLOs, and error budgets.
  • Experience building internal tooling and automation to eliminate operational toil.
  • Experience supporting multi-region AWS environments.
  • AWS cost optimization or FinOps experience.
  • Network, application security, and compliance experience.
  • Experience introducing AI tools or processes into engineering and operational workflows.

This position requires access to items and data that are developed under U.S. government contracts and subject to dissemination controls that limit access to U.S. citizens only.

We’re a growing team that works collaboratively, is supportive of each other, and is highly energized by the opportunity for a large impact. We actively work to promote an inclusive environment, valuing passion and the ability to learn. You’re encouraged to apply even if your experience doesn’t precisely match the job description!

The following is a representative annual base salary range for this position within the Bay Area: $140,000-$165,000. Job level and salary opportunities are evaluated through our interview process - we review the experience, knowledge, skills, and abilities of each applicant.

Instrumental is proud to offer a highly-rated variety of benefits, including health, vision, dental, commuter plans, and parental leave.

At Instrumental, protecting company and customer information is a shared responsibility. Employees are expected to comply with company engineering, security, access control, and privacy policies, and promptly report suspected security incidents or policy violations.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
435,278 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Palo Alto
Remote/Hybrid • 5+ years exp
Python
JavaScript
TypeScript
Node JS
Databases
PostgreSQL
pgvector
AI/ML
Claude
Claude Code
Embeddings
Prompt Engineering
Function Calling
LLM
RAG
Semantic Search
Anthropic
Structured Outputs
Semantic Search
Tool Use
Frontend
React.js
DevOps
Azure
Apply
In office • 5+ years exp
Python
AI/ML
Copilot
Claude
ChatGPT
Prompt Engineering
AI Agents
Midjourney
Instructor
TensorFlow
PyTorch
Perplexity
DevOps
GitHub
Apply
Remote/Hybrid • 4+ years exp
Python
JavaScript
C#
C#
.NET
AI/ML
LangChain
Prompt Engineering
AI Agents
LLM
OpenAI
Hugging Face
Frontend
React.js
Apply
Senior AI Engineer 1 day ago
Remote/Hybrid • 9+ years exp
Python
JavaScript
C#
C#
.NET
AI/ML
LangChain
Prompt Engineering
AI Agents
LLM
OpenAI
Hugging Face
Frontend
React.js
Apply
Remote/Hybrid
Python
C#
C#
.NET
Cybersecurity
Threat Modeling
Apply
$175k – $205k per year • Remote/Hybrid • Full-Time • 5+ years exp • Palo Alto
DevOps
AWS
Kubernetes
Apply
$116k – $176k per year • In office • Full-Time • Cincinnati
Apply
$140k – $200k per year • In office • Full-Time • Palo Alto
Apply
Lead Recruiter 22 days ago
$170k – $221k per year • Remote/Hybrid • Full-Time • Palo Alto
Apply
$204k – $240k per year • Remote/Hybrid • Full-Time • 10+ years exp • Master's Degree • Palo Alto
Design
SolidWorks
Apply
$213k – $285k per year • In office • 8+ years exp • Master's Degree • Palo Alto
Python
SQL
AI/ML
Claude
Multimodal AI
AI Agents
LLM
Multi-Agent Systems
QA
Rest-Assured
Apply
$218k – $365k per year • In office • Full-Time • 12+ years exp • Master's Degree • San Francisco • Palo Alto • Chicago • New York • Washington
AI/ML
AI Agents
LLM
Agentforce
Apply
Workplace Manager 1 day ago
$68k – $146k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Palo Alto
Apply
Chief of Staff 1 day ago
$120k – $150k per year • Equity 0.4–0.7% • In office • Full-Time • 3+ years exp • Palo Alto
AI/ML
AI Agents
Management
Stripe
Apply
$48k – $96k per year • In office • Internship • Palo Alto
Apply
See all jobs
This is one of many
435,278 more open roles from verified company boards, updated every day.