368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$139k – $262k per year (Estimated)
Location
Remote (United States)
Seniority
Staff · 12+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Filevine is a leading provider of legal case management software. We are committed to helping law firms streamline their workflows and improve their efficiency. Filevine is a global organization with a mission to help law firms streamline their workflows and improve their efficiency.

Role Summary

As a Staff Site Reliability Engineer at Filevine, you are the senior technical authority on the SRE team

and a strategic partner to engineering leadership. You don’t just maintain systems - you shape

engineering culture, define the technical standard for how Filevine runs in production, and bridge the

gap between high-level business goals and robust, internet-scale technical execution. You bring a

forward-looking perspective - actively shaping how AI and machine learning drive the future of

reliability practice.

You own the roadmap across two critical SRE domains - Observability & Alerting and Platform

Infrastructure - and are accountable for ensuring the team solves reliability problems permanently

rather than absorbing them as toil. You operate as the senior IC counterpart to the Engineering

Manager: technical correctness lives with you. You partner with the Reliability Architect and engineering

leadership on significant technical decisions, mentor engineers across experience levels, and influence

reliability strategy across the broader organization. Reliability at Filevine protects revenue. You are the

senior technical voice responsible for ensuring that uptime, incident response, and every production

change meet the operational standard the business demands.

This role does not participate in on-call rotation, but you are deeply invested in the engineers who do -

shaping the on-call strategy, tooling, and culture that make production support sustainable and

effective.

Who You Are

The Technical Authority

  • Master of the Craft: You bring deep expertise in distributed systems, cloud infrastructure,

observability, and reliability engineering. You raise the technical standard for every engineer

around you and thrive where the challenges are complex and the stakes are real.

  • Technical Leader and Mentor: You are passionate about mentoring engineers and investing in

their growth. You influence technical direction and communicate production risk clearly across

engineering, product, and executive audiences.

  • Forward-Thinking & AI/ML Fluent: You bring deep knowledge of AIOps and drive the use of

AI and machine learning in observability, anomaly detection, incident response, automated

remediation, and resource optimization.

  • Production-Scale Problem Solver: You turn ambiguous, complex reliability challenges into

durable solutions for systems where availability, performance, and production changes carry

meaningful business impact.

  • Software-Minded Builder: You use software, automation, Infrastructure as Code, and platform

capabilities to eliminate toil and make systems safer, more scalable, and easier to operate.

What you will do

Define and execute the technical strategy for Observability & Alerting, Platform Infrastructure,

and operational excellence.

  • Lead the evolution of reliable, scalable, secure, and efficient cloud platforms and distributed

systems.

  • Champion SLIs, SLOs, error budgets, capacity planning, operational readiness, and automation

across the service lifecycle.

  • Lead the organization through complex production incidents and turn post-incident learning into

permanent engineering improvements.

  • Build self-service platform capabilities that reduce toil, improve engineering safety and velocity,

and make every team more capable of owning their own reliability.

  • Mentor engineers and serve as a trusted technical authority for long-term reliability and platform

direction.

Qualifications

12+ years of experience in software engineering, infrastructure, platform engineering, or SRE,

including 6+ years in SRE and 3+ years leading complex, cross-functional technical initiatives

for distributed production systems.

  • Expert-level depth in observability and platform infrastructure, with broad expertise in incident

response, capacity planning, automation, and reliability engineering.

  • Advanced experience with a major container-orchestration platform, preferably Kubernetes, and

an observability platform such as New Relic, Datadog, or equivalent.

  • Strong software-engineering ability in Python, Go, Bash, or another general-purpose language,

with experience building production tooling, automation, or platform capabilities.

  • Proven ability to mentor engineers and communicate technical risk clearly to engineering,

product, and executive audiences.

  • Experience in a regulated environment such as FedRAMP, CJIS, HIPAA, SOC 2, or PCI is

strongly preferred.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
United States
$23k – $57k per year (Estimated) • Remote • Full-Time • 6+ years exp
Python
Databases
OpenSearch
DevOps
AIOps
ArgoCD
AWS
Azure
CI/CD
Datadog
Docker
Dynatrace
GCP
GitHub Actions
Grafana
Incident Management
Jaeger
Jenkins
Kubernetes
New Relic
OpenTelemetry
Platform Engineering
Prometheus
SLI/SLO/SLA
Splunk
Terraform
GitHub
Apply
$18k – $46k per year (Estimated) • Remote • Full-Time • Tula
C#
JavaScript
Node JS
SQL
TypeScript
C#
.NET
Node JS
InversifyJS
Databases
DynamoDB
MySQL
AI/ML
Claude
Copilot
Cursor
OpenAI Codex
Frontend
Angular
React.js
Tailwind CSS
Mobile
Dependency Injection
DevOps
AWS
AWS Lambda
CI/CD
OpenTelemetry
Rest API
Terraform
Amazon CloudWatch
Amazon S3
API Gateway
GitHub
Cybersecurity
HIPAA
Apply
$19k – $53k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Pune
Bash
JavaScript
Python
TypeScript
Frontend
Angular
React.js
DevOps
AWS
Azure
Datadog
Docker
GCP
Grafana
Kubernetes
Prometheus
Splunk
IAM
Cybersecurity
Keycloak
Apply
$162k – $180k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree
Python
SQL
Databases
Snowflake
AI/ML
dbt
DevOps
AWS
Cybersecurity
HIPAA
Zero Trust
Analytics
ETL/ELT
Apply
$37k – $80k per year (Estimated) • Remote/Hybrid • Full-Time • 7+ years exp • Pune
PowerShell
Python
AI/ML
Anomaly Detection
DevOps
AppDynamics
AWS
Azure
CI/CD
GCP
Grafana
Incident Management
Prometheus
Splunk
Apply
$99k – $201k per year (Estimated) • In office • Full-Time • 6+ years exp • Bachelor's Degree • Salt Lake City
SQL
Databases
Snowflake
Analytics
Power BI
Tableau
Marketing
HubSpot
Salesforce
Apply
Remote/Hybrid • Contractor • 3+ years exp • Prague
Python
TypeScript
JavaScript
Python
Django
Frontend
React.js
Svelte
DevOps
AWS
Kubernetes
Apply
Remote/Hybrid • Contractor • 3+ years exp • Bratislava
Python
TypeScript
JavaScript
Python
Django
Frontend
React.js
Svelte
DevOps
AWS
Kubernetes
Apply
$76k – $164k per year (Estimated) • In office • Full-Time • 3+ years exp • Salt Lake City
Management
Slack
Apply
$109k – $212k per year (Estimated) • Remote • Full-Time • 8+ years exp • United States
Python
DevOps
AWS
CI/CD
Kubernetes
Apply
$78k – $130k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Cary • San Jose
Analytics
Power BI
Apply
$91k – $182k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • United States
Design
SolidWorks
Apply
$45k – $60k per year • Remote • Full-Time • 2+ years exp • PhD • United States
Cybersecurity
HIPAA
Apply
$117k – $258k per year (Estimated) • Equity • Remote • Full-Time • United States
C++
Java
Python
Cybersecurity
Crowdstrike
Apply
$90k – $184k per year (Estimated) • Equity • Remote • Full-Time • 3+ years exp • United States
AI/ML
Red Teaming
Cybersecurity
Burp Suite
Cobalt Strike
Crowdstrike
Metasploit
MITRE ATT&CK
Nessus
Nmap
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.