408,465open jobs
14,167companies
73,523added this week
Browse all
Salary
$217k – $304k per year
Location
Remote (United States)
Seniority
Staff · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Jobgether is a Belgian recruitment platform built entirely around remote and flexible work, aggregating openings from thousands of employers that allow work from outside an office. Its matching engine ranks roles against a candidate's skills, seniority and stated preferences on location and flexibility, rather than leaving people to filter a keyword search, and it verifies how genuinely remote each posting is. The company also runs an AI screening layer that shortlists applicants for employers, and publishes research and guidance on distributed work practices alongside the job marketplace itself.

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Site Reliability Engineer, Ads based in United States.

As a Staff Site Reliability Engineer, you will provide technical leadership for reliability across a large-scale advertising technology ecosystem.

You will help ensure critical systems remain highly available, scalable, performant, and resilient as traffic and business demands grow.

Your scope will span ad serving, auctions, targeting, reporting, measurement, and billing, with direct impact on revenue and customer experience.

You will partner with engineering leaders and multiple technical teams to establish reliability roadmaps, architecture standards, and operational practices.

The role combines hands-on engineering, automation, incident leadership, observability, and long-term resilience initiatives.

You will also mentor engineers and influence technical decisions across a broad organization.

This is an opportunity to shape reliability strategy for high-QPS, low-latency systems where operational excellence directly drives business outcomes.

Accountabilities

    • Lead reliability initiatives across multiple advertising technology domains, including ad serving, auctions, targeting, reporting, measurement, and billing.
    • Partner with engineering leadership to define and execute a roadmap focused on reliability, scalability, operational excellence, and developer productivity.
    • Design and build platforms, tooling, automation, and engineering solutions that improve system resilience and developer efficiency at scale.
    • Lead architecture reviews and influence technical decisions affecting critical, revenue-generating distributed systems.
    • Participate in on-call rotations, lead complex production investigations, and coordinate cross-functional responses to major incidents.
    • Identify systemic reliability risks and drive durable engineering solutions that strengthen platform resilience.
    • Establish and monitor reliability metrics for critical customer and advertiser journeys, including campaign creation, ad delivery, auction participation, reporting, attribution, and billing.
    • Improve operational maturity through SLOs, automation, incident management, observability, and performance optimization.
    • Troubleshoot complex issues across modern distributed system stacks and drive root-cause analysis toward long-term improvements.
    • Mentor engineers and provide technical leadership across multiple teams and initiatives.
    • Influence roadmap and investment decisions by ensuring reliability requirements are incorporated into product and infrastructure planning.
    • Requirements

      • 8+ years of experience in Site Reliability Engineering, Infrastructure Engineering, or a related discipline, with experience operating large-scale distributed systems.
      • Strong experience evolving and supporting high-traffic, user-facing production environments.
      • Deep expertise in distributed systems, scale engineering, cloud-native architectures, and highly available system design.
      • Strong software engineering capabilities in a general-purpose backend language such as Go.
      • Solid understanding of observability practices and technologies, including metrics, logging, tracing, and alerting.
      • Proven experience improving reliability through SLOs, automation, incident management, performance optimization, and operational best practices.
      • Demonstrated ability to diagnose and resolve complex issues across modern distributed technology stacks.
      • Strong cross-functional leadership skills, with the ability to influence technical direction and drive operational improvements across multiple teams.
      • Excellent written and verbal communication skills and the ability to collaborate effectively with engineering leadership and technical stakeholders.
      • Experience with Kubernetes, cloud infrastructure, and large-scale distributed systems is highly valued.
      • Experience with advertising technology or other revenue-critical platforms is a plus, particularly ad serving, real-time auctions, budget pacing, campaign delivery, measurement, attribution, or billing systems.
      • Experience operating high-QPS, low-latency services where performance directly affects business outcomes is advantageous.
      • Familiarity with large-scale data technologies such as Kafka, ClickHouse, Spark, Flink, or BigQuery is a plus.
      • Experience partnering with Product, Data Science, and advertising engineering teams is beneficial.
      • Exposure to machine learning inference or recommendation systems operating at scale is also valued.
      • Benefits

        • Comprehensive health benefits, including medical, dental, and vision coverage.
        • 401(k) program with employer matching.
        • Equity compensation in the form of restricted stock units, subject to the position offered.
        • Flexible vacation policy and global company days off.
        • 4+ months of paid parental leave.
        • Family planning support.
        • Workspace benefits and support for a home office.
        • Personal and professional development funds.
        • Paid volunteer time off.
        • Flexible-first workforce approach with remote-friendly work arrangements.
        • Reasonable accommodations available for qualified candidates with disabilities and disabled veterans.
        • Base salary range of $217,000-$303,900 USD, with final compensation determined by factors such as skills, depth of experience, relevant credentials, level, and location.
        • Certain positions may also include additional variable compensation depending on the role.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
408,465 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$51k per year • Remote • Full-Time • 5+ years exp
Go
DevOps
AWS
Azure
Docker
GCP
Kubernetes
Web3
Avalanche
Ethereum
Polygon
Reown
Smart Contracts
Apply
$51k per year • Remote • Full-Time
Go
Kotlin
Rust
Solidity
TypeScript
JavaScript
Databases
Azure Cosmos DB
PostgreSQL
Frontend
Chakra UI
Next.js
React.js
DevOps
Argo Workflows
ArgoCD
Azure
CI/CD
GitHub
GitHub Actions
Grafana
Kubernetes
OpenTelemetry
Web3
Cosmos
Ethereum
Apply
$51k per year • Remote • Full-Time
C#
C++
Go
Java
Kotlin
Rust
Solidity
TypeScript
JavaScript
Java
Spring Boot
Kotlin
Ktor
Databases
Amazon Aurora
Frontend
ESLint
MSW
React.js
Rollup
Storybook
Vite
Vue.js
DevOps
Amazon Kinesis
AWS
AWS Lambda
Azure
CI/CD
CloudFormation
GCP
Terraform
Web3
DeFi
Layer 2
Rollup
Smart Contracts
Zk-SNARKs
Zk-STARKs
QA
Vitest
Apply
$51k per year • Remote • Full-Time
Go
TypeScript
JavaScript
Databases
Amazon Aurora
Frontend
ESLint
MSW
Storybook
Vite
DevOps
Amazon ECS
Amazon Kinesis
AWS
AWS CDK
Azure
CI/CD
GCP
Grafana
Platform Engineering
Terraform
Web3
Smart Contracts
Management
Slack
QA
Sentry
Vitest
Apply
$51k per year • Remote • Full-Time
Go
TypeScript
JavaScript
Databases
Amazon Aurora
Frontend
Chakra UI
ESLint
GraphQL
MSW
React.js
Storybook
Svelte
tRPC
Vite
Vue.js
DevOps
Amazon Kinesis
AWS
AWS Lambda
CI/CD
Git
Rest API
Web3
Smart Contracts
QA
Vitest
Apply
Remote • Full-Time • 10+ years exp
Apply
Remote • Full-Time • 10+ years exp
Apply
$28k – $71k per year (Estimated) • Remote • Contractor • 5+ years exp • Bachelor's Degree
Python
AI/ML
Amazon SageMaker
Evidently AI
MLFlow
PyTorch
Recommender Systems
TensorFlow
DevOps
AWS
CI/CD
GitLab
GitLab CI
Grafana
Prometheus
Apply
$50k – $127k per year (Estimated) • Remote • Contractor • 5+ years exp • Bachelor's Degree
Python
AI/ML
Amazon SageMaker
Evidently AI
MLFlow
PyTorch
Recommender Systems
TensorFlow
DevOps
AWS
CI/CD
GitLab
GitLab CI
Grafana
Prometheus
Apply
Remote/Hybrid • Internship
Apply
See all jobs
This is one of many
408,465 more open roles from verified company boards, updated every day.