368,657open jobs
9,442companies
50,883added this week
Browse all
Salary
$145k – $295k per year (Estimated)
Location
Remote (United States, Canada)
Seniority
Staff · 8+ years exp
Employment
Internship
Overview
Company
Impact
Profile match
Reddit is an American social platform founded in 2005 and organised into hundreds of thousands of user-created communities known as subreddits. Members post links, text, images and video, and the community ranks each submission through upvotes and downvotes, which decides what surfaces on the front page and inside each community. Headquartered in San Francisco and listed on the New York Stock Exchange since 2024, the company earns most of its revenue from advertising and has added data licensing deals that let AI developers train on its public conversation archive.

Reddit is a community of communities. It’s built on shared interests, passion, and trust, and is home to the most open and authentic conversations on the internet. Every day, Reddit users submit, vote, and comment on the topics they care most about. With 100,000+ active communities and approximately 130 million daily active unique visitors, Reddit is one of the internet’s largest sources of information. For more information, visit www.redditinc.com.

This role is remote friendly. Reddit has a flexible first workforce

The Ads organization powers Reddit's advertising platform, enabling advertisers to reach highly engaged communities while helping Reddit grow its business. The reliability of our Ads systems directly impacts advertiser success, revenue generation, and user experience.

The Ads Reliability team partners closely with Ads Engineering teams to improve reliability, scalability, operational excellence, and developer productivity across Reddit's advertising ecosystem.

We're looking for a Staff Site Reliability Engineer who will define and provide technical leadership for reliability initiatives across the Ads organization and help shape the future of Ads infrastructure at Reddit.

What you’ll do:

  • Lead reliability initiatives across multiple Ads domains including ad serving, auctions, targeting, reporting, measurement, and billing.
  • Partner with engineering leadership to develop a roadmap to improve reliability, scalability, operational excellence, and engineering efficiency across the Ads organization.
  • Design and build platforms, tooling, and automation that improve reliability and developer productivity at scale.
  • Drive architecture reviews and influence technical decisions impacting critical revenue-generating systems.
  • Participate in on-call rotations, lead complex incident investigations and coordinate cross-functional response efforts during major production events.
  • Identify systemic reliability risks and drive long-term solutions that improve platform resilience.
  • Establish reliability metrics around advertiser-critical user journeys such as campaign creation, ad delivery, auction participation, reporting, attribution, and billing.
  • Mentor engineers and provide technical leadership across multiple teams.
  • Influence roadmap planning and ensure reliability considerations are incorporated into product and infrastructure investments.

What We’re Looking For

  • 8+ years of experience in Site Reliability Engineering, Infrastructure Engineering, or related roles operating large scale distributed systems.
  • Strong experience evolving high traffic, user-facing production environments.
  • Strong cross-functional collaborations skills to lead and influence projects driving operational excellence.
  • Deep expertise in modern distributed systems, scale engineering, and cloud-native architectures.
  • Experience designing highly-available systems with strong operational and reliability practices.
  • Strong software engineering skills in languages like general-purpose backend languages like Go.
  • Strong understanding of observability systems including metrics, logging, tracing, and alerting.
  • Experience improving reliability through SLOs, automation, incident management, and performance optimization.
  • Demonstrated ability to troubleshoot complex issues across a modern distributed system stack.
  • Strong collaboration and communication skills with the ability to influence technical direction across teams.

Nice to Have

  • Experience supporting advertising technology platforms or other large-scale revenue-critical systems.
  • Deep understanding of reliability challenges associated with ad-serving, real-time auctions, budget pacing, campaign delivery, measurement, attribution, or billing systems.
  • Experience operating high-QPS, low-latency services where latency directly impacts business outcomes.
  • Experience establishing reliability programs that deliver meaningful, measurable business outcomes
  • Experience with Kubernetes, cloud infrastructure, and large-scale distributed systems.
  • Familiarity with Kafka, ClickHouse, Spark, Flink, BigQuery, or similar large-scale data platforms.
  • Experience partnering with Product, Data Science, and Ads Engineering organizations.
  • Experience supporting machine learning inference or recommendation systems at scale.

Benefits:

  • Comprehensive Health benefits
  • 401k Matching 
  • Workspace benefits for your home office
  • Personal & Professional development funds
  • Family Planning Support
  • Flexible Vacation & Reddit Global Days Off
  • 4+ months paid Parental Leave  
  • Paid Volunteer time off

Pay Transparency:

This job posting may span more than one career level.

In addition to base salary, this job is eligible to receive equity in the form of restricted stock units, and depending on the position offered, it may also be eligible to receive a commission. Additionally, Reddit offers a wide range of benefits to U.S.-based employees, including medical, dental, and vision insurance, 401(k) program with employer match, generous time off for vacation, and parental leave. To learn more, please visit https://www.redditinc.com/careers/.

To provide greater transparency to candidates, we share base salary ranges for all US-based job postings regardless of state. We set standard base pay ranges for all roles based on function, level, and country location, benchmarked against similar stage growth companies. Final offer amounts are determined by multiple factors including, skills, depth of work experience and relevant licenses/credentials, and may vary from the amounts listed below.

The base salary range for this position is:

$217,000—$303,900 USD

In select roles and locations, the interviews will be recorded, transcribed and summarized by artificial intelligence (AI). You will have the opportunity to opt out of recording, transcription and summarization prior to any scheduled interviews.

During the interview, we will collect the following categories of personal information: Identifiers, Professional and Employment-Related Information, Sensory Information (audio/video recording), and any other categories of personal information you choose to share with us. We will use this information to evaluate your application for employment or an independent contractor role, as applicable.  We will not sell your personal information or disclose it to any third party for their marketing purposes.  We will delete any recording of your interview promptly after making a hiring decision.  For more information about how we will handle your personal information, including our retention of it, please refer to our Candidate Privacy Policy for Potential Employees and Contractors.

Reddit is proud to be an equal opportunity employer, and is committed to building a workforce representative of the diverse communities we serve.  Reddit is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If, due to a disability, you need an accommodation during the interview process, please let your recruiter know.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,657 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
In office • Full-Time • Noida
C#
SQL
TypeScript
JavaScript
C#
.NET
Databases
Oracle
Frontend
Angular
DevOps
Incident Management
SLI/SLO/SLA
Apply
$71k – $149k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Jacksonville • King of Prussia
Go
Python
Java
Java
Flyway
Liquibase
DevOps
Amazon EKS
ArgoCD
CI/CD
Git
GitHub Actions
GitLab CI
Helm
Jenkins
JFrog Artifactory
Kubernetes
Platform Engineering
SRE
Terraform
AWS
GitHub
GitLab
Cybersecurity
SonarQube
Apply
$17k – $21k per year (Estimated) • Remote/Hybrid • Full-Time • Bucharest
DevOps
Azure
GCP
Incident Management
Splunk
Cybersecurity
Google SecOps
MITRE ATT&CK
Apply
$39k – $83k per year (Estimated) • In office • Full-Time • 12+ years exp • Pune
C++
Go
Java
Java
Spring Boot
Databases
Apache Kafka
NATS
DevOps
CI/CD
Docker
Jenkins
Kubernetes
Rest API
Apply
$96k – $227k per year (Estimated) • In office • Full-Time • Melbourne • Sydney
PowerShell
Python
DevOps
AWS
Azure
CI/CD
GCP
Incident Management
Kubernetes
Platform Engineering
Service Mesh
Terraform
Apply
$163k – $292k per year (Estimated) • Equity • Remote • Internship • 5+ years exp • Master's Degree
Java
Python
Databases
Apache Kafka
Google BigQuery
Redis
AI/ML
Spark
Apply
$127k – $240k per year (Estimated) • Equity • Remote • Internship • 6+ years exp • Master's Degree
Apply
$141k – $249k per year (Estimated) • Equity • Remote • Internship • 4+ years exp • Master's Degree
Python
SQL
AI/ML
Anomaly Detection
Apply
$187k – $310k per year (Estimated) • Equity • Remote • Internship • 5+ years exp
Apply
$177k – $299k per year (Estimated) • Equity • Remote • Contractor • 10+ years exp • Bachelor's Degree
Python
SQL
Apply
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 1 hour ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
See all jobs
This is one of many
368,657 more open roles from verified company boards, updated every day.