368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$109k – $245k per year (Estimated)
Location
Remote (United States)
Employment
Full-Time
Overview
Company
Impact
Profile match
PostHog is a product analytics company headquartered in San Francisco, California, and founded in 2020 by James Hawkins and Tim Glaser. The company builds an open source platform combining product analytics, session replay, feature flags, A/B testing, surveys, and error tracking that teams can self-host or run on its cloud. It went through Y Combinator, operates as a fully remote organisation, and publishes its internal handbook and pricing openly.

About PostHog

Product development used to mean manually writing code, running analysis, diagnosing bugs, and rolling out changes using dozens of tools.

PostHog is the only platform that acts like a co-pilot for you (and your AI agents) to do it all - autonomously.

We started with open-source product analytics, launched out of Y Combinator's W20 cohort. We've since shipped more than a dozen products, including:

  • PostHog Code, the only AI devtool that understands your product, not just your codebase.

  • A built-in data warehouse, so users can query product and customer data together using custom SQL insights.

  • PostHog AI, an AI-powered analyst that answers product questions, helps users find useful session recordings, and writes custom SQL queries.

We are:

  • Product-led. More than 450,000 organizations have installed PostHog, mostly driven by word-of-mouth. We have intensely strong product-market fit.

  • Default alive. Revenue is growing incredibly quickly, and we're very efficient. We raise money to push ambition and grow faster, not to keep the lights on.

  • Well-funded. We've raised more than $180m from some of the world's top investors. We're set up for a long, ambitious journey.

We're focused on building an awesome product for end users, hiring exceptional teammates, shipping fast, and being as weird as possible.

Things we care about

  • Transparency: Everyone can read about our roadmap, how we pay (or even let go of) people, our strategy, and how we work, in our public company handbook. Internally, we share revenue, notes and slides from board meetings, and fundraising plans, so everyone has the context they need to make good decisions.

  • Autonomy: We don’t tell anyone what to do. Everyone chooses what to work on next based on what's going to have the biggest impact on our customers, and what they find interesting and motivating to work on. Engineers lead product teams and make product decisions. Teams are flexible and easy to change when needed.

  • Shipping fast: Why not now? We want to build a lot of products; we can't do that shipping at a normal pace. We've built the company around small teams - autonomous, highly-efficient groups of cracked engineers who can outship much larger companies because they own their products end-to-end.

  • Time for building: Nothing gets shipped in a meeting. We're a natively remote company. We default to async communication - PRs > Issues > Slack. Tuesdays and Thursdays are meeting-free days, and we prioritize heads down building time over perfect coordination. This will be the most productive job you've ever had.

  • Ambition: We want to solve big problems. We strongly believe that aiming for the best possible upside, and sometimes missing, is better than never trying. We're optimistic about what's possible and our ability to get there.

  • Being weird: Weird means redesigning an already world-class website for the 5th time. It means shipping literally every product that relates to customer data. It means building an objectively unnecessary developer toy with dubious shareholder value. Doing weird stuff is a competitive advantage. And it's fun.

Who we're looking for

We’re looking for people (US - Central/Eastern Timezone) that like deep ownership of production systems, people that are not afraid of working with stateful infrastructure and love working in AWS, VMs, automation, and making messy systems reliable.

In general we seek SRE’s who are:

  • Enthusiastic drivers. We need proactive people that can fully own projects and get them done, and know to get help when needed. "Are we there yet?" is the wrong question.

  • Optimistic problem solvers. Things get hard here sometimes, whether it's scaling, shipping complex products, handling a stream of support requests, or trying to ship something that touches multiple teams. We need people who won't get disheartened, and will collaborate, iterate, and ship their way out of anything.

  • Grown ups. We’re an international bunch of weirdos, but one thing unites us: everyone is kind, considerate, and professional towards each other.This isn't about age or experience, it's about being low-ego, flexible, and respectful.

  • Genuine builders. PostHog is full of people who just love building stuff, people who would still be building software even if there wasn't a paycheck at the end. If this sounds like you, we should talk.

What you'll be doing

You won’t be in a typical “keep the lights on” SRE role. The work is about turning a fast-growing, stateful system into a predictable, well-automated platform. (provisioning, scaling, rebalancing, recovery)

That means reducing operational stress, designing safe automation for traffic-heavy workloads, and building the tooling and patterns that let the system scale without scaling human effort.

You'll work on the kind of problems that only show up at large scale (petabytes of data, thousands of cores, constant ingestion) across a multi-region, multi-account AWS platform running many services on Kubernetes.

  • Operating EKS clusters across several environments with Karpenter autoscaling, Cilium networking, and ArgoCD-driven GitOps deployments

  • Managing and evolving a multi AWS account organization, provisioning, networking, access control, and cross-account connectivity

  • Maintaining the Terraform/Terragrunt IaC platform - modules, automated plan-on-PR / apply-on-merge pipelines, and safe patterns for shared infrastructure

  • Improving operational tooling around deploys, schema changes, backups, restores, and incident response

  • Reducing operational load by identifying repeat pain points and eliminating them through code and self-healing automation

  • Optimizing cloud spend as you go

  • Participating in on-call and incident response, with a strong focus on making incidents rarer over time

You'll have room to design and automate, not just respond to alerts. You should join this team if you like deep ownership of production systems and enjoy building the platform layer that everything else runs on.

Requirements

  • Deep hands-on experience with Kubernetes in production (EKS preferred). You've debugged node pressure, networking issues, and deployment failures at scale (thousands of nodes)

  • Strong experience operating production infrastructure on AWS. Not just one account, but understanding organizational boundaries, IAM, and networking between many

  • Experience automating infrastructure using Terraform or Terragrunt at scale, including module design and state management

  • Solid understanding of Linux systems (disk, memory, networking, failure modes)

  • Experience supporting stateful systems (databases, queues, storage systems, etc.)

  • Ability to debug and reason about performance and reliability issues in production

  • You're comfortable owning systems end-to-end, including on-call responsibilities

You don't need to be an expert in every system we run on day one. But you do need to enjoy owning complex infrastructure and learning how the pieces fit together.

Nice to have

  • Experience with GitOps workflows (ArgoCD) and CI/CD pipelines (GitHub Actions)

  • Experience with building AI agent-enabled base-level infra services for teams that move fast

  • Familiarity with multi-region infrastructure and the consistency/availability tradeoffs that come with it

We are committed to ensuring a fair and accessible interview process. If you need any accommodations or adjustments, please let us know.

#LI-DNI

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$25k – $60k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Chennai
Java
PowerShell
Python
TypeScript
JavaScript
Java
Spring Boot
Databases
Amazon Redshift
DynamoDB
Frontend
Angular
DevOps
Amazon EC2
Amazon EKS
Ansible
ArgoCD
AWS
AWS Lambda
Azure
Azure DevOps
Bicep
Chef
CI/CD
CloudFormation
Configuration Management
Docker
GCP
GitLab CI
Jenkins
Kubernetes
Platform Engineering
Puppet
Terraform
Amazon S3
GitLab
IAM
Apply
$14k – $35k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Phoenix
SQL
AI/ML
AI Agents
Edge AI
DevOps
AWS
Apply
$13k – $29k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Pune
JavaScript
Apex
Apex
MuleSoft
AI/ML
AI Agents
Edge AI
DevOps
AWS
Azure
Management
Draw.io
Marketing
Salesforce
Apply
$180k – $225k per year • Equity • In office • Full-Time • 4+ years exp • Bachelor's Degree • Seattle
C#
Go
Java
Kotlin
TypeScript
JavaScript
AI/ML
Claude
Claude Code
OpenAI Codex
Frontend
React.js
DevOps
AWS
AWS Step Functions
Apply
Sr. UX Designer 3 hours ago
$76k – $161k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Toronto
Python
AI/ML
AI Agents
Hallucination
Human-in-the-Loop
LLM
LLM Guardrails
Design
Figma
Apply
AI Research Engineer 13 days ago
$100k – $210k per year (Estimated) • Remote/Hybrid • Full-Time • PhD
Rust
SQL
AI/ML
AI Agents
CUDA
CUDA Toolkit
Jupyter Notebook
PyTorch
Transformers
Management
Slack
Apply
$84k – $207k per year (Estimated) • Remote • Full-Time
Go
Node JS
Rust
SQL
JavaScript
Databases
Apache Kafka
ClickHouse
PostgreSQL
Redis
AI/ML
AI Agents
DevOps
Amazon S3
Management
Slack
Apply
Remote • Full-Time
SQL
AI/ML
AI Agents
Analytics
A/B Testing
Management
Slack
Apply
$102k – $265k per year (Estimated) • In office • Full-Time • San Francisco
SQL
AI/ML
AI Agents
DevOps
Amazon EKS
ArgoCD
AWS
CI/CD
Cilium
GitHub Actions
GitOps
Karpenter
Kubernetes
Self-Healing
Terraform
Terragrunt
GitHub
IAM
Management
Slack
Apply
$123k – $283k per year (Estimated) • Remote • Full-Time
SQL
AI/ML
AI Agents
Analytics
A/B Testing
Management
Slack
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.