368,657open jobs
9,442companies
50,883added this week
Browse all
Salary
$53k – $74k per year
Location
Remote (India)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
SigNoz is an observability company headquartered in San Francisco, California, and founded in 2020. The company develops an open source, OpenTelemetry-native platform that unifies traces, metrics, and logs in a single application, offered both self-hosted and as a managed cloud service. It targets engineering teams looking for an alternative to per-seat observability vendors, and went through Y Combinator before raising venture funding.

About SigNoz

SigNoz is an open-source observability platform that helps modern engineering teams monitor, debug, and optimize their applications with deep visibility into metrics, traces, and logs - all in one place. We're built natively on OpenTelemetry and offer both self-hosted and cloud options, so teams can run observability the way they want, without vendor lock-in.

We are growing fast and building core developer infra products. And we are not fooling around:

  • 27,000+ GitHub stars

  • 800+ customers

  • 7,000+ members in our Slack community

Role: Sr Site Reliability Engineer (SRE)

We're looking for an SRE to own the reliability, scalability, and operability of the SigNoz cloud platform. You'll keep a petabyte-scale observability system fast and dependable - making sure the people who trust us to watch their systems can always trust ours. The platform team handles infra, scalability of SaaS, ingest pipelines, staging environments, automation, and the operational backbone of the product.

This is a deeply hands-on role for someone who understands what actually breaks in production at scale - and enjoys fixing it for good.

What we're looking for

  • Kubernetes at scale - not just "I've deployed to k8s," but real fluency with the nuances and gotchas: resource tuning, autoscaling behavior, networking, stateful workloads, upgrades, and the failure modes that only show up under load

  • Working knowledge of ClickHouse - operating it, tuning queries, and understanding its behavior at scale - is a strong plus

  • Knowledge of Golang is a plus (most of our stack and tooling is in Go)

  • Familiarity with OpenTelemetry and running large-scale data ingest pipelines is a plus

What you'll work on

You'll work with a high-caliber team across areas like:

  • Reliability of the SigNoz cloud platform: SLOs/SLIs, error budgets, incident response, and on-call practices that don't burn people out

  • Scaling the ingest path - making it robust to bursts while maintaining data freshness

  • SaaS auto-scalability and capacity planning across a petabyte-scale system

  • Operating and tuning ClickHouse and the data layer for performance and cost

  • Kubernetes infrastructure: cluster operations, upgrades, multi-tenancy, and the automation that keeps it boring

  • Observability of SigNoz itself - we dogfood our own product, so you'll help make it world-class

  • Infrastructure-as-code, CI/CD, and the tooling that lets a small team operate big systems

What will make you successful

  • 5-8 years in SRE, infrastructure, or platform/backend roles operating production systems at scale

  • Deep, practical Kubernetes experience - you know where the bodies are buried

  • Strong grasp of distributed systems failure modes, performance debugging, and capacity planning

  • Comfortable in code (Go preferred) - you automate and fix things, not just configure them

  • Loves open source - ideally with prior contributions to OSS projects (any size)

  • Comfortable in a high-ownership, fast-moving, remote-first environment

  • Strong communication - can write clear runbooks and tech docs and explain trade-offs

Nice-to-haves

  • Past experience on platform/infra/SRE teams of Series B+ startups

  • Hands-on experience operating ClickHouse, Kafka, or similar high-throughput data systems

  • Experience in observability (monitoring / logging / tracing) and with OpenTelemetry

Why you'll love working at SigNoz

  • Work on a globally used open-source project that engineers actually love

  • Huge scope and ownership - your work directly shapes how teams adopt SigNoz

  • Collaborate with a high-caliber team who just can't stop shipping

  • Remote-first, async-friendly culture

  • Opportunity to help define the future of open-source observability

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,657 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$80k – $169k per year (Estimated) • Equity • Remote • Full-Time • 5+ years exp
Go
Python
Databases
PostgreSQL
RabbitMQ
DevOps
Alertmanager
Ansible
Atlantis
Backstage
Chef
CI/CD
containerd
Docker
GCP
GitOps
Google GKE
Grafana
Helm
Incident Management
Kubernetes
Loki
Platform Engineering
Prometheus
Puppet
Terraform
Thanos
IAM
Cybersecurity
Checkov
SOC 2
Least Privilege
Apply
$112k – $217k per year (Estimated) • Equity • Remote • Full-Time • 5+ years exp
Go
Python
Databases
PostgreSQL
RabbitMQ
DevOps
Alertmanager
Ansible
Atlantis
Backstage
Chef
CI/CD
containerd
Docker
GCP
GitOps
Google GKE
Grafana
Helm
Incident Management
Kubernetes
Loki
Platform Engineering
Prometheus
Puppet
Terraform
Thanos
IAM
Cybersecurity
Checkov
SOC 2
Least Privilege
Apply
$42k – $106k per year (Estimated) • Equity • Remote • Full-Time • 5+ years exp
Go
Python
Databases
PostgreSQL
RabbitMQ
DevOps
Alertmanager
Ansible
Atlantis
Backstage
Chef
CI/CD
containerd
Docker
GCP
GitOps
Google GKE
Grafana
Helm
Incident Management
Kubernetes
Loki
Platform Engineering
Prometheus
Puppet
Terraform
Thanos
IAM
Cybersecurity
Checkov
SOC 2
Least Privilege
Apply
$41k – $103k per year (Estimated) • Remote • Full-Time • 5+ years exp
Go
Python
AI/ML
Reinforcement Learning
Edge AI
DevOps
Ansible
AWS
Azure
Chef
CI/CD
Docker
GCP
GitHub Actions
Google GKE
Jenkins
Kubernetes
Platform Engineering
Puppet
Terraform
GitHub
Apply
$20k – $49k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
Ruby
SQL
Databases
Amazon Neptune
Neo4j
AI/ML
Hallucination
LangChain
LangGraph
LLM
Model Context Protocol
Spark
AI Agents
LLM Guardrails
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
GitHub Actions
GitLab CI
Jenkins
Kubernetes
Rest API
Terraform
GitHub
GitLab
QA
Playwright
Postman
Selenium
Swagger
Apply
$21k – $69k per year (Estimated) • Remote • Full-Time
TypeScript
JavaScript
Frontend
React.js
DevOps
SigNoz
GitHub
Management
Slack
Apply
AI GTM Engineer 2 months ago
$26k – $72k per year (Estimated) • Remote • Full-Time • 4+ years exp
SQL
Management
n8n
Zapier
Marketing
HubSpot
Mixpanel
Apply
Sr Product Designer 3 months ago
$31k – $70k per year (Estimated) • Remote • Full-Time • 4+ years exp
DevOps
OpenTelemetry
SigNoz
GitHub
Management
Slack
Apply
$30k – $75k per year (Estimated) • Remote • Full-Time • 5+ years exp
Go
Python
Python
Asyncio
FastAPI
Pydantic
Databases
Apache Kafka
ClickHouse
AI/ML
Function Calling
LLM
Prompt Engineering
Model Context Protocol
DevOps
Kubernetes
OpenTelemetry
SigNoz
GitHub
Management
Slack
Apply
Exceptional Engineer 3 months ago
$19k – $87k per year (Estimated) • Remote • Full-Time
Databases
ClickHouse
DevOps
Kubernetes
OpenTelemetry
SigNoz
GitHub
Management
Slack
Apply
See all jobs
This is one of many
368,657 more open roles from verified company boards, updated every day.