368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$95k – $224k per year (Estimated)
Location
Remote (United Kingdom, Ireland, Portugal)
Seniority
Staff
Overview
Company
Impact
Profile match
Gradle by Develocity is the open source build system for Java, Android, and Kotlin developers.

Who We Are

AI is changing how software gets built. Code production is becoming a commodity. The focus is shifting from writing code to orchestrating, verifying, and governing change - and the toolchain is the new constraint.

We are at the center of this shift. We build Develocity, a toolchain observability and intelligence platform used by some of the world's leading software organizations - Netflix, Airbnb, Spotify, SAP, major global banks, and hundreds more. Develocity helps software teams achieve delivery excellence through deep observability, build and test acceleration, and AI-powered intelligence across the entire toolchain - with current support for Gradle Build Tool, Apache Maven™, sbt, npm, and Python.

We are an AI-native company. AI is not a feature we're bolting on - it's central to how we work, how we think about our product, and where we're heading. We're investing deeply in making Develocity's unique data and decades of domain expertise accessible to both humans and AI agents, with trust, evidence, and explainability at the core of everything we build.

We have partnered with the Apache Software Foundation, the Commonhaus Foundation, the Micronaut Foundation, and other OSS projects such as Spring, Quarkus, Kotlin, JUnit, AndroidX, and many more to bring the values of Develocity also to the OSS Community.

Our Values

Seek to Understand: Everything starts with listening and understanding; we strive to understand diverse viewpoints, problems, and motivations. Before we take action, we ensure we truly grasp the challenges, perspectives, and goals. 

Know the Why: We approach our work with a clear sense of purpose, ensuring every step is deliberate and focused. We take meaningful action with urgency, but never at the expense of thoughtful consideration. 

Innovate & Iterate: We embrace challenges and are not afraid to try new things, even if they might fail. With a deep understanding and a clear purpose, we can develop creative, bold solutions to tackle challenges.

Own the Outcome: We are empowered to take initiative, and we maintain transparency in our work and its outcomes. When we execute, we take responsibility for our decisions, measure the success of our innovations, and learn from the results.

Who You Are

We're building a new SRE team and looking for founding members to help shape how we operate. As a Lead SRE, you’ll be a technical and operational leader for reliability across Develocity. You’ll help define our SRE vision, set standards for how we operate production services, and mentor other SREs as the team grows. This is a hands-on role with broad influence across engineering, cloud platform, and customer-facing teams.

The SRE team will be responsible for the reliability, performance, and availability of Develocity instances serving paying customers, open-source projects, and public-facing services, plus supporting infrastructure like artifact registries.

You'll work on our internally-built Cloud Application Platform, Kubernetes on AWS, and develop deep expertise in it. When incidents happen, you'll troubleshoot issues across the stack, from application to infrastructure. You'll collaborate with the Cloud Platform team to improve the tooling you depend on, and with engineering teams to build reliability into how we ship software. If you like automating things and hate doing the same task twice, you'll fit in well.

You'll be part of a distributed, remote-first team that values asynchronous communication and written documentation. Strong self-direction and clear communication across time zones are essential.

Responsibilities

  • Operate and maintain all Develocity instances and supporting services in production.
  • Define and evolve SRE standards, practices, and operating models, including on-call, incident response, postmortems, and SLOs.
  • Participate in an on-call rotation, acting as a technical escalation point for complex or high-severity incidents.
  • Lead incident response and blameless retrospectives, ensuring learnings result in measurable reliability improvements.
  • Set reliability priorities using risk, customer impact, business goals, SLOs, and error budgets.
  • Identify systemic reliability risks and continuously evolve Develocity’s SaaS operations as the platform and customer base grow.
  • Lead and influence architectural and design reviews to ensure reliability, scalability, and operability.
  • Drive automation across deployment, upgrades, monitoring, self-healing, recovery, and operational workflows.
  • Build and maintain comprehensive observability for all managed services, including logging, metrics, tracing, and alerting.
  • Own disaster recovery, backups, and business continuity planning and execution.
  • Partner with engineering leadership to balance feature delivery with reliability and operational excellence.
  • Mentor and coach SREs, supporting technical growth and strong operational practices.
  • Help onboard new SREs and contribute to hiring by defining and assessing SRE excellence at Develocity.
  • Communicate clearly with customers during incidents and maintenance windows.
  • Optimize performance, resource utilization, and operational costs.

Minimum qualifications

  • 7+ years in SRE, DevOps, or an equivalent role operating production services at scale.
  • Experience leading reliability initiatives across multiple teams or services.
  • Demonstrated ability to influence technical direction without direct authority.
  • Experience designing and operating systems with SLOs and error budgets, and exercising strong judgment in balancing reliability, velocity, and cost.
  • Strong Kubernetes experience in production environments.
  • Cloud infrastructure expertise, preferably AWS (EKS, RDS, S3, EC2).
  • Proficiency with observability tools (Prometheus, Grafana) and Infrastructure as Code (Terraform).
  • Track record of incident management and response in a 24/7 on-call environment.
  • Scripting proficiency (Python, Bash) for automation.
  • Strong written and verbal English communication skills.

Preferred qualifications

  • Experience as a founding or early SRE establishing practices in a growing SaaS organization.
  • Familiarity with Develocity.
  • JVM language experience (Java, Kotlin).
  • Experience with customer-facing and executive-level incident communications.

What We Offer

  • A ground-floor role in a new SRE team - you'll shape how we do things, not inherit someone else's decisions.
  • Real ownership of production systems used by engineers at companies you've heard of.
  • Direct interaction with customers when things go wrong (and when they go right).
  • A culture that values automation over heroics.
  • In-person meetings, such as our annual company offsite and team meetings.
  • Work from home in a remote-first environment.
  • Competitive salaries and equity grants.

Location

  • Remote and must be located in GMT timezone (UK, Ireland, Portugal, Canary Islands).
  • While our team works remotely and is spread across the globe, we deeply value daily interactions and collaboration.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$42k – $104k per year (Estimated) • In office • Internship • 5+ years exp
Bash
PowerShell
Python
DevOps
Amazon EC2
Amazon EKS
Ansible
AWS
AWS Lambda
Azure
CentOS Stream
Chef
CI/CD
CloudFormation
Configuration Management
Datadog
FinOps
GCP
Git
GitHub Actions
GitLab CI
Grafana
Hyper-V
Istio
Jenkins
Kubernetes
KVM
Linkerd
Platform Engineering
Prometheus
Puppet
Service Mesh
Splunk
Terraform
Ubuntu
VMWare
Windows Server
Amazon CloudWatch
Amazon ECS
Amazon S3
API Gateway
AWS Step Functions
GitHub
GitLab
IAM
Cybersecurity
GDPR
ISO 27001
SOC 2
Apply
$12k – $35k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru
Databases
MySQL
Oracle
PostgreSQL
AI/ML
AI Agents
DevOps
AIOps
Amazon EC2
AWS
CI/CD
CloudFormation
GitHub Actions
Kubernetes
Terraform
Amazon CloudWatch
Amazon S3
GitHub
IAM
Apply
$24k – $64k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Chennai
Python
Scala
SQL
TypeScript
JavaScript
Java
Python
pySpark
Java
Spring Boot
Databases
Apache Kafka
Databricks
AI/ML
AI Agents
Copilot
Google ADK
LLM
NLP
Prompt Engineering
Spark
Devin
Model Context Protocol
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
OpenShift
GitHub
Analytics
ETL/ELT
Apply
Backend Engineer 1 day ago
$45k – $57k per year • In office • Full-Time • 3+ years exp • Tokyo
Python
TypeScript
JavaScript
Python
FastAPI
Databases
PostgreSQL
Frontend
Next.js
React.js
DevOps
Amazon EC2
AWS
AWS CDK
CI/CD
Docker
GitHub Actions
Vercel
Amazon CloudWatch
Amazon S3
GitHub
IAM
Management
Linear
Apply
$23k – $32k per year (Estimated) • Remote/Hybrid • Full-Time • Bucharest
Java
SQL
Kotlin
Java
Gradle
Spring Boot
Kotlin
Mockito
Databases
Apache Kafka
MySQL
PostgreSQL
RabbitMQ
Redis
AI/ML
Claude
Copilot
Cursor
Prompt Engineering
Mobile
JUnit
DevOps
AWS
CI/CD
Docker
Git
Kubernetes
GitHub
Apply
$114k – $237k per year (Estimated) • Equity • Remote
Java
Kotlin
Python
JavaScript
Java
Gradle
Maven
Micronaut
Quarkus
AI/ML
AI Agents
Context Engineering
Frontend
npm
Mobile
JUnit
DevOps
AWS
Docker
Kubernetes
Apply
$180k – $205k per year • Equity • Remote
Java
Kotlin
Python
JavaScript
Java
Gradle
Maven
Micronaut
Quarkus
AI/ML
AI Agents
Frontend
npm
Mobile
JUnit
DevOps
Amazon EC2
Amazon EKS
AWS
Grafana
Incident Management
Kubernetes
Prometheus
Self-Healing
Terraform
Amazon S3
Apply
$68k – $172k per year (Estimated) • Equity • Remote
Java
Kotlin
Python
JavaScript
Java
Gradle
Maven
Micronaut
Quarkus
AI/ML
AI Agents
Frontend
npm
Mobile
JUnit
DevOps
Amazon EC2
Amazon EKS
AWS
Grafana
Incident Management
Kubernetes
Prometheus
Self-Healing
Terraform
Amazon S3
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.