368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$103k – $200k per year (Estimated)
Location
In office (Chicago)
Seniority
Senior · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Ripple is a financial technology company that provides enterprise-grade blockchain infrastructure for real-time cross-border payments, liquidity, and asset tokenization. Its platform allows banks, payment providers, and institutions to execute instant, low-cost international transactions using its global network and digital assets like XRP as bridge liquidity. Beyond its core payment rails, the company offers digital asset custody solutions, stablecoin infrastructure, and central bank digital currency (CBDC) platform tools.

At Ripple, we’re building a world where value moves like information does today. It’s big, it’s bold, and we’re already doing it. Through our crypto solutions for financial institutions, businesses, governments and developers, we are improving the global financial system and creating greater economic fairness and opportunity for more people, in more places around the world. And we get to do the best work of our career and grow our skills surrounded by colleagues who have our backs. 

If you’re ready to see your impact and unlock incredible career growth opportunities, join us, and build real world value.

At Ripple, we’re building a world where value moves like information does today. Through our crypto solutions for financial institutions, businesses, governments, and developers, we are improving the global financial system and creating greater economic fairness and opportunity for more people, in more places around the world.

Ripple Treasury, now a Ripple solution acquired in 2025, marks a significant expansion into the multi-trillion-dollar corporate finance arena. With more than 40 years of experience supporting some of the world’s largest and most sophisticated companies, Ripple Treasury integrates a treasury command center into Ripple’s technology stack-giving corporates the ability to move, manage, and optimize liquidity in real-time, across traditional and digital assets, under one expanded umbrella.

THE WORK:

This is an engineering-first role with a coaching dimension-not the other way around. You will spend the majority of your time doing hands-on observability and reliability engineering work: building instrumentation, designing alert configurations, authoring Terraform, and troubleshooting production systems. Alongside that, you will coach and consult with stream-aligned product teams, helping them build operational maturity over time.

You will join Ripple’s Technical Operations team and work across Azure (80%) and AWS (20%) environments supporting infrastructure that is predominantly Windows-based (80%), handling significant payment volume for enterprise treasury customers. The incident management program you will help build is early-stage-you will be establishing practices, not inheriting a mature playbook.

WHAT YOU’LL DO:

  • Observability Engineering

    • Design and implement monitoring, alerting, and dashboards in New Relic (APM, Infrastructure, Logs, Synthetics) across Azure and AWS; write NRQL queries for troubleshooting, analysis, and reporting.
    • Define and implement SLOs/SLIs and error budgets; coach teams on using them to balance feature velocity with reliability and communicate system health to stakeholders.
    • Lead alert noise reduction and signal quality engineering-tune thresholds, eliminate false positives, and ensure every alert is actionable.
    • Optimize observability costs through log ingestion management, pipeline rules, and New Relic configuration governance.
    • Partner with engineering teams to improve observability maturity: structured logging, metrics instrumentation (RED/USE methods), distributed tracing, and effective dashboard patterns.

    Infrastructure & IaC

    • Develop and maintain Terraform infrastructure as code for provisioning and managing monitoring resources, alert configurations, and observability infrastructure-this is a primary engineering responsibility, not an occasional task.
    • Establish and enforce IaC governance standards for observability infrastructure across teams, providing a repeatable, auditable model for how monitoring resources are managed.
    • Author and troubleshoot Azure DevOps pipelines; support teams with deployment visibility, change tracking, and release hygiene as it relates to production reliability.

    Incident Management

    • Administer and configure Incident.IO: alert routing, notification workflows, Slack and OpsGenie integration, and runbook management-operationalizing what exists today and expanding from there.
    • Build out incident management foundations that are largely yours to establish: PIR/postmortem processes, on-call rotation design, escalation policies, incident severity classification, and response playbooks.
    • Track and report on MTTR, MTTD, and incident frequency; identify trends and drive continuous improvement in partnership with engineering teams.
    • Respond to and debrief on production incidents-providing real-time troubleshooting support and facilitating structured post-incident reviews.

    Cross-Functional Enablement

    • Enable stream-aligned engineering teams to adopt improved observability and incident management practices through workshops, consultation, and hands-on guidance.
    • Collaborate with the Subsystems Platform Team to translate common needs into self-service observability and incident management capabilities.
    • Build lasting team competency through documentation, training materials, and knowledge-sharing sessions that outlast any individual engagement.

WHAT YOU'LL BRING: 

  • Core SRE Experience

    • 7+ years in Site Reliability Engineering, DevOps, or Platform Engineering with a strong focus on observability and production operations.
    • Proven ability to deliver hands-on engineering work while coaching and mentoring teams-comfortable switching between builder and consultant modes.
    • Experience working in Agile/Scrum environments and collaborating effectively with cross-functional teams.

    Observability & Incident Management Expertise - Required

    • Expert-level hands-on experience with New Relic (APM, Infrastructure, Logs, Synthetics, Alerts) and strong NRQL proficiency for troubleshooting and analysis.
    • Deep understanding of structured logging, metrics collection (RED/USE methods), distributed tracing, and designing effective dashboards and alerts.
    • Expertise defining and implementing SLOs/SLIs and error budgets for reliability management.
    • Hands-on experience with incident management platforms (Incident.IO, PagerDuty, OpsGenie, or similar).
    • Experience designing incident response workflows, on-call rotations, escalation policies, and facilitating post-incident reviews that drive actionable improvements.
    • Demonstrated ability to troubleshoot complex production issues using observability data across distributed systems.

    Infrastructure & Tools - Required

    • Strong Terraform experience: developing and maintaining IaC for cloud infrastructure and monitoring resources; familiarity with IaC governance patterns.
    • Proficiency with PowerShell scripting (required given the 80% Windows environment).
    • Strong experience with Azure cloud (App Services, Virtual Machines, Azure SQL, networking, monitoring) and working knowledge of AWS.
    • Experience with Azure DevOps for CI/CD pipeline authoring and troubleshooting.
    • Experience with Octopus Deploy for deployment management and release orchestration.
    • Comfort working across both Windows and Linux server environments.
    • Familiarity with Slack for operational workflows, alert routing, and incident communication.

    Desired / Additional

    • Experience with alert noise reduction strategies and observability cost optimization (log ingestion, pipeline rules, cardinality management).
    • Background facilitating chaos engineering, game day exercises, or failure injection to build team resilience.
    • Knowledge of VM-hosted SQL Server monitoring and performance optimization.
    • Familiarity with FinTech compliance requirements (SOC 2, ISO 27001) and audit evidence collection.
    • Experience measuring and improving key reliability metrics (MTTR, MTTD, availability, error budgets) at an organizational level.
    • Python or Bash scripting experience in addition to PowerShell.
    • Familiarity with Jira for incident tracking and workflow automation.

    Other common names for this role: Senior Site Reliability Engineer, Observability Engineer, Incident Management Engineer

For positions that will be based in IL, the annual salary range for this position is below. Actual salaries may vary based on numerous factors including, among other things, an individual applicant’s experience and qualifications for the position. This range does not include equity or additional compensation, such as bonuses or commissions. 

IL Annual Base Salary Range

$160,000—$200,000 USD

WHO WE ARE:

Do Your Best Work

  • The opportunity to build in a fast-paced start-up environment with experienced industry leaders
  • A learning environment where you can dive deep into the latest technologies and make an impact.  A professional development budget to support other modes of learning.
  • Thrive in an environment where no matter what race, ethnicity, gender, origin, or culture they identify with, every employee is a respected, valued, and empowered part of the team.
  • In-office collaboration for moments that matter is important to our culture, and we give managers and teams the flexibility to decide which 10+ days a month they come in. 
  • Bi-weekly all-company meeting - business updates and ask me anything style discussion with our Leadership Team
  • We come together for moments that matter which include team offsites, team bonding activities, happy hours and more!

Take Control of Your Finances

  • Competitive salary, bonuses, and equity
  • Competitive benefits that cover physical and mental healthcare, retirement, family forming, and family support
  • Employee giving match
  • Mobile phone stipend

Take Care of Yourself

  • R&R days so you can rest and recharge
  • Generous wellness reimbursement and weekly onsite & virtual programming
  • Generous vacation policy - work with your manager to take time off when you need it
  • Industry-leading parental leave policies. Family planning benefits.
  • Catered lunches, fully-stocked kitchens with premium snacks/beverages, and plenty of fun events

Benefits listed above are for full-time employees. 

Ripple is an Equal Opportunity Employer. We’re committed to building a diverse and inclusive team. We do not discriminate against qualified employees or applicants because of race, color, religion, gender identity, sex, sexual identity, pregnancy, national origin, ancestry, citizenship, age, marital status, physical disability, mental disability, medical condition, military status, or any other characteristic protected by local law or ordinance.

Please find our UK/EU Applicant Privacy Notice and our California Applicant Privacy Notice for reference.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Chicago
$25k – $42k per year • Equity 0–0.2% • Remote • Full-Time • 3+ years exp
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
$100k – $210k per year • Equity 0–0.5% • Remote • Full-Time • 3+ years exp • San Francisco
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
$230k – $260k per year • Equity • Remote • Internship • Bachelor's Degree
Python
DevOps
Amazon EKS
AWS
Azure
CI/CD
GCP
Helm
Kubernetes
Terraform
Cybersecurity
FedRAMP
Orca Security
Apply
$22k – $52k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
DevOps
AWS
Azure
GCP
Analytics
Power BI
Tableau
ETL/ELT
Apply
Team Lead DevOps 1 day ago
$23k – $62k per year (Estimated) • Remote • 5+ years exp • Moscow
Bash
Python
Erlang
Erlang
EMQX
Databases
Apache Kafka
ClickHouse
PostgreSQL
RabbitMQ
Redis
Redpanda
Trino
DevOps
Ansible
AWS
AWX
FinOps
HAProxy
Hetzner
Kubernetes
SLI/SLO/SLA
Terraform
Yandex Cloud
Amazon S3
Apply
In office • Contractor • Dublin
DevOps
Azure
Cybersecurity
Okta
Management
Google Workspace
Jira
Slack
Apply
$75k – $181k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Toronto
Java
Python
DevOps
AWS
CI/CD
CloudFormation
GitOps
Kubernetes
OpenTelemetry
Prometheus
Rancher
Terraform
Apply
$132k – $235k per year (Estimated) • In office • Full-Time • 3+ years exp • Chicago
Analytics
A/B Testing
Apply
$150k – $305k per year (Estimated) • In office • Full-Time • 8+ years exp • New York
JavaScript
AI/ML
LangChain
LLM
Anthropic
OpenAI
Frontend
Next.js
React.js
DevOps
CI/CD
Cloudflare
Vercel
Analytics
A/B Testing
Marketing
GA4
Marketo
Salesforce
Apply
$161k – $291k per year (Estimated) • In office • Internship • 5+ years exp • San Francisco
C#
Go
Java
Kotlin
Python
Rust
Java
Spring Boot
Databases
Apache Kafka
DevOps
AWS
gRPC
Kubernetes
Amazon ECS
Apply
$118k – $162k per year • Remote • Full-Time • 10+ years exp • Bachelor's Degree • Louisville • Fort Lauderdale • Washington • Chicago • Tampa
DevOps
Azure
GCP
Cybersecurity
HIPAA
Apply
$158k – $288k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Chicago
Python
SQL
Python
pySpark
Databases
Databricks
Snowflake
AI/ML
Spark
DevOps
AWS
Apply
$120k – $160k per year • Equity • Remote/Hybrid • 5+ years exp • Chicago
TypeScript
JavaScript
AI/ML
ChatGPT
Claude
Claude Code
Copilot
Edge AI
Frontend
Angular
Webpack
DevOps
CI/CD
Datadog
GitHub
GitLab
GitLab CI
QA
Cypress
Jest
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • Chicago
AI/ML
AI Agents
Apply
$97k – $184k per year • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Chicago
Cybersecurity
CAPA
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.