412,234open jobs
14,465companies
72,289added this week
Browse all
Salary
$150k – $185k per year
Location
Remote (United States)
Seniority
Senior · 5+ years exp
Overview
Company
Impact
Profile match
Headquartered in Silver Spring, Maryland, Rocket Money is a personal finance technology company and subsidiary of Rocket Companies. The company provides an all-in-one mobile and web platform offering subscription management, automated bill negotiation, credit monitoring, and personal budgeting tools. By consolidating financial account data and automating recurring cost-reduction workflows, it enables consumers to optimize spending and maintain control over their personal finances.

ABOUT ROCKET MONEY

Rocket Money’s mission is to empower people to live their best financial lives. Rocket Money offers members a unique understanding of their finances and a suite of valuable services that save them time and money - ultimately giving them a leg up on their financial journey.

ABOUT THE TEAM

We're looking to expand our Cloud Infrastructure team with a Senior Infrastructure Engineer, SRE to lead the reliability and operational evolution of our platform. We run hundreds of services in production, which enable us to process billions of transactions, consume multiple terabytes of data, and produce hundreds of millions of logs per day, and our reliability practice needs to evolve to match our growing scale. This includes:

  • Building and improving the reliability and resiliency of our systems and services
  • Establishing SLIs, SLOs, and error budgets for our most critical services and user journeys, and reviewing them regularly with the teams that own them
  • Owning and evolving our disaster recovery strategy: recovery objectives, failover and restore paths, and regular exercises that prove they work
  • Partnering with product engineering teams so they can own and operate their own services, with metrics that reflect real user experience
  • Evolving our observability platform and standards across metrics, tracing, and logs: including instrumentation paved roads, alert quality, and observability cost
  • Strengthening our incident practice: tuning paging thresholds, keeping runbooks current, and following through on postmortem action items
  • Contributing to day-to-day Cloud Infrastructure work alongside your reliability specialty - infrastructure build-outs, platform backlog, and a shared on-call rotation (1 week out of every 6 weeks)

You'll join the Cloud Infrastructure team and partner with engineering and internal support teams to drive this work.

We support millions of people to improve their financial lives, and this role ensures we can continue to do so reliably and at scale.

ABOUT YOU

  • You have 5+ years of hands-on cloud or infrastructure engineering experience, with substantial time spent on reliability and production operations at scale
  • You have defined SLIs and SLOs for real production services, and can talk about what changed as a result. What got fixed, what got deprioritized, and what you got wrong the first time
  • You have hands-on experience with an observability platform in production; Datadog strongly preferred
  • You're comfortable writing code (Python, Go, TypeScript, or similar) for internal tooling, production debugging, and automation
  • You write production Terraform and are comfortable in AWS, and when production breaks you can find the problem and fix it
  • You have built or operated a disaster recovery plan: you set the recovery goals, wrote the failover and restore steps, and ran the drills that proved it works
  • You have been on-call for services you helped build, and you have opinions about what makes an alert worth waking someone for
  • You prefer giving teams paved roads and good defaults over mandates, so they can own their own instrumentation

BONUS POINTS 

  • You have led a reliability or observability modernization project where you defined the vision, approach, and delivered the implementation
  • You have built internal tooling, libraries, or instrumentation standards that made it easier for other teams to operate their services well
  • You have run game days, chaos experiments, or DR exercises, and fixed the problems they uncovered
  • You have cut observability spend while keeping the coverage you needed

WE OFFER

  • Health, Dental & Vision Plans
  • Competitive Pay
  • 401k Matching
  • Unlimited PTO
  • Lunch daily (in-office only)
  • Snacks & Coffee (in-office only)
  • Commuter benefits (in-office only)

Additional information: Salary range of $150,000 - $185,000/year + bonus + benefits. Base pay offered may vary depending on job-related knowledge, skills, and experience.

This job description is an outline of the primary responsibilities of this position and may be modified at the discretion of the company at any time.  Decisions related to employment are not based on race, color, religion, national origin, sex, physical or mental disability, sexual orientation, gender identity or expression, age, military or veteran status or any other characteristic protected by state or federal law.  The company provides reasonable accommodations to qualified individuals with disabilities in accordance with applicable state and federal laws.

The information regarding compensation and other benefits included in this paragraph is the company’s current, good faith estimate at the time of posting. [Compensation and benefits are subject to modification from time to time as the Company, in its sole and exclusive discretion, deems appropriate.] The Company may determine during its future reviews of the proposed compensation and benefits provided for this position, that the compensation and benefits for such position should be reduced. In no event will the Company reduce the compensation for the position to a level below the applicable jurisdictional minimum wage rate for the position. Los Angeles County and San Francisco Candidates only: qualified applicants with arrest or conviction records will be considered for employment per the Fair Chance Ordinance and the Fair Chance Initiative for Hiring.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
412,234 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$26k – $66k per year (Estimated) • Remote/Hybrid • 11+ years exp • Noida
Python
JavaScript
Java
TypeScript
SQL
Node JS
Java
Hibernate
Spring MVC
Databases
PostgreSQL
Frontend
React.js
JQuery
DevOps
Rest API
CI/CD
Jenkins
AWS
Docker
Kubernetes
Platform Engineering
AWS Lambda
Amazon S3
IAM
Amazon CloudWatch
Apply
In office • 3+ years exp
Python
SQL
Databases
Snowflake
Google BigQuery
Amazon Redshift
Microsoft Fabric
BigQuery
AI/ML
Spark
dbt
DevOps
GCP
AWS
Analytics
Power BI
ETL/ELT
Apply
DevOps Engineer II 1 day ago
In office • 5+ years exp • Bachelor's Degree
Python
Go
JavaScript
Ruby
DevOps
CI/CD
AWS
Docker
Kubernetes
Platform Engineering
Configuration Management
Apply
$56k – $143k per year (Estimated) • Remote/Hybrid • Sheffield
DevOps
AWS
Cybersecurity
Nessus
Qualys Cloud Platform
Microsoft Defender
Microsoft Defender for Cloud
Apply
In office • 5+ years exp
DevOps
GCP
Azure
AWS
FinOps
Analytics
Power BI
Apply
$130k – $170k per year • In office • Full-Time • 3+ years exp • Los Angeles
JavaScript
TypeScript
Databases
PostgreSQL
Frontend
GraphQL
React.js
Mobile
React Native
Apply
$130k – $170k per year • Remote • 3+ years exp • San Francisco
JavaScript
TypeScript
Databases
PostgreSQL
Frontend
GraphQL
React.js
Mobile
React Native
Apply
$115k – $140k per year • In office • 4+ years exp • New York
Apply
$40k – $44k per year • Remote
Apply
$150k – $185k per year • Remote • 8+ years exp • San Francisco
JavaScript
TypeScript
Node JS
Databases
PostgreSQL
Frontend
GraphQL
React.js
Management
Stripe
Apply
$85k – $105k per year • Equity 0–0.1% • In office • Full-Time • San Francisco
Management
Slack
Marketing
HubSpot
LinkedIn
Apply
$260k – $310k per year • Equity 0.1–0.4% • In office • Full-Time • 3+ years exp • San Francisco
Management
Slack
Marketing
HubSpot
Apply
$65k – $100k per year • Remote • Contractor • San Francisco
AI/ML
Claude
Apply
$71k – $95k per year • In office • 2+ years exp • Bachelor's Degree • San Francisco
Apply
Fullstack Engineer 3 hours ago
$100k – $150k per year • Equity 0.2–2% • Remote • Full-Time • 1+ year exp • San Francisco
Python
JavaScript
C++
DevOps
AWS
Apply
See all jobs
This is one of many
412,234 more open roles from verified company boards, updated every day.