368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$84k – $204k per year (Estimated)
Location
Remote/Hybrid (Toronto, Canada)
Seniority
Architect
Employment
Full-Time
Overview
Company
Impact
Profile match
Rootly is an AI-native incident response and on-call management SaaS platform based in San Francisco, California. Founded in 2020 by Quentin Rousseau and JJ Tang (graduates of Y Combinator S21), the platform automates end-to-end incident response workflows directly within enterprise collaboration tools like Slack and Microsoft Teams.

About Rootly

At Rootly, we are on a mission to be the go-to way companies respond when things go wrong, helping every organization be more reliable. We do this by building an industry-leading incident management platform that allows companies around the world to consistently and quickly resolve incidents. We are not simply transforming an industry, we are carving an entirely new +$B segment ourselves and need incredible talent to achieve this ambitious goal together.

Customers love Rootly. Some of the fastest growing companies around the world such as NVIDIA, Figma, Canva, Tripadvisor, Squarespace and more rely on Rootly to power their critical incident management process. They obsess over our delightful enterprise-ready platform and unique partnership model. See why our customers have reviewed us 5 stars on G2.

Investors love Rootly. We are backed by some of the most respected funds in the world from Y Combinator to operators like the CTO of Dropbox and GitHub. We'd be happy to disclose our entire funding and profitability picture live during the interview. As a culture we relentlessly put transparency first. We conduct monthly financial reviews as a team so everyone has a pulse on the health of the business and publish what we are building in our weekly changelog.

About the Role

This is a rare opportunity to join Rootly as the founding leader of Platform Engineering and to shape the foundation that powers incident response and on call for some of the world’s most forward thinking engineering teams.

As Head of Platform, you will own two critical outcomes for the company:

  • Infrastructure Reliability and Scale, building a rock solid, redundant, scalable, operationally mature, and cost efficient infrastructure platform that supports our next tenfold of growth
  • Developer Experience and Velocity, crafting a world class developer experience that enables product engineers to move extremely fast, with safety and confidence, and that shapes how Rootly builds, tests, deploys, and operates
  • This is not a traditional ops role. This is a high leverage engineering leadership position for someone who combines deep technical skill, systems thinking, taste, and the ability to inspire teams to raise the bar. You will define strategy, hire the team, own the roadmap, and build the platform that makes Rootly engineering world class.

If you thrive where ownership is real, intensity is celebrated, and the platform you build directly determines the company’s trajectory, you will love it here.

What You'll Do

Platform leadership and vision

  • Own the vision, strategy, and roadmap for Rootly’s infrastructure and developer platform
  • Build and lead a high performing Platform Engineering organization that may include SRE, infrastructure, DevEx, and internal tooling
  • Establish a culture where reliability, performance, and developer experience are non negotiables
  • Act like an owner, spotting problems early, mobilizing teams, and driving solutions from concept to completion

Infrastructure reliability and scale

  • Architect a highly available, redundant, and scalable infrastructure foundation
  • Lead capacity planning, cost management, performance tuning, and long term infrastructure scaling
  • Drive operational maturity through infrastructure as code, declarative infrastructure, configuration management, and repeatable automation

Developer experience and productivity

  • Enable product engineers to move extremely quickly by optimizing local dev environments, ephemeral cloud environments, fast CI and CD, and reliable canaries
  • Provide tooling that abstracts infrastructure complexity and removes friction from development
  • Ensure every engineer can ship confidently, frequently, and safely

Reliability, observability, and incident response

  • Own platform wide SLOs, SLIs, and error budgets and use them to drive prioritization
  • Oversee observability tooling, monitoring, alerting, and incident response processes
  • Partner with product engineering teams to ensure services meet reliability and performance goals and to improve runbooks and postmortems

Execution, communication, and leadership

  • Drive high quality execution with urgency while balancing long term bets with tactical wins
  • Raise the bar and inspire engineers to think bigger, move faster, and deliver exceptional results
  • Collaborate closely with Product, Engineering, and leadership to align platform investments with company strategy
  • Recruit, mentor, and develop top tier platform engineers and create a culture of excellence

What We Are Looking For

We care about impact, ownership, and technical judgement. Degrees and big name logos are not required.

Minimum Qualifications

  • 10+ years in platform, infrastructure, SRE, or DevOps roles, with increasing leadership responsibility
  • Experience leading platform or SRE teams, including hiring, mentoring, and building culture
  • Deep expertise with cloud infrastructure, AWS preferred, distributed systems, scaling, and redundancy
  • Proven experience designing or operating high scale production systems and delivering operational maturity
  • Strong background in observability, performance tuning, and scaling strategies
  • Comfortable writing production grade software to solve infrastructure problems, Ruby or Go is a plus
  • Strong architectural judgement and systems thinking that anticipates scaling pain before it becomes real

Preferred Qualifications

  • Experience delivering DevEx tooling that materially improved developer velocity
  • Experience navigating startup to hypergrowth transitions and scaling infra and teams accordingly
  • High standards for taste and craftsmanship in platform engineering
  • Exceptional communicator, able to translate complex technical decisions for technical and non technical audiences
  • Bias toward action with the judgement to optimize versus ship at the right times

What Success Looks Like

By 6 to 12 months you will have

  • Built a small but elite Platform team with clear ownership and high morale
  • Dramatically accelerated engineering velocity with faster deploys, shorter tests, and fewer bottlenecks
  • Established high availability infrastructure with clear SLOs and stronger reliability across the board
  • Delivered developer tooling that makes engineering faster and more enjoyable
  • Positioned Rootly to scale tenfold in customers, traffic, and complexity

Why Rootly?

We’re not just another startup. We’re building something category-defining and want teammates who crave ownership, love solving hard problems, and thrive in a high-bar, high-impact environment.

Here’s what you can expect when you join Rootly:

  • Competitive compensation and early equity in a fast-growing, venture-backed company.
  • Comprehensive medical, dental, and vision coverage.
  • 3 weeks of vacation, plus unlimited sick and mental health days, and a company-wide end-of-year shutdown to recharge.
  • $500 stipend for home office setup.
  • Unlimited token usage and access to AI tools
  • A fast-moving, high-impact environment where your leadership and ideas directly shape the future of the company.

If this sounds like the kind of challenge and opportunity you’re looking for, apply now and let’s build something great together.

Rootly is an equal opportunity employer. We aim to create an environment where every team member at Rootly feels like they belong so they can have a greater impact on our business and customers. We do not discriminate on the basis of race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Toronto
$110k – $131k per year • Remote • Full-Time • 10+ years exp • Bachelor's Degree
DevOps
AWS
Incident Management
VMWare
Apply
$217k – $304k per year • Equity • Remote • Full-Time • 8+ years exp
Go
Databases
Apache Kafka
ClickHouse
Google BigQuery
AI/ML
Flink
Recommender Systems
DevOps
Incident Management
Kubernetes
Apply
$105k – $252k per year • Remote • Full-Time • 18+ years exp • Bachelor's Degree
Python
Java
Java
Gradle
DevOps
Ansible
AWS
CI/CD
CloudFormation
Configuration Management
Docker
GitHub Actions
GitLab CI
Helm
Jenkins
Kubernetes
Platform Engineering
Terraform
GitHub
GitLab
Cybersecurity
Sonatype Nexus IQ
Management
Confluence
Jira
Apply
$169k – $321k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Phoenix
AI/ML
AI Agents
Anomaly Detection
LLM Guardrails
DevOps
AWS
Kong
Amazon S3
API Gateway
Cybersecurity
Zero Trust
Apply
$90k – $115k per year • Remote • Full-Time • 8+ years exp
AI/ML
Runway
Design
Adobe After Effects
Adobe Photoshop
Figma
Apply
Remote/Hybrid • Internship • Toronto
AI/ML
LLM
Anthropic
DevOps
AIOps
AWS
Incident Management
Platform Engineering
GitHub
Design
Figma
Apply
$73k – $168k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Toronto
DevOps
Incident Management
GitHub
Design
Figma
Apply
$73k – $153k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Toronto
Go
Ruby
DevOps
CI/CD
Incident Management
GitHub
Design
Figma
Apply
$82k – $167k per year (Estimated) • Remote/Hybrid • Full-Time • Toronto
Ruby
Ruby
Ruby on Rails
DevOps
AWS
Incident Management
Kubernetes
GitHub
Design
Figma
Apply
$78k – $164k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Toronto
DevOps
Incident Management
GitHub
Design
Figma
Apply
Actuarial Analyst 25 min ago
$73k – $146k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Quebec • Waterloo • Toronto
Visual Basic
Apply
$47k – $109k per year (Estimated) • In office • Full-Time • Toronto
SQL
Visual Basic
Apply
$45k – $106k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Toronto
AI/ML
Copilot
Analytics
Power BI
Management
Power Apps
Apply
$48k – $112k per year (Estimated) • In office • Full-Time • Master's Degree • Toronto
C++
MATLAB
Python
Apply
Sr. UX Designer 1 hour ago
$76k – $161k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Toronto
Python
AI/ML
AI Agents
Hallucination
Human-in-the-Loop
LLM
LLM Guardrails
Design
Figma
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.