368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$195k – $270k per year
Location
Remote (United States)
Seniority
Architect · 15+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Everbridge is a critical event management software company headquartered in Burlington, Massachusetts, and founded in 2002. The company runs a platform that detects threats such as severe weather, active shooters, or IT outages, then locates affected people and sends coordinated alerts across phone, SMS, email, and public warning channels. It serves corporations, hospitals, universities, and government agencies including national public warning systems, and was taken private by Thoma Bravo in 2024.

Everbridge is seeking a Vice President, Global Production Operations & Reliability to lead the strategy, execution, and continuous improvement of our global production operations. Reporting to the Chief Technology Officer, this executive will be responsible for ensuring the reliability, scalability, security, and operational excellence of our cloud-native SaaS platform.

Everbridge empowers organizations to keep people safe and operations running during critical events. Our Critical Event Management platform is trusted by enterprises, governments, healthcare providers, financial institutions, and public safety organizations around the world to deliver resilient, mission-critical services when they matter most.

This role will lead global Site Reliability Engineering (SRE) and Development & Reliability Engineering (DRE) teams while driving operational excellence across production operations, incident management, change governance, disaster recovery, observability, and platform engineering. The successful candidate will play a key role in strengthening customer trust by delivering highly available, resilient services on a global scale.

What you'll do:

    • Define and execute Everbridge's global production operations and reliability strategy.
    • Lead and develop high-performing global SRE and DRE teams responsible for platform reliability and operational engineering.
    • Own the operational excellence of Everbridge's AWS and Kubernetes-based cloud platform, ensuring scalability, resilience, security, and performance.
    • Establish best practices for service reliability, including Service Level Objectives (SLOs), error budgets, production readiness, capacity planning, observability, and operational automation.
    • Drive disciplined incident management, change governance, release management, and post-incident reviews to continuously improve platform stability and reduce operational risk.
    • Lead disaster recovery planning, resilience testing, business continuity initiatives, and operational readiness across global production environments.
    • Partner closely with Engineering, Product, Security, Customer Support, and Customer Success to embed reliability into the software development lifecycle and improve customer outcomes.
    • Champion automation, cloud-native engineering practices, and continuous improvement to enhance operational efficiency and platform performance.
    • Provide executive leadership and reporting on operational health, reliability metrics, customer-impacting incidents, and strategic initiatives.

What you'll bring:

  • 15+ years of experience in production operations, cloud infrastructure, platform engineering, SRE, or related technology leadership roles.
  • Proven experience leading global production operations for large-scale, mission-critical SaaS or cloud platforms.
  • Deep expertise in AWS, Kubernetes, cloud-native architectures, distributed systems, and high-availability environments.
  • Strong understanding of SRE principles, incident management, observability, disaster recovery, change management, CI/CD, and Infrastructure as Code.
  • Demonstrated success building and leading high-performing global engineering and operations teams.
  • Excellent executive communication, stakeholder management, and cross-functional leadership skills.

Preferred Qualifications

  • Experience in mission-critical industries such as public safety, critical communications, healthcare, financial services, security, or enterprise SaaS.
  • Familiarity with modern SRE practices, progressive delivery, platform engineering, service mesh, and cloud cost optimization.
  • Knowledge of regulatory and operational frameworks including ISO 27001, SOC 2, NIST, or FedRAMP.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$25k – $42k per year • Equity 0–0.2% • Remote • Full-Time • 3+ years exp
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
$100k – $210k per year • Equity 0–0.5% • Remote • Full-Time • 3+ years exp • San Francisco
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
Team Lead DevOps 1 day ago
$23k – $62k per year (Estimated) • Remote • 5+ years exp • Moscow
Bash
Python
Erlang
Erlang
EMQX
Databases
Apache Kafka
ClickHouse
PostgreSQL
RabbitMQ
Redis
Redpanda
Trino
DevOps
Ansible
AWS
AWX
FinOps
HAProxy
Hetzner
Kubernetes
SLI/SLO/SLA
Terraform
Yandex Cloud
Amazon S3
Apply
$100k – $200k per year • Equity 0.5–5% • In office • Full-Time • 1+ year exp • New York
Python
TypeScript
JavaScript
Python
FastAPI
Databases
DynamoDB
PostgreSQL
AI/ML
Claude
LLM
OpenAI
AI Agents
Frontend
Next.js
Tailwind CSS
React.js
DevOps
AWS
Docker
Vercel
GitHub
Management
Slack
Apply
$19k – $28k per year (net) • Remote • Full-Time • Moscow
C#
C++
C++
CMake
DevOps
CI/CD
Git
Management
Jira
Slack
Apply
$60k – $144k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree
JavaScript
Python
TypeScript
Java
Java
Spring Boot
Databases
Apache Kafka
MariaDB
PostgreSQL
AI/ML
Claude
Claude Code
Copilot
Cursor
OpenAI Codex
Frontend
Angular
DevOps
Ansible
Docker
Helm
Kubernetes
GitHub
Apply
$27k – $67k per year (Estimated) • Remote • Full-Time • 5+ years exp
Java
JavaScript
Java
Spring Boot
AI/ML
Claude
Claude Code
Copilot
Cursor
Ignite
Windsurf
PyTorch
OpenAI Codex
Frontend
React.js
DevOps
AWS
CI/CD
Docker
Incident Management
Kubernetes
Platform Engineering
GitHub
Apply
$27k – $66k per year (Estimated) • Remote • Full-Time • 5+ years exp • Bachelor's Degree
C#
JavaScript
TypeScript
C#
.NET
Databases
Apache Kafka
Frontend
npm
React.js
DevOps
AWS
CI/CD
Docker
Kubernetes
GitLab
Apply
Remote • Full-Time
Databases
PostgreSQL
DevOps
Incident Management
Platform Engineering
Apply
Platform Engineer 18 days ago
$76k – $161k per year (Estimated) • Remote • Full-Time
PowerShell
SQL
Databases
PostgreSQL
DevOps
Ansible
Azure
Azure DevOps
CI/CD
Git
Hyper-V
Platform Engineering
Proxmox VE
VMWare
Windows Server
GitLab
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.