686,079open jobs
39,763companies
97,004added this week
Browse all
Salary
$110k – $150k per year
Location
Remote (United States)
Employment
Full-Time
Overview
Company
Impact
Profile match

About Us

TherapyNotes is the go-to superhero for behavioral health Practice Management and EHR software! Our top-notch SaaS solution handles scheduling, billing, documenting, telehealth, and more so clinicians can focus on awesome patient care.

We're a dynamic team of pros who love to innovate and push the envelope, keeping our software cutting-edge. Join us, and let's revolutionize behavioral health software together while making a real difference!

About The Position

We are seeking a Site Reliability Engineer to improve the reliability and operability of the production services and shared platforms supporting our growing 24×7 SaaS environment. In this role, you will apply software and systems engineering practices to improve availability, performance, scalability, resilience, observability, incident response, and operational automation. You will partner with software development, infrastructure, database, security, and other technology teams to establish measurable reliability goals, reduce operational toil, and ensure services are supportable throughout their lifecycle. If you are passionate about building reliable systems, solving complex production problems, and driving continuous improvement, we want to hear from you.

Requirements

  • BS degree in Information Systems, Engineering, or equivalent experience.
  • 5+ years of engineering experience in Systems Engineering, Cloud or Platform Engineering, DevOps, Software Engineering, and/or SRE.
  • Experience designing and operating production systems using cloud-based compute, storage, networking, and containerization technologies; Azure and Kubernetes preferred.
  • Strong Linux systems and networking fundamentals, with experience troubleshooting complex distributed systems in production.
  • Expertise with an observability platform; Datadog experience strongly preferred. Experience with Prometheus, Grafana, New Relic, or equivalent platforms is also valuable.
  • Experience with scripting and operational automation using tools such as Bash, PowerShell, or Python, along with infrastructure-as-code and configuration-management practices.
  • Experience participating in production on-call rotations, incident response, root cause analysis, and post-incident improvement.
  • Experience working in Agile/DevOps environments and operating production services using ITSM practices where applicable.
  • Prior software development experience-or experience investigating application behavior through code, logs, and distributed traces-is a plus

Responsibilities

  • Own and continuously improve how we use Datadog to make reliability visible and actionable across metrics, logs, traces, dashboards, monitors, alerts, and service-level views.
  • Design, implement, and maintain high-availability, high-throughput, data- and compute-intensive critical systems supporting a growing 24×7 SaaS platform.
  • Partner with service owners to define and improve reliability through meaningful SLIs, SLOs, error budgets, actionable alerting, and operational-readiness practices.
  • Participate in and help drive incident management for production events, serving as an incident commander or technical responder as needed. Coordinate triage, service restoration, escalation, communication, incident documentation, root cause analysis, and completion of corrective actions.
  • Partner with development teams to investigate issues across the infrastructure and application layers using metrics, logs, distributed traces, and code-level context.
  • Improve deployment safety and service resilience through automated validation, recovery and rollback capabilities, reliability testing, and analysis of system failure modes.
  • Partner with other technical leaders to ensure all newly introduced systems are supportable and maintainable by both development and operations.
  • Provide escalated technical guidance and support to other technology teams throughout the organization.
  • Provide on-call coverage for production support and other duties as required.
  • Ensure supported systems and operational activities comply with organizational security, HIPAA, and operating policies.
  • Identify and eliminate repetitive operational toil using Bash, PowerShell, Python, or Ansible. Manage infrastructure as code using Terraform/OpenTofu and configuration automation using Ansible.

Benefits

  • Competitive salary - $110,000-$150,000
  • Employer sponsored health, dental, vision, life, and disability insurance
  • Retirement plan with company contribution
  • Annual company profit sharing
  • Personal development/training budget
  • Open, collaborative work environment
  • Extensive 2-week onboarding plan
  • Comprehensive mentorship program

Equal Opportunity Employer Statement & Applicant Rights

TherapyNotes LLC is an Equal Opportunity Employer and does not discriminate based on race, color, religion, sex, national origin, age, disability, genetic information, or any other protected status under federal, state, or local law. We are committed to providing a workplace free of discrimination and harassment.For more information about your rights under federal employment laws, please review the following:

If you require a reasonable accommodation during the application process, please contact [email protected].

9/17/2026

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
686,079 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Philadelphia
$110k – $150k per year • Remote • Full-Time • Bachelor's Degree • Philadelphia
Python
JavaScript
TypeScript
PowerShell
C#
Node JS
C#
.NET
Frontend
Angular
pnpm
DevOps
Terraform
GitHub Actions
Docker Swarm
Azure
CI/CD
ArgoCD
Docker
Kubernetes
Platform Engineering
Octopus Deploy
GitHub
Cybersecurity
HIPAA
Management
Agile
Apply
Dev Ops Engineer 2 hours ago
In office • TS/SCI • 3+ years exp
Python
Rust
SQL
PowerShell
Rust
Rocket
Frontend
WebAssembly
DevOps
Ansible
VMWare
Windows Server
Docker
Ubuntu
GitHub
GitLab
Management
Agile
Apply
$48k – $128k per year (Estimated) • Remote • Part-Time
Python
AI/ML
Vertex AI
Embeddings
Scikit-learn
TensorFlow
PyTorch
DevOps
GCP
Apply
$53k – $149k per year (Estimated) • Remote/Hybrid • Full-Time • Master's Degree • Dublin
Python
C++
DevOps
CI/CD
Git
Apply
$91k – $139k per year • In office • Top Secret • Full-Time • 8+ years exp • Bachelor's Degree • Albuquerque
Python
MATLAB
MATLAB
Simulink
Apply
$110k – $150k per year • Remote • Full-Time • Bachelor's Degree • Philadelphia
Python
JavaScript
TypeScript
PowerShell
C#
Node JS
C#
.NET
Frontend
Angular
pnpm
DevOps
Terraform
GitHub Actions
Docker Swarm
Azure
CI/CD
ArgoCD
Docker
Kubernetes
Platform Engineering
Octopus Deploy
GitHub
Cybersecurity
HIPAA
Management
Agile
Apply
$120k – $140k per year • Remote • Full-Time • Bachelor's Degree • Philadelphia
JavaScript
TypeScript
SQL
C#
C#
ASP.NET Core
Entity Framework Core
Databases
PostgreSQL
AI/ML
AI Agents
LLM
LLM Guardrails
Agentic Workflows
Frontend
Angular
Sass
Mobile
Ionic
Capacitor
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
HIPAA
Management
Agile
Apply
$110k – $150k per year • Remote • Full-Time • Bachelor's Degree • Philadelphia
DevOps
Terraform
OpenTofu
Azure
CI/CD
GitOps
AWS
Kubernetes
Cloudflare
Azure AKS
Cybersecurity
ISO 27001
PCI DSS
HIPAA
Zero Trust
Microsoft Entra ID
Apply
$110k – $150k per year • Remote • Full-Time • Bachelor's Degree • Philadelphia
DevOps
Terraform
GitHub Actions
Azure
CI/CD
AWS
GitHub
Cybersecurity
Snyk
HIPAA
Zero Trust
Threat Modeling
Apply
$95k – $125k per year • Remote • Full-Time • Bachelor's Degree • Philadelphia
JavaScript
C#
DevOps
CI/CD
Management
Agile
QA
Selenium
Cucumber
JMeter
Playwright
Gatling
Appium
k6
Apply
$255k – $497k per year • In office • Full-Time • 14+ years exp • Bachelor's Degree • New York • San Francisco • Los Angeles • Dallas • Pittsburgh
Apply
$201k – $229k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • McLean • New York • Wilmington • Philadelphia
Apply
$115k – $150k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • Philadelphia • Allentown
Apply
$55k – $142k per year (Estimated) • Equity • In office • Full-Time • Philadelphia • Pittsburgh
Management
Outlook
Apply
IT Support Engineer-1 7 hours ago
$77k – $128k per year • Remote/Hybrid • Full-Time • Philadelphia
Cybersecurity
Microsoft Entra ID
Management
Jira
ServiceNow
Apply
See all jobs
This is one of many
686,079 more open roles from verified company boards, updated every day.