368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$81k – $187k per year
Location
In office
Seniority
Senior
Overview
Company
Impact
Profile match
Oracle Corporation is an American multinational computer technology corporation headquartered in Austin, Texas. Founded in 1977 by Larry Ellison, Bob Miner, and Ed Oates, Oracle is one of the world's largest enterprise software and cloud computing infrastructure providers.

We are seeking a Site Reliability Engineer to help build and operate reliable, scalable cloud-native platforms and services that support Oracle Health’s next-generation healthcare technology initiatives. This role works across platform engineering, cloud infrastructure, distributed systems, integration services, healthcare applications, and operational tooling.

The ideal candidate is passionate about reliability, automation, observability, and performance. You will collaborate closely with software engineering, product, security, operations, and customer-facing teams to design resilient systems, improve service health, respond to incidents, and support modernization efforts across healthcare workflows, interoperability, data platforms, and AI-driven capabilities.

This role bridges software engineering and operations: designing infrastructure and services for reliability, developing automation, improving monitoring and alerting, supporting incident response and root cause analysis, and contributing to secure, scalable systems running on Oracle Cloud Infrastructure.

Responsibilities

  • Design, build, test, and operate reliable cloud infrastructure, platform capabilities, and services on Oracle Cloud Infrastructure and legacy deployment models.
  • Partner with software engineering teams to develop scalable, resilient services, APIs, integrations, and distributed systems.
  • Forecast capacity needs, analyze service trends, and take proactive steps to ensure systems can support current and future workloads.
  • Monitor service health, availability, latency, performance, and capacity using observability and reporting tools.
  • Define and maintain meaningful SLIs, SLOs, KPIs, dashboards, alerts, and runbooks for production services.
  • Improve service resilience through backup and restore validation, disaster recovery planning, secrets handling, patching, and least-privilege access practices.
  • Participate in incident response, troubleshooting, root cause analysis, postmortems, and follow-up remediation.
  • Develop automation, scripts, and tooling to support provisioning, deployment, monitoring, metrics collection, mitigation, and remediation.
  • Support safe release practices, including CI/CD, infrastructure automation, canary or blue-green deployments, rollback planning, and operational readiness reviews.
  • Investigate and debug issues across applications, infrastructure, services, and dependencies to help teams meet service level objectives.
  • Identify performance bottlenecks and reliability risks, then recommend and implement improvements.
  • Collaborate with product managers, architects, engineers, security, operations, and customer teams to deliver secure, customer-focused healthcare solutions.
  • Support modernization efforts involving cloud-native architectures, healthcare interoperability, large-scale healthcare data platforms, and AI-enabled capabilities.
  • Communicate service health, operational risks, capacity concerns, and the potential impact of infrastructure, feature, or tooling changes.
  • Contribute to documentation, runbooks, incident records, operational standards, and knowledge sharing.
  • Participate in on-call rotations and operational support for production services.

Required Skills

  • Linux and networking: Processes, filesystems, systemd, DNS, TCP/IP, TLS, HTTP, load balancers, proxies, and basic database behavior.
  • Kubernetes operations: Deployments, Services/Ingress, ConfigMaps/Secrets, RBAC, resource requests/limits, probes, autoscaling, persistent storage, Helm/Kustomize, container troubleshooting, and effective kubectl troubleshooting.
  • Cloud and infrastructure-as-code: Oracle Cloud Infrastructure OCI or other cloud experience (AWS, GCP, Azure) , IAM, networks, compute, managed Kubernetes, Terraform, and configuration automation such as Ansible.
  • Delivery engineering: Git, GitHub, container images/registries, CI/CD, safe release practices including canary deployments, blue-green deployments, rollback strategies, operational readiness, and ideally GitOps.
  • Observability and reliability: Metrics, logs, traces, dashboards, useful alerts, SLIs/SLOs, error budgets, KPIs, incident response, on-call support, runbooks, postmortems, capacity planning, and performance tuning.
  • Automation: Strong Bash/Shell scripting and Python experience; PowerShell and Go are useful differentiators. Focus on eliminating recurring toil through code, scripting, and repeatable automation.
  • Security and recovery: Least-privilege access, secrets handling, image/dependency hygiene, patching, vulnerability remediation, backup/restore, and disaster-recovery testing.
  • Healthcare and data systems are strongly preferred: SQL, healthcare technology operations, healthcare interoperability, and familiarity with FHIR, HL7, or large-scale healthcare data platforms.
  • Collaboration and communication: Calm incident communication, clear root-cause analysis, collaboration with developers, technical communication, knowledge sharing, and influencing systems toward simpler, safer operations.
Disclaimer:

Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.

Range and benefit information provided in this posting are specific to the stated locations only

US: Hiring Range in USD from: $81,100 to $187,000 per annum. May be eligible for bonus and equity.

Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.

Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.

Oracle US offers a comprehensive benefits package which includes the following:

1. Medical, dental, and vision insurance, including expert medical opinion

2. Short term disability and long term disability

3. Life insurance and AD&D

4. Supplemental life insurance (Employee/Spouse/Child)

5. Health care and dependent care Flexible Spending Accounts

6. Pre-tax commuter and parking benefits

7. 401(k) Savings and Investment Plan with company match

8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.

9. 11 paid holidays

10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.

11. Paid parental leave

12. Adoption assistance

13. Employee Stock Purchase Plan

14. Financial planning and group legal

15. Voluntary benefits including auto, homeowner and pet insurance

The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.

Career Level - IC3

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$140k – $225k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Seattle
C++
Go
Python
Rust
Python
FastAPI
Databases
Neo4j
pgvector
Qdrant
PostgreSQL
AI/ML
LLM
SGLang
vLLM
Knowledge Graph
AI Agents
DevOps
AWS
Azure
Bicep
GCP
Karpenter
KEDA
Kubernetes
OpenTelemetry
OpenTofu
Terraform
Vector
Cybersecurity
Least Privilege
Apply
$105k – $252k per year • Remote • Full-Time • 18+ years exp • Bachelor's Degree
Python
Java
Java
Gradle
DevOps
Ansible
AWS
CI/CD
CloudFormation
Configuration Management
Docker
GitHub Actions
GitLab CI
Helm
Jenkins
Kubernetes
Platform Engineering
Terraform
GitHub
GitLab
Cybersecurity
Sonatype Nexus IQ
Management
Confluence
Jira
Apply
$68k – $85k per year • In office • Full-Time • Master's Degree • San Jose
Python
AI/ML
AI Agents
DevOps
Amazon EC2
AWS
AWS Lambda
Bitbucket
CI/CD
CloudFormation
Docker
Git
Kubernetes
Terraform
Amazon S3
IAM
HPC
Cybersecurity
Least Privilege
Apply
$185k – $260k per year • Remote • Full-Time • 8+ years exp • Bachelor's Degree
DevOps
AWS
CI/CD
GCP
Kubernetes
GitHub
Cybersecurity
Clair
Dependabot
OWASP Top 10
OWASP ZAP
Snyk
Trivy
Apply
$110k – $131k per year • Remote • Full-Time • 10+ years exp • Bachelor's Degree
DevOps
AWS
Incident Management
VMWare
Apply
$71k – $166k per year • Equity • In office • Bachelor's Degree
Apply
$115k – $235k per year • Equity • In office • Bachelor's Degree • Nashville
Go
Java
Python
AI/ML
AutoGen
Claude
Claude Code
Copilot
CrewAI
Cursor
LangChain
LangGraph
LlamaIndex
LLM
RAG
AI Agents
Function Calling
Human-in-the-Loop
LLM Guardrails
OpenAI Codex
Structured Outputs
DevOps
AWS
Azure
Docker
Kubernetes
Apply
In office • Bachelor's Degree
DevOps
AWS
Azure
CI/CD
Kubernetes
Cybersecurity
FedRAMP
Apply
In office
DevOps
SLI/SLO/SLA
Management
Slack
Apply
In office
DevOps
Incident Management
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.