738,400open jobs
44,340companies
105,262added this week
Browse all
Salary
$96k – $264k per year
Location
In office (Vienna)
Seniority
Principal · 8+ years exp

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Sep 15, 2026. Oracle scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Oracle is an American enterprise technology company founded in 1977 by Larry Ellison, Bob Miner and Ed Oates, and headquartered in Austin, Texas. It built its business on the Oracle Database, still the reference relational engine for large transactional systems, and has expanded into a full applications suite covering finance, human resources, supply chain and customer experience through Fusion Cloud and NetSuite. Its fastest growing segment is Oracle Cloud Infrastructure, which the company has positioned aggressively for AI training and inference workloads through large multi-year capacity contracts.

We are seeking a highly motivated Site Reliability Engineer (SRE) to support a strategic customer's cloud platform and mission-critical applications. The SRE will be responsible for ensuring high availability, operational excellence, automation, and continuous improvement of production environments.

The successful candidate will partner with application development, platform engineering, cloud infrastructure, networking, cybersecurity, and operations teams to build resilient, secure, and highly automated systems. This role emphasizes reliability engineering, observability, incident response, capacity planning, and proactive operational improvements.

Serves as a consultant and leads the design and architecture of infrastructure and service, ensuring alignment with reliability and functionality standards. Takes full ownership of forecasting of demands and responding to capacity needs. Owns collaborations with software development teams to develop reliable and scalable infrastructures. Oversees incident response and/or maintenance tasks. Provides strategic, future-oriented health and performance reporting. Contributes to strategies for automation and reviews the development and implementation of automation. Leads the implementation of innovative tools and provides expertise in site reliability trends.

Key Responsibilities

  • Maintain high availability, reliability, and performance of enterprise applications and cloud infrastructure.
  • Design and implement automation to reduce manual operational effort and improve deployment consistency.
  • Develop Infrastructure as Code (IaC) using Terraform and automation scripts using Python, Bash, or similar languages.
  • Build and maintain CI/CD pipelines to support automated deployments and release management.
  • Monitor applications and infrastructure using observability platforms, including metrics, logs, traces, and alerting.
  • Define and maintain Service Level Indicators (SLIs), Service Level Objectives (SLOs), and error budgets.
  • Participate in on-call rotations and respond to production incidents with urgency and professionalism.
  • Conduct root cause analysis (RCA) and implement corrective and preventive actions to eliminate recurring issues.
  • Perform capacity planning, performance tuning, and scalability assessments.
  • Support Kubernetes clusters, containerized workloads, and cloud-native applications.
  • Collaborate with development teams to improve application resiliency, fault tolerance, and operational readiness.
  • Partner with security teams to ensure systems comply with enterprise security and regulatory requirements.
  • Create operational runbooks, documentation, and standard operating procedures.
  • Continuously improve platform reliability through automation, monitoring, and operational best practices.
  • Takes full ownership of forecasting infrastructure demands and strategically responds to capacity needs, ensuring systems have sufficient resources to handle current and future workloads, anticipating resource gaps.
  • Seeks opportunities for prototyping and encourages the implementation of prototyping initiatives (e.g., testing new applications or infrastructures, assisting in onboarding), driving new approaches.
  • Identifies and projects resource gaps using cost calculators, identifying alternative solutions to lower costs, when necessary.
  • Provides guidance to others when monitoring services, maintains up-to-date knowledge of their performance, and meticulously documents their condition.
  • Oversees root cause analyses for incidents and/or maintenance on assigned services (e.g., software installs, version upgrades, security updates, backup and recovery), ensuring efficient execution and preventing incident reoccurrence.
  • Provides strategic, future-oriented health and performance reporting, and anticipates actions needed based on trends in data.
  • Provides expert-level release notes, communication, and/or guidance on the scale, capacity, security, performance attributes, and requirements of services and technology to customers, cross-functional teams, leadership, and external stakeholders.
  • Provides leadership in the on-call shifts.
  • Leads the resolution of multifaceted technical issues spanning multiple services, functions, and customers, collaborating with cross-functional teams and leveraging advanced investigation and debugging techniques to ensure the achievement of SLOs (service level objectives).
  • Provides input on strategic initiatives to address opportunities to improve performance bottlenecks and deployments, maximizing resource utilization, cost efficiency, speed, and scalability across their line of business.
  • Provides expertise in site reliability trends, guiding the creation and sharing of insights and best practices to shape the future of building, testing, deploying, and running services.

Required Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent experience.
  • 8+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, or Systems Engineering.
  • Experience supporting production environments with strict availability requirements.
  • Hands-on experience with Oracle Cloud Infrastructure (OCI) or another major cloud provider (AWS, Azure, or GCP).
  • Experience with Kubernetes, Docker, and container orchestration platforms.
  • Proficiency in Infrastructure as Code (Terraform preferred).
  • Experience with CI/CD tools such as Jenkins, GitHub Actions, GitLab CI, or Azure DevOps.
  • Strong scripting skills in Python, Bash, or PowerShell.
  • Experience with Linux system administration.
  • Strong understanding of networking concepts including DRG, DNS, load balancing, routing, APIs, and Endpoints.
  • Experience with Observability tools such as Prometheus, Grafana, ELK/OpenSearch, Splunk, Datadog, New Relic, or OCI Native Observability services.
  • Strong troubleshooting and problem-solving skills.
  • Experience supporting financial services or other highly regulated industries.
  • Knowledge of disaster recovery, backup strategies, and business continuity.

Desired Skills

  • Site Reliability Engineering (SRE)
  • Oracle Cloud Infrastructure (OCI)
  • Kubernetes
  • Docker
  • Terraform
  • Python
  • Bash
  • Linux Administration
  • CI/CD Pipelines
  • Git
  • Monitoring & Observability
  • Incident Management
  • Root Cause Analysis (RCA)
  • Automation
  • Capacity Planning
  • Performance Tuning
  • Networking
  • Cloud Security
  • High Availability
  • Disaster Recovery
  • DevOps
  • Agile Methodologies

Success Measures

  • Achieve and maintain service availability and reliability targets.
  • Reduce mean time to detect (MTTD) and mean time to recover (MTTR).
  • Increase deployment frequency while minimizing change failure rates.
  • Reduce operational toil through automation and self-service capabilities.
  • Improve platform observability and proactive issue detection.
  • Meet defined SLOs and error budget objectives.
  • Deliver resilient, secure, and scalable cloud platforms that support client's critical business services.

This role is ideal for an engineer who is passionate about automation, operational excellence, and building highly reliable cloud platforms. The successful candidate will play a key role in ensuring client's production environments remain secure, resilient, and capable of supporting mission-critical financial services at scale.

Disclaimer:

Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.

Range and benefit information provided in this posting are specific to the stated locations only

US: Hiring Range in USD from: $96,300 to $264,100 per annum. May be eligible for bonus, equity, and compensation deferral.

Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.

Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.

Oracle US offers a comprehensive benefits package which includes the following:

1. Medical, dental, and vision insurance, including expert medical opinion

2. Short term disability and long term disability

3. Life insurance and AD&D

4. Supplemental life insurance (Employee/Spouse/Child)

5. Health care and dependent care Flexible Spending Accounts

6. Pre-tax commuter and parking benefits

7. 401(k) Savings and Investment Plan with company match

8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.

9. 11 paid holidays

10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.

11. Paid parental leave

12. Adoption assistance

13. Employee Stock Purchase Plan

14. Financial planning and group legal

15. Voluntary benefits including auto, homeowner and pet insurance

The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.

Career Level - IC5

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
738,400 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Vienna
$29k – $41k per year • In office • Full-Time • Osaka
Python
JavaScript
Databases
Redis
Frontend
React.js
DevOps
Jenkins
Git
AWS
Nginx
GitHub
Linux
Management
Slack
Trello
Apply
$35k – $98k per year (Estimated) • In office • Full-Time • Master's Degree • Tokyo
DevOps
AWS
Apply
Equity • Remote/Hybrid • Full-Time • 4+ years exp • Master's Degree
Python
JavaScript
PHP
SQL
Apex
AI/ML
Model Context Protocol
AI Agents
Frontend
JQuery
DevOps
GCP
Azure
AWS
Management
Slack
Apply
$36k – $99k per year (Estimated) • In office • Full-Time • Master's Degree • Tokyo
DevOps
GCP
AWS
Platform Engineering
Management
Slack
Google Workspace
Zapier
Apply
$36k – $100k per year (Estimated) • In office • Full-Time • Master's Degree • Tokyo
Chips/EDA
PoC Library
Management
Slack
Google Workspace
Apply
In office • Full-Time • Bachelor's Degree • Pune
Python
JavaScript
Java
TypeScript
Databases
Redis
Frontend
Angular
DevOps
Splunk
New Relic
Incident Management
IAM
Management
ITIL
Apply
Аналитик 1 day ago
$20k – $50k per year (Estimated) • In office • Full-Time • Moscow
Python
SQL
Databases
ClickHouse
Management
Telegram
Apply
System Analyst 1 day ago
$49k – $125k per year (Estimated) • In office • Full-Time • 4+ years exp • Tbilisi
Python
Rust
TypeScript
DevOps
Rest API
SOAP
Management
UML
Apply
$71k – $185k per year (Estimated) • In office • Full-Time • PhD • Singapore
Python
Robotics
Agisoft Metashape
Apply
$22k – $47k per year (Estimated) • In office • Full-Time • Pune
Python
SQL
C++
MATLAB
SAS
Apply
$27k – $59k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Bengaluru
Java
DevOps
Rest API
Terraform
CI/CD
Management
Agile
Apply
$27k – $59k per year (Estimated) • In office • Bengaluru
Apply
$113k – $217k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Seattle
Python
C++
Databases
Oracle
DevOps
Azure
AWS
Linux
Management
Agile
Apply
$115k – $223k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Santa Clara
Python
Go
Java
SQL
Bash
Databases
Redis
RabbitMQ
Memcached
Apache Kafka
DevOps
Terraform
Ansible
GCP
Helm
Jaeger
OpenTelemetry
CloudFormation
Prometheus
Azure
AWS
Docker
Kubernetes
Grafana
IAM
Linux
Unix
TCP/IP
DNS
Cybersecurity
ISO 27001
SOC 2
Apply
In office • Bangkok
Apply
$127k – $222k per year • Equity • In office • Full-Time • 8+ years exp • Vienna
AI/ML
AI Agents
Management
ServiceNow
Apply
$180k – $282k per year • Equity • In office • Full-Time • 15+ years exp • Vienna
Management
ServiceNow
Apply
$148k – $232k per year • Equity • In office • Full-Time • 12+ years exp • Master's Degree • Vienna
Management
ServiceNow
Apply
$171k – $220k per year • Equity • In office • Full-Time • Vienna
Management
ServiceNow
Apply
Shift Leader - 4318 2 days ago
$35k – $84k per year (Estimated) • In office • Full-Time • Master's Degree • Vienna
Apply
See all jobs
This is one of many
738,400 more open roles from verified company boards, updated every day.