746,581open jobs
44,818companies
107,195added this week
Browse all
Location
In office (Singapore)

Confirmed on the employer's own hiring board on Sep 24, 2026. First seen by Alion on Sep 16, 2026.

Overview
Company
Impact
Profile match

SGX

Get Official Stock Quotes, Share Prices, Market Data & Many Other Investment Tools & Information From Singapore Exchange Ltd.

BUILD MARKETS. SHAPE ECONOMIES.

At SGX Group, we create markets. We turn ideas into products and expand access to new opportunities. We are building the future of the exchange in-house: from architecture and platforms to the critical systems that power markets. The biggest decisions are still open to people who join us now.

The opportunity

The ED, Site Reliability Engineering is a senior strategic executive leadership role responsible for shaping SGX’s enterprise reliability, observability, automation, and operational resilience agenda across critical platforms and services. This leader will help createa and own and own the long-term strategic view on strategic view on SRE vision and operating model, elevate engineering and resilience standards across the organisation across the organisation, and ensure reliability is embedded as a strategic differentiator that supports business growth, trusted operations, regulatory confidence, and customer outcomes. Working closely with senior stakeholders across business, technology, infrastructure, security, risk, compliance, and audit, the role leads senior SRE leaders, and teams and drives alignment between reliability investments, critical service commitments, and enterprise transformation priorities.

This is a rare opportunity to shape enterprise reliability at one of Asia’s leading market infrastructures, where technology resilience, market integrity, and customer trust are inseparable. The role offers meaningful executive visibility, direct impact on business-critical outcomes, and the platform to build a world-class SRE capability in a highly regulated, innovation-driven environment.

The Site Reliability Engineering team

Site Reliability Engineering keeps the platforms behind SGX Group's business and market infrastructure available, observable, and recoverable.

The team sets the standards other engineering teams work to across service levels, error budgets, observability, incident response and automation. It works across engineering, infrastructure, security, and product, and it owns the practice as well as seen as the SME for ensuring resilience and proactive maintenance and continuous improvement of the estate.

The next phase is about establishing SRE as a discipline rather than a function, moving reliability decisions upstream into design, and reducing the manual work that currently sits behind keeping services up.

What you will do

  • Reliability Engineering: Define enterprise reliability strategy and resilience objectives.
  • Observability: Establish enterprise observability strategy and investment roadmap.
  • Incident Management: Define the enterprise incident management framework and resilience standards.
  • SLO / SLA Management: Define enterprise service performance strategy.
  • Automation: Define the function automation strategy and lead the build-out of agentic AI adoption and improvements.
  • Performance Engineering: Define the organisational scalability and performance roadmap.
  • Capacity & Resilience Planning: Own enterprise resilience and continuity planning.
  • Cloud & Platform Operations: Define cloud operations and platform reliability strategy.
  • Operational Risk Management: Help build the function operational risk strategy.
  • Engineering Leadership: Build enterprise SRE capability and workforce development strategy.
  • Agentic AI for Reliability & Operations: Lead defining the function strategy for AI-enabled reliability engineering, including how AI agents support resilience, incident management, observability, operational risk reduction and measurable improvements in MTTD, MTTR and service stability.
  • Partner with executive stakeholders across business, engineering, infrastructure, operations, security, risk, compliance, and audit to align reliability priorities with business strategy, regulatory obligations, and critical service commitments.
  • Provide executive leadership during major incidents and crisis scenarios, ensuring decisive cross-functional action, strong stakeholder communication, and continuous improvement.

What we are looking for

  • Experience: Distinguished executive-level experience leading Site Reliability Engineering, DevSecOps, platform engineering, production engineering or large-scale technology operations within complex, always-on, mission-critical environments.
  • Track record: Proven track record of defining enterprise SRE strategies, scaling leadership teams and embedding reliability standards, observability practices and engineering disciplines across critical platforms, with strong fluency in DORA, SLIs, SLOs and error budgets.
  • Technical stack: Strong experience with observability platforms such as Datadog, New Relic or Grafana Cloud, and cloud-native infrastructure including AWS multi-account, Kubernetes/EKS and Terraform.
  • Standards & environment: Experience operating in regulated, high-availability environments, with strong understanding of operational resilience, technology risk, governance, audit and compliance expectations.
  • Communication & leadership: Exceptional executive presence, stakeholder management and communication skills, with strong commercial judgement, vendor management capability and the ability to attract, inspire and retain top-tier engineering talent.
  • Education & certifications: Bachelor’s degree in Computer Science, Engineering, Information Systems, or a related discipline; advanced qualifications are advantageous.
  • Exposure to financial market infrastructure, including securities and derivatives trading, clearing, settlement, market operations, or exchange-related systems, is a significant advantage.

Why this role matters

You will work on technology that underpins critical market infrastructure, where reliability is not a quality attribute of the product. It is the product. When these platforms work, participants trade and capital moves. When they do not, everyone knows within seconds.

The mandate is real. You will decide what reliability means here, how it is measured, and what the organisation is willing to trade for it. The work is demanding and the practice is still being built, which is exactly where the opportunity sits. There are not many chances in a career to establish a discipline rather than inherit one.

About SGX Group

SGX Group is one of the world's most trusted international marketplaces, known for its stability and openness. Anchored in Singapore, we enable price discovery, capital formation and risk management across asset classes, supported by resilient infrastructure and robust clearing. We convene issuers, investors and intermediaries to create and grow markets that stand the test of time. Find out more at www.SGXGroup.com.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
746,581 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Singapore
In office • Singapore
DevOps
CI/CD
Windows
Management
Agile
Apply
In office • Bachelor's Degree • Singapore
DevOps
DNS
DHCP
BGP
Apply
In office • Bachelor's Degree • Singapore
Python
Bash
DevOps
Ubuntu
CentOS Stream
Linux
TCP/IP
DNS
Apply
In office • Internship • Singapore
Python
DevOps
Linux
Apply
$343k – $686k per year • In office • Tokyo
DevOps
Terraform
CloudFormation
AWS
Kubernetes
Apply
≈ $34k – $72k per year (Estimated) • In office • Full-Time • 15+ years exp • Bachelor's Degree • Bengaluru
Python
JavaScript
SQL
C#
Cython
C#
ASP.NET Core
Entity Framework Core
Cython
Bottleneck
AI/ML
LangChain
AI Agents
LLM
RAG
DevOps
Rest API
GitHub Actions
Datadog
CI/CD
Jenkins
AWS
Docker
Kubernetes
QA
Selenium
JMeter
Swagger
Apply
$245k – $279k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco • New York • San Jose
Python
Go
JavaScript
Rust
TypeScript
C#
Scala
AI/ML
AI Agents
Machine Learning
DevOps
GCP
Azure
AWS
HPC
Apply
≈ $12k – $22k per year (Estimated) • Remote (EAEU) • Full-Time • 2+ years exp • Bachelor's Degree • Vladivostok
SQL
Databases
PostgreSQL
RabbitMQ
DevOps
Rest API
Puppet
Ansible
Docker Compose
Zabbix
VMWare
Chef
etcd
Prometheus
Docker
Kubernetes
Ubuntu
Nginx
Grafana
Graylog
KVM
CentOS Stream
SLI/SLO/SLA
Linux
Astra Linux
TCP/IP
DNS
SOAP
Apache HTTP Server
Management
Confluence
ITIL
ITSM
Apply
≈ $14k – $27k per year (Estimated) • Remote (EAEU) • 2+ years exp • Bachelor's Degree • Moscow
SQL
Databases
PostgreSQL
RabbitMQ
DevOps
Rest API
Puppet
Ansible
Docker Compose
Zabbix
VMWare
Chef
etcd
Prometheus
Docker
Kubernetes
Ubuntu
Nginx
Grafana
Graylog
KVM
CentOS Stream
SLI/SLO/SLA
Linux
Astra Linux
TCP/IP
DNS
SOAP
Apache HTTP Server
Management
Confluence
ITIL
ITSM
Apply
Security Engineer 9 hours ago
≈ $75k – $162k per year (Estimated) • In office • Full-Time • Bachelor's Degree • London
DevOps
Incident Management
Apply
≈ $100k – $240k per year (Estimated) • In office • Bachelor's Degree • Singapore
AI/ML
AI Agents
DevOps
Terraform
New Relic
Datadog
AWS
Kubernetes
Grafana
Platform Engineering
Amazon EKS
Incident Management
SLI/SLO/SLA
Apply
≈ $100k – $240k per year (Estimated) • In office • 5+ years exp • Singapore
Python
Go
Rust
Kotlin
TypeScript
AI/ML
Agentic Workflows
DevOps
Terraform
GCP
CI/CD
AWS
Kubernetes
Apply
≈ $108k – $269k per year (Estimated) • In office • Bachelor's Degree • Singapore
AI/ML
AI Agents
DevOps
CI/CD
Kubernetes
Platform Engineering
Incident Management
Apply
≈ $86k – $235k per year (Estimated) • In office • Bachelor's Degree • Singapore
Python
Java
TypeScript
AI/ML
AI Agents
DevOps
GCP
OpenTelemetry
Prometheus
Azure
CI/CD
GitOps
AWS
Grafana
Apply
≈ $97k – $252k per year (Estimated) • In office • Singapore
Python
Rust
TypeScript
SQL
Databases
PostgreSQL
Neo4j
pgvector
AI/ML
LangGraph
LangChain
Claude
Model Context Protocol
Vertex AI
Embeddings
Gemini
LLM
RAG
Google ADK
OpenAI Agents SDK
GraphRAG
Knowledge Graph
LLM Evaluation
DevOps
Rest API
GCP
Platform Engineering
Google Cloud Run
Cybersecurity
Least Privilege
Management
Confluence
SharePoint
Apply
≈ $54k – $115k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Singapore
Management
Outlook
Apply
≈ $66k – $145k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Singapore
Analytics
Microsoft Excel
Apply
In office • Bachelor's Degree • Singapore
Robotics
Digital Twin
IoT
OPC UA
Apply
≈ $48k – $132k per year (Estimated) • In office • Internship • Singapore
SQL
Apply
In office • Bachelor's Degree • Singapore
Apply
See all jobs
This is one of many
746,581 more open roles from verified company boards, updated every day.