368,746open jobs
9,444companies
47,506added this week
Browse all
Location
In office (Kuala Lumpur)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

GBG

Global Benefits Group was founded in 1981. This company provides multiple insurance packages and options as well as other financial services. Their headquarters are located in Foothill Ranch, California.

Enabling safe and rewarding digital lives for genuine people, everywhere

We make it our mission to ensure more genuine people have digital access to opportunities, and businesses have access to more genuine people. Our technology draws on diverse and reliable data to create a single point of truth for identity and address verification.

With over 30 years of experience behind us our team and technology are focused on enabling safe and rewarding digital lives for everyone. Regardless of age, location or background, genuine people everywhere should be able to digitally prove who they are and where they live.

About the team and role

Global Fraud Solutions

The team provides decision support solutions to address business objectives in risk prevention and fraud detection. We deliver software solutions and offer client support using our expertise and a client-focused approach.

Site Reliability Engineer

The SRE will build and operate the reliability, observability, and operational excellence infrastructure underpinning the GFS managed fraud detection platforms. You will work across deployment pipelines, cloud infrastructure, monitoring, and incident management - ensuring GBG can deliver on high availability SLAs for banking and fintech customers who depend on real-time fraud detection at scale.

What you will do

  • Design and operate the SRE practice for Managed oferings, including on-call processes, SLA frameworks, incident response playbooks, and post-incident review (PIR) processes.
  • Build and maintain observability infrastructure: centralised logging (correlation IDs), metrics dashboards, distributed tracing, and alerting for the Predator/Instinct platform stack.
  • Define and track SLOs (Service Level Objectives) and error budgets for real-time transaction processing pipelines, targeting high TPS and low round-trip latency.
  • Manage cloud infrastructure provisioning and configuration using IaC tooling (Terraform, Helm), supporting both AWS/Azure cloud deployments and on-premises customer environments.
  • Implement and maintain CI/CD pipelines for GFS solutions (Jenkins, etc.)
  • Work with Engineering teams to ensure security and compliance readiness for Managed services - including PCI DSS, ISO 27001, SOC 1/2/3, PDPA/GDPR - in close coordination with InfoSec teams.
  • Drive platform resilience improvements: high availability, auto-scaling, disaster recovery, backup/restore procedures, and chaos engineering practices.
  • Manage secrets, certificate rotation, identity/access controls (OAuth/RBAC), and vulnerability management for the hosted environment.
  • Support performance testing methodology and baseline establishment for our products.
  • Contribute to the Architecture Review Committee (ARC) with SRE and operational perspectives on technology choices.
  • Collaborate with engineering squads to embed reliability and DevSecOps practices across the SDLC.

Skills we’re looking for

  • Minimum 5 years of solid hands-on experience in a Site Reliability, Platform Engineering, or DevOps role, ideally supporting mission-critical real-time processing systems in banking, payments, or fintech.
  • Strong proficiency with cloud platforms (AWS preferred; Azure/GCP acceptable) including networking, compute, storage, and managed services.
  • Deep expertise with containerisation and orchestration: Docker, Kubernetes (EKS/AKS/GKE), Helm, and associated tooling.
  • Infrastructure as Code experience: Terraform (required), and familiarity with Ansible or Pulumi.
  • Observability stack proficiency: Prometheus, Grafana, ELK/OpenSearch, Jaeger/Zipkin, or equivalent enterprise-grade tooling.
  • CI/CD pipeline design and management: GitHub Actions, Jenkins, ArgoCD, or equivalent.
  • Experience with security and compliance frameworks applicable to hosted financial services: PCI DSS, ISO 27001, SOC 1/2/3, GDPR/PDPA.
  • Familiarity with database reliability practices for SQL Server, PostgreSQL, and Oracle - including replication, read replicas, and failover.
  • Working knowledge of secrets management (HashiCorp Vault, AWS Secrets Manager) and zero-trust identity principles.
  • Experience supporting real-time streaming or event-driven architectures (Kafka, RisingWave, or similar) in production environments.
  • Scripting and automation proficiency: Python, Bash, or Go for operational tooling.
  • Strong sense of operational ownership and accountability - comfortable being on-call and driving incidents to resolution.
  • Excellent communication skills - able to produce clear incident reports, runbooks, and architecture documentation for both technical and executive audiences.
  • Proactive mindset: identifies reliability risks before they become incidents and champions a culture of blameless post-mortems.
  • Collaborative and effective working with software engineers, product managers, and InfoSec teams.
  • Continuous improvement orientation - always looking to reduce toil, automate repetitive tasks, and improve platform resilience.
  • Flexible and adaptable - able to support a globally distributed product with customers across multiple time zones.

To find out more

As an equal opportunity employer, we are dedicated to creating a diverse and inclusive workplace where everyone feels valued and empowered. Please inform your GBG Talent Attraction Partner if you require any reasonable adjustments to the interview process.

To chat to the Talent Attraction team and find out more about our benefits and why we’re a great place to work, drop an email to [email protected] and we’ll be in touch. You can also find out more about careers at GBG and check out our current opportunities at gbgplc.com/careers.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Kuala Lumpur
$96k – $134k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • New York
JavaScript
Swift
TypeScript
Java
Java
Spring Framework
Databases
Apache Kafka
PostgreSQL
AI/ML
AI Agents
Claude
Copilot
Fine-tuning
Flink
LangChain
LangGraph
Llama
LlamaIndex
Prompt Engineering
PyTorch
RAG
TensorFlow
Transformers
Devin
Hugging Face
OpenAI
Frontend
Angular
React.js
Mobile
MVC
DevOps
AWS
CI/CD
Docker
Kubernetes
OpenShift
Splunk
Vector
GitHub
Analytics
Tableau
Apply
$38k – $91k per year (Estimated) • Remote • Full-Time • 8+ years exp • PhD • Guadalajara
PowerShell
Python
Databases
Amazon Aurora
DynamoDB
AI/ML
Amazon SageMaker
AWS Bedrock
AWS Bedrock AgentCore
Ray
DevOps
Amazon CloudWatch
Amazon EC2
Amazon ECS
Amazon EKS
Amazon EventBridge
Amazon S3
AWS
AWS Lambda
AWS Step Functions
Azure
CI/CD
Datadog
FinOps
GCP
Git
GitLab
GitLab CI
IAM
Jenkins
JFrog Artifactory
Kubernetes
New Relic
Service Mesh
Splunk
Terraform
Cybersecurity
HIPAA
ISO 27001
PCI DSS
SOC 2
Apply
Senior AI Architect 1 hour ago
$113k – $189k per year • In office • Full-Time • Bachelor's Degree • Barcelona
AI/ML
Anomaly Detection
Computer Vision
CrewAI
LangGraph
RAG
LangChain
LLMOps
AI Agents
Time Series Forecasting
DevOps
AWS
Azure
GCP
Management
n8n
Apply
$158k – $317k per year • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Durham
DevOps
AWS
Azure
GCP
Cybersecurity
Zero Trust
Apply
$256k – $320k per year • Remote/Hybrid • Full-Time • Seattle
Go
Python
AI/ML
AI Agents
CrewAI
LangChain
LangGraph
DevOps
AWS
GCP
Kubernetes
Apply
$96k – $195k per year (Estimated) • Remote • Full-Time
C++
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
Anomaly Detection
Computer Vision
Image Segmentation
LoRA
OpenCV
PyTorch
QLoRA
TensorFlow
PEFT
DevOps
CI/CD
Apply
$165k – $302k per year (Estimated) • Remote • Full-Time • Atlanta
Apply
In office • Full-Time • 3+ years exp • Kuala Lumpur
Python
AI/ML
Anomaly Detection
MLFlow
Vertex AI
Amazon SageMaker
Feature Store
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Analytics
A/B Testing
Apply
$139k – $270k per year (Estimated) • In office • Full-Time • Manchester
C++
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
Anomaly Detection
Computer Vision
Image Segmentation
LoRA
OpenCV
PyTorch
QLoRA
TensorFlow
PEFT
DevOps
CI/CD
Apply
In office • Full-Time • Kuala Lumpur
Apply
Remote/Hybrid • Full-Time • 1+ year exp • Kuala Lumpur
Apply
In office • Bachelor's Degree • Kuala Lumpur
Apply
In office • Full-Time • 10+ years exp • Kuala Lumpur
COBOL
SQL
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Kuala Lumpur
ABAP
ABAP
ABAP Objects
DevOps
Rest API
SLI/SLO/SLA
Apply
In office • Full-Time • 8+ years exp • Bachelor's Degree • Kuala Lumpur
ABAP
DevOps
SLI/SLO/SLA
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.