713,007open jobs
42,513companies
99,377added this week
Browse all
Salary
$18k – $44k per year (Estimated)
Location
In office (Petaling Jaya)
Seniority
Middle · 3+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 23, 2026. First seen by Alion on Sep 23, 2026. Grab scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Grab is a Southeast Asian technology company founded in 2012 by two Harvard Business School classmates as a taxi-hailing app for Malaysia, and now the region's dominant everyday services platform. It operates across eight countries in ride hailing, food and grocery delivery, courier services and digital payments, and famously absorbed Uber's Southeast Asian operations in 2018 in exchange for a stake in the company. Headquartered in Singapore and listed on Nasdaq since 2021, Grab has expanded into financial services through GXS Bank and its lending and insurance arms, and builds its own mapping stack through GrabMaps.

About Grab and Our Workplace

Grab is Southeast Asia's leading superapp. From getting your favourite meals delivered to helping you manage your finances and getting around town hassle-free, we've got your back with everything. In Grab, purpose gives us joy and habits build excellence, while harnessing the power of Technology and AI to deliver the mission of driving Southeast Asia forward by economically empowering everyone, with heart, hunger, honour, and humility.

Get to Know the Team

The Data Compute Platform team is an important contributor to Grab's data ecosystem, allowing growth through democratization of data at scale. We build and operate Grab's data infrastructure and efficient platform that supports internal data processes and company-wide data lake access. Our tech stack uses industry-leading distributed compute engines like Apache Spark, Ray, Trino, and Starrocks, orchestrated by Airflow and Michelangelo and backed by AWS S3. Our evolving Data Lake storage architecture uses modern open-source formats like Apache Iceberg and Delta in addition to traditional Apache Hive Parquet tables.

Underneath these engines sits a large, multi-tenant Kubernetes and AWS Infrastructure that we own end to end. This infrastructure includes the clusters, the operators, the autoscaling, the networking, the identity and access model, the observability, and the cost controls. These components work together to keep the platform fast, reliable, and affordable for thousands of pipelines and queries every day.

Get to Know the Role

You will be an important contributor to the infrastructure layer of the Data Compute Platform. This layer consists of the Kubernetes, AWS, and infrastructure-as-code foundations. Spark, Ray, Trino, Starrocks, Airflow, and Michelangelo run on these foundations. You will design how compute is provisioned, scaled, secured, observed and paid for, and you will drive the reliability and cost-efficiency of the platform as it grows. As a senior engineer, you will take ownership of well-scoped infrastructure projects from design through rollout and operation. You will also contribute to the team's engineering and SRE standards. Additionally, you will support other engineers through code review and knowledge sharing. You will also explore new developments in the cloud-native and data infrastructure space and integrate them into our ecosystem to the benefit of the data community at Grab.

You will report to our Data Engineering Manager II, and you will based onsite in our Petaling Jaya office.

The Critical Tasks You Will Perform

  • You will design, build, and operate the multi-tenant Kubernetes (EKS) platform. This platform runs Grab's workloads, including Spark, Ray, Trino, Starrocks, Airflow, and Michelangelo. Additionally, your responsibilities will include cluster lifecycle, node provisioning, and autoscaling, and scheduling and resource isolation.
  • You will build Kubernetes operators and custom resources (kubebuilder / controller-runtime, Go) that automate the provisioning and lifecycle of compute engines and their tenants.
  • You will lead the infrastructure-as-code (Terraform) and CI/CD (GitLab) that provision and change our AWS estate. This estate includes EKS, S3, IAM, RDS, VPC, and networking. You will drive it towards safe, reviewable, automated change.
  • You will design and implement the platform's identity, access and security model across AWS IAM, Kubernetes RBAC and service identities, working with the storage access and security teams.
  • You will build the observability, alerting, capacity planning, and incident tooling for the platform. You will contribute to the SRE practice, which includes SLOs, runbooks, on-call, and post-incident reviews. You will reduce toil and MTTR.
  • You will own compute cost efficiency: instance and storage strategy, spot and right-sizing, bin-packing, idle reclamation, and cost attribution back to tenants.
  • You will drive architectural improvements and migrations (for example engine version upgrades, cluster consolidation, new execution backends), managing the design, phased rollout and rollback plan with guidance from senior team members.

What Essential Skills You Will Need

  • Software Engineering, Computer Science, or related undergraduate degree.
  • You have 3 or more years of experience, with at least 2 years building and operating production infrastructure or platform services at scale.
  • You have programming proficiency in Go and/or Python, with the habit of treating infrastructure as software: tested, reviewed, versioned and automated.
  • You have deep, hands-on Kubernetes experience in production: cluster operations, scheduling, autoscaling, networking, storage, RBAC and multi-tenancy. We prefer experience building custom controllers or operators.
  • You have experience with AWS (EKS, EC2, S3, IAM, VPC) and infrastructure as code with Terraform.
  • You have proficiency in CI/CD tooling (GitLab CI, Jenkins or similar) and GitOps-style delivery.
  • You have solid SRE fundamentals: observability (Datadog, Prometheus, Grafana or equivalent), SLOs, incident management and capacity planning, with experience running reliable services.

Skills that are Good to have

  • Working knowledge of at least one of Spark, Ray, Airflow, Trino or Starrocks, and a appetite to learn how distributed data engines behave on Kubernetes.
  • Experience running Apache Spark on Kubernetes at scale (Spark Operator, dynamic allocation, shuffle services) and tuning its interaction with the resource manager.
  • Experience with Kubernetes autoscaling and scheduling tooling such as Karpenter, Cluster Autoscaler, Yunikorn or Volcano.
  • Experience deploying and operating Ray on Kubernetes (KubeRay), including cluster autoscaling and isolation for ML and batch workloads.
  • Experience operating distributed query engines like Trino or Starrocks, including cluster sizing, fault tolerance and workload isolation.
  • Experience with container runtimes and internals (containerd, cgroups, OverlayFS, image build and distribution).
  • Experience with service mesh, ingress and network policy (Istio, Envoy, Cilium or similar) in multi-tenant clusters.
  • Experience with FinOps for compute platforms: cost attribution, and spot / reserved capacity strategy.
  • Contributions to open-source cloud-native or data infrastructure projects.

Life at Grab

We care about your well-being at Grab, here are some of the global benefits we offer:

  • We have your back with Term Life Insurance and comprehensive Medical Insurance.
  • With GrabFlex, create a benefits package that suits your needs and aspirations.
  • Celebrate moments that matter in life with loved ones through Parental and Birthday leave, and give back to your communities through Love-all-Serve-all (LASA) volunteering leave
  • We have a confidential Grabber Assistance Programme to guide and uplift you and your loved ones through life's challenges.
  • Balancing personal commitments and life's demands are made easier with our FlexWork arrangements such as differentiated hours

What We Stand For At Grab

We are committed to building an inclusive and equitable workplace that provides equal opportunity for Grabbers to grow and perform at their best. We consider all candidates fairly and equally regardless of nationality, ethnicity, race, religion, age, gender, family commitments, physical and mental impairments or disabilities, and other attributes that make them unique.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
713,007 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Petaling Jaya
$17k – $45k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Petaling Jaya
Python
SQL
AI/ML
Spark
Reinforcement Learning
TensorFlow
PyTorch
Recommender Systems
Machine Learning
Apply
$34k – $75k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Zagreb
Python
DevOps
Rest API
Windows
Cybersecurity
PCI DSS
SIEM
Apply
$91k – $181k per year (Estimated) • Remote/Hybrid • Full-Time • London
JavaScript
TypeScript
SQL
PowerShell
C#
Frontend
Redux
GraphQL
Tailwind CSS
React.js
React Query
DevOps
Terraform
WebSockets
CI/CD
Git
AWS
AWS Fargate
AWS Lambda
API Gateway
Windows
Management
Agile
Apply
$95k – $178k per year (Estimated) • In office • Full-Time • 10+ years exp • Melbourne
Python
Java
SQL
Java
Spring Boot
Databases
Couchbase
Apache Kafka
Mobile
JUnit
DevOps
Rest API
OpenShift
CI/CD
Docker
Kubernetes
API Gateway
Cybersecurity
SonarQube
Management
Agile
Apply
$98k – $184k per year (Estimated) • In office • Full-Time • 10+ years exp • Melbourne
Python
Go
Java
C++
Bash
Java
Spring Boot
DevOps
Rest API
Docker
Kubernetes
Linux
Unix
Apply
$17k – $45k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Petaling Jaya
Python
SQL
AI/ML
Spark
Reinforcement Learning
TensorFlow
PyTorch
Recommender Systems
Machine Learning
Apply
$23k – $55k per year (Estimated) • In office • Full-Time • 15+ years exp • Pasig
Apply
$70k – $178k per year (Estimated) • In office • Full-Time • 6+ years exp • Singapore
Apply
$66k – $149k per year (Estimated) • In office • Full-Time • 8+ years exp • PhD • Singapore
Apply
$86k – $226k per year (Estimated) • In office • Full-Time • 6+ years exp • Bachelor's Degree • Jakarta
Python
Go
JavaScript
Rust
TypeScript
C#
C++
Node JS
Scala
AI/ML
Copilot
LLM
Frontend
React.js
Apply
$22k – $46k per year (Estimated) • In office • Full-Time • 8+ years exp • Petaling Jaya
DevOps
Git
GitHub
GitLab
Management
Confluence
Apply
$11k – $26k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Petaling Jaya
Management
Agile
Apply
$16k – $38k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Petaling Jaya
DevOps
Ansible
CI/CD
Apply
$19k – $48k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Petaling Jaya
Apply
In office • Full-Time • 2+ years exp • Bachelor's Degree • Petaling Jaya
Management
Microsoft Office
Marketing
LinkedIn
Instagram
Apply
See all jobs
This is one of many
713,007 more open roles from verified company boards, updated every day.