996,562open jobs
59,446companies
165,965added this week
Browse all
Location
In office
Seniority
Senior · 4+ years exp

Confirmed on the employer's own hiring board on Sep 29, 2026. First seen by Alion on Sep 23, 2026.

Overview
Company
Impact
Profile match
StarHub is a Singaporean telecommunications company offering mobile, broadband, pay television and enterprise services. It has expanded into cybersecurity and regional managed network businesses. The company is listed on the Singapore Exchange.

Job Description

You will design, develop, and deploy AI-powered, cloud-based products. As a Data Engineer, you’ll work with large-scale, heterogeneous datasets and hybrid cloud architectures to support analytics and AI solutions. Collaborate with data scientists, infra engineers, sales specialists, and stakeholders to ensure data quality, build scalable pipelines, and optimize performance. Your work will integrate telco data with other verticals (retail, healthcare), automate DataOps/MLOps/LLMOps workflows, and deliver production-grade systems.

As a Data Engineer, you will:

  • Ensure Data Quality & Consistency
  • Validate, clean, and standardize data (e.g., geolocation attributes) to maintain integrity.
  • Define and implement data quality metrics (completeness, uniqueness, accuracy) with automated checks and reporting.
  • Build & Maintain Data Pipelines
  • Develop ETL/ELT workflows (PySpark, Airflow) to ingest, transform, and load data into warehouses (S3, Postgres, Redshift, MongoDB).
  • Automate DataOps/MLOps/LLMOps pipelines with CI/CD (Airflow, GitLab CI/CD, Jenkins), including model training, deployment, and monitoring.
  • Design Data Models & Schemas
  • Translate requirements into normalized/denormalized structures, star/snowflake schemas, or data vaults.
  • Optimize storage (tables, indexes, partitions, materialized views, columnar encodings) and tune queries (sort/distribution keys, vacuum).
  • Integrate & Enrich Telco Data
  • Map 4G/5G infrastructure metadata to geospatial context, augment 5G metrics with legacy 4G, and create unified time-series datasets.
  • Consume analytics/ML endpoints and real-time streams (Kafka, Kinesis), designing aggregated-data APIs with proper versioning (Swagger/OpenAPI).
  • Manage Cloud Infrastructure
  • Provision and configure resources (AWS S3, EMR, Redshift, RDS) using IaC (Terraform, CloudFormation), ensuring security (IAM, VPC, encryption).
  • Monitor performance (CloudWatch, Prometheus, Grafana), define SLAs for data freshness and system uptime, and automate backups/DR processes.
  • Collaborate Cross-Functionally & Document
  • Clarify objectives with data owners, data scientists, and stakeholders; partner with infra and security teams to maintain compliance (PDPA, GDPR).
  • Document schemas, ETL procedures, and runbooks; enforce version control and mentor junior engineers on best practices.

Qualifications

Qualifications

· Bachelor’s or Master’s in Computer Science, Software Engineering, Data Science, or equivalent experience

· 4+ years in data engineering, analytics, or related AI/ML role

· Proficient in Python for ETL/data engineering and Spark (PySpark) for large-scale pipelines

· Experience with Big Data frameworks and SQL engines (Spark SQL, Redshift, PostgreSQL) for data marts and analytics

· Hands-on with Airflow (or equivalent) to orchestrate ETL workflows and GitLab CI/CD or Jenkins for pipeline automation

· Familiar with relational (PostgreSQL, Redshift) and NoSQL (MongoDB) stores: data modeling, indexing, partitioning, and schema evolution

· Proven ability to implement scalable storage solutions: tables, indexes, partitions, materialized views, columnar encodings

· Skilled in query optimization: execution plans, sort/distribution keys, vacuum maintenance, and cost-optimization strategies (cluster resizing, Spectrum)

· Experience with cloud platforms (AWS): S3/EMR/Glue, Redshift and containerization (Docker, Kubernetes)

· Infrastructure as Code using Terraform or CloudFormation for provisioning and drift detection

· Knowledge of MLOps/LLMOps: auto-scaling ML systems, model registry management, and CI/CD for model deployment

· Strong problem-solving, attention to detail, and the ability to collaborate with cross-functional teams

Nice to Have

· Exposure to serverless architectures (AWS Lambda) for event-driven pipelines

· Familiarity with vector databases, data mesh, or lakehouse architectures

· Experience using BI/visualization tools (Tableau, QuickSight, Grafana) for data quality dashboards

· Hands-on with data quality frameworks (Deequ) or LLM-based data applications (NL-->SQL generation)

· Participation in GenAI POCs (RAG pipelines, Agentic AI demos, geomobility analytics)

· Client-facing or stakeholder-management experience in data-driven/AI projects

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
996,562 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
In your city
Data Scientist 10 min ago
≈ $23k – $51k per year (Estimated) • Remote (likely EAEU) • 3+ years exp • Saint Petersburg
Python
SQL
Python
pySpark
AI/ML
Hadoop
Spark
MLlib
DevOps
Git
Apply
≈ $72k – $126k per year (Estimated) • Hybrid • Poland
SQL
Databases
Databricks
MS SQL
Azure Cosmos DB
Azure SQL Database
AI/ML
AI Agents
RAG
DevOps
Azure
Management
Agile
Apply
Data Engineer, Staff 3 months ago
≈ $132k – $262k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Newberg
Python
SQL
Databases
Snowflake
Databricks
DevOps
Azure
Analytics
ETL/ELT
Apply
≈ $70k – $155k per year (Estimated) • In office • 1+ year exp • Bachelor's Degree • Rochester
Python
Java
SQL
DevOps
GCP
Azure
AWS
Docker
Linux
Analytics
ETL/ELT
Apply
≈ $22k – $44k per year (Estimated) • In office • Bengaluru
Python
Rust
SQL
C#
DevOps
GCP
Azure
CI/CD
AWS
Analytics
ETL/ELT
Apply
up to $17k per year (gross) • In office • Full-Time • Moscow
Python
SQL
Databases
PostgreSQL
Apache Kafka
AI/ML
LangGraph
LangChain
Hadoop
AI Agents
LLM
DevOps
Rest API
Apply
≈ $101k – $196k per year (Estimated) • Hybrid • Home
Python
PowerShell
Bash
DevOps
Terraform
Azure DevOps
Crossplane
Azure
CI/CD
Kubernetes
Platform Engineering
Bicep
Azure AKS
Cybersecurity
Microsoft Entra ID
Apply
≈ $30k – $73k per year (Estimated) • Hybrid • 1+ year exp • Moscow
SQL
Databases
PostgreSQL
AI/ML
Function Calling
LLM
OpenAI
Tool Use
DevOps
Rest API
gRPC
WebSockets
CI/CD
Git
Docker
Linux
Apply
≈ $22k – $43k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Moscow
SQL
Databases
PostgreSQL
DevOps
Ansible
Zabbix
Jenkins
Grafana
Bitbucket
Unix
Management
Confluence
Jira
Apply
$53k – $80k per year • Equity • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Madrid
Python
C#
C#
.NET
Databases
MS SQL
SQLite
Design
SolidWorks
AutoCAD
Apply
≈ $21k – $46k per year (Estimated) • In office • Full-Time • 8+ years exp • Petaling Jaya
Databases
Snowflake
AI/ML
Amazon SageMaker
Machine Learning
DevOps
Splunk
CloudFormation
PagerDuty
Prometheus
CI/CD
AWS
Docker
Kubernetes
Blue-Green Deployment
Platform Engineering
Self-Healing
Amazon EKS
AWS Lambda
Amazon EC2
Cortex
Amazon S3
IAM
Amazon CloudWatch
Cybersecurity
GDPR
CVE
Management
Slack
ServiceNow
Apply
In office • 4+ years exp • Bachelor's Degree
Python
Scala
Python
pySpark
Databases
Apache Kafka
AI/ML
Spark
Flink
LLM
Machine Learning
DevOps
GCP
Azure
AWS
Analytics
Tableau
Management
Agile
Apply
≈ $21k – $46k per year (Estimated) • In office • 7+ years exp • Petaling Jaya
Python
SQL
Databases
Snowflake
AI/ML
Airflow
AI Agents
Streamlit
Anomaly Detection
Amazon SageMaker
Machine Learning
DevOps
Prometheus
CI/CD
AWS
Platform Engineering
AWS Lambda
Cortex
Amazon S3
Amazon CloudWatch
Analytics
ETL/ELT
Fivetran
Apply
≈ $21k – $46k per year (Estimated) • In office • 6+ years exp • Petaling Jaya
Databases
Snowflake
AI/ML
Streamlit
Amazon SageMaker
Machine Learning
DevOps
Splunk
CloudFormation
PagerDuty
Prometheus
CI/CD
AWS
Docker
Kubernetes
Blue-Green Deployment
Platform Engineering
Self-Healing
Amazon EKS
AWS Lambda
Amazon EC2
Cortex
Amazon S3
IAM
Amazon CloudWatch
Cybersecurity
GDPR
CVE
Management
Slack
ServiceNow
Apply
Data Scientist 22 days ago
In office • 3+ years exp • Bachelor's Degree
Python
Scala
Python
pySpark
Databases
Apache Kafka
AI/ML
Spark
AI Agents
Flink
LLM
Multi-Agent Systems
Machine Learning
DevOps
GCP
Azure
AWS
Analytics
Tableau
Management
Agile
Apply
See all jobs
This is one of many
996,562 more open roles from verified company boards, updated every day.