969,314open jobs
58,495companies
159,022added this week
Browse all
Salary
≈ $13k – $32k per year (Estimated)
Location
In office (Chennai)
Seniority
Middle · 6+ years exp

Confirmed on the employer's own hiring board on Sep 30, 2026. First seen by Alion on Sep 28, 2026. Pearson scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Pearson is a London-based learning company that publishes educational courseware, runs assessments and qualifications, and provides digital learning, English-language and workforce skills services. Founded in 1844 and listed in London and New York, it employs more than 20,000 people, and its UK Assessment and Qualifications arm delivers nearly four million exam results a year, including GCSE, A-Level and BTEC. It hires for software and data engineering, product management, channel sales, partnership management, payroll, HR operations and tutoring roles across India, the UK, the US, Poland, Spain, Israel and the Philippines.

Data Engineer - Global Data Team

Role Overview

We are looking for an experienced Data Engineer with 6-10 years of hands-on experience to join our Global Data team and help build reliable, scalable, and self-service data platforms that power enterprise-wide analytics and data products.

In this role, you will design, develop, and optimize modern cloud-based data platforms, lakehouse architectures, and distributed data pipelines. You will work extensively with technologies such as Google Cloud Platform (GCP), Google BigQuery, GCP Data engineering services, Data flow, Data Fusion, Data Streams, Cloud Storage, Cloud Composer/Airflow, dbt, Python, and SQL, while contributing to data governance, observability, automation, and AI-powered data engineering capabilities.

The ideal candidate is a hands-on engineer who enjoys solving complex data challenges, building reusable engineering frameworks, improving platform reliability, and enabling data teams through scalable self-service capabilities.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines supporting enterprise data products and analytics.
  • Build and optimize data solutions using Google BigQuery, GCS, Cloud Composer/Airflow, dbt, Python, and SQL.
  • Develop robust data ingestion and transformation pipelines from databases, SaaS applications, APIs, files, and other enterprise data sources.
  • Implement and maintain API integrations and data ingestion frameworks for cloud and enterprise applications.
  • Develop reusable dbt models, transformation frameworks, and data quality processes.
  • Build and manage workflow orchestration using Apache Airflow / Google Cloud Composer.
  • Work with CDC and data ingestion technologies such as Informatica CDC, Airflow, Composer, dbt, or similar platforms.
  • Design and implement modern data lakehouse architectures using BigQuery and related cloud technologies.
  • Optimize data pipelines and BigQuery workloads for performance, scalability, reliability, and cost efficiency.
  • Implement engineering best practices including CI/CD, version control, automated testing, deployment automation, and DevOps practices.
  • Contribute to data observability and monitoring frameworks, including pipeline health, data quality, SLA monitoring, and operational metrics.
  • Partner with data architects, analysts, data scientists, product teams, and business stakeholders to deliver high-quality data solutions.
  • Contribute to data governance, metadata management, lineage, and data discovery capabilities.
  • Explore and implement AI/ML and GenAI capabilities to improve data engineering workflows, automation, data discovery, and developer productivity.
  • Troubleshoot complex data pipeline and platform issues and drive root-cause analysis and long-term improvements.
  • Help establish engineering standards, reusable frameworks, and best practices across the Global Data organization.
  • Champion self-service data capabilities that improve the experience of data consumers and engineering teams.

Required Skills and Experience

  • 6-10 years of professional experience in Data Engineering, Data Platform Engineering, or a closely related field.
  • Strong hands-on experience developing ETL/ELT pipelines and data transformation workflows.
  • Strong programming experience with Python.
  • Strong SQL skills, including experience with complex queries, optimization, and large-scale data processing.
  • Hands-on experience with dbt and modern data transformation practices.
  • Hands-on experience with Apache Airflow and/or Cloud Composer for workflow orchestration.
  • Strong experience working with Google BigQuery or a comparable cloud data warehouse.
  • Experience with Google Cloud Storage (GCS) and cloud-based data platforms.
  • Experience building and integrating data pipelines using REST APIs and other data integration mechanisms.
  • Strong understanding of data modeling, data warehousing, and modern lakehouse architectures.
  • Experience working with large-scale distributed data pipelines and production data environments.
  • Experience with Git, CI/CD, automated testing, and DevOps practices.
  • Strong problem-solving and analytical skills with the ability to troubleshoot complex data engineering issues.
  • Ability to work effectively in a global, collaborative, and cross-functional environment.
  • Excellent written and verbal communication skills.

Preferred Skills

  • Strong experience with GCP data services, particularly BigQuery, GCS, Cloud Composer, and related services.
  • Experience with Informatica CDC / Mass Ingestion or other change-data-capture technologies.
  • Experience with additional cloud platforms such as AWS or Azure.
  • Experience with data governance and metadata management platforms, such as Collibra or similar tools.
  • Experience implementing data observability frameworks and tools.
  • Experience with data quality, lineage, cataloging, and metadata management.
  • Experience with Terraform or Infrastructure as Code.
  • Experience with containerization and modern DevOps technologies.
  • Exposure to AI/ML and GenAI technologies, including using LLMs to automate or enhance data engineering workflows.
  • Experience building self-service data platforms, reusable engineering frameworks, or data products.
  • Experience with data security, privacy, access controls, and enterprise data governance.
  • Experience optimizing cloud data platforms for performance and cost.

What You Will Bring

  • A hands-on engineering mindset with a passion for building production-grade data solutions.
  • Strong ownership and accountability for the reliability and quality of data pipelines and platforms.
  • A continuous-improvement mindset and enthusiasm for automation, standardization, and reusable frameworks.
  • Ability to simplify complex technical problems and develop scalable solutions.
  • Curiosity and willingness to learn and adopt emerging technologies, particularly AI/ML and GenAI.
  • A strong focus on platform reliability, data quality, observability, and developer/user experience.
  • Ability to collaborate effectively with globally distributed engineering, architecture, product, and business teams.
  • Strong communication skills and the ability to explain complex technical concepts to both technical and non-technical audiences.
  • Passion for building self-service capabilities that enable teams to discover, access, understand, and use data effectively.

Why Join Us

  • Build and scale next-generation cloud data platforms powering enterprise-wide analytics, data products, and AI initiatives.
  • Join Pearson, a global leader transforming lives through learning and innovation.
  • Work hands-on with GenAI, AI-powered data engineering, and intelligent automation.
  • Shape the future of self-service data, data products, governance, and observability at scale.
  • Collaborate with high-performing global data and technology teams solving real-world, high-impact problems.
  • Work with modern technologies across GCP, BigQuery, dbt, Airflow/Composer, Python, and AI/ML.
  • Have the opportunity to influence data engineering standards, architecture, and platform strategy.
  • Accelerate your career in a fast-evolving, innovation-driven global data ecosystem.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
969,314 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
Chennai
Sr. Data Engineer 1 year ago
$21k – $26k per year • In office • Full-Time • 5+ years exp • Ahmedabad
Python
Java
SQL
Scala
Databases
Snowflake
Apache Kafka
Google BigQuery
Amazon Redshift
BigQuery
AI/ML
Spark
Flink
DevOps
GCP
Azure
CI/CD
AWS
Analytics
ETL/ELT
Apply
≈ $21k – $43k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Gurgaon
Python
SQL
Databases
Databricks
AI/ML
Hadoop
Spark
XGBoost
Scikit-learn
TensorFlow
PyTorch
RAG
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Analytics
Tableau
Power BI
A/B Testing
Management
Agile
Apply
≈ $18k – $42k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Python
JavaScript
Java
SQL
Node JS
Scala
Databases
Amazon Redshift
DevOps
AWS
AWS Lambda
Amazon S3
IAM
Amazon Kinesis
Analytics
ETL/ELT
AWS Glue
Apply
Senior Data Engineer 3 hours ago
≈ $22k – $44k per year (Estimated) • In office • Full-Time • 4+ years exp • Bengaluru • Mumbai • Thiruvananthapuram
Python
SQL
Scala
Python
pySpark
Databases
Snowflake
Databricks
AI/ML
Spark
DevOps
GCP
Azure DevOps
Azure
AWS
Analytics
Dimensional Modeling
Apply
≈ $23k – $47k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Bengaluru
Python
AI/ML
Machine Learning
Apply
Hybrid • Full-Time • 2+ years exp • Sydney
Python
Go
JavaScript
Databases
PostgreSQL
Frontend
React.js
DevOps
Terraform
GCP
Prometheus
GitLab CI
CI/CD
Jenkins
AWS
Docker
Kubernetes
Grafana
TeamCity
Management
Agile
Apply
AI Engineer 4 hours ago
≈ $13k – $32k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Bengaluru
Python
Databases
Chroma
Pinecone
FAISS
AI/ML
LangChain
LlamaIndex
Prompt Engineering
LLM
RAG
Multi-Agent Systems
Machine Learning
DevOps
Rest API
Git
Docker
Apply
≈ $16k – $38k per year (Estimated) • In office • Bachelor's Degree • Moscow
SQL
Databases
PostgreSQL
DevOps
Rest API
Management
Scrum
Apply
In office • 3+ years exp • Batumi
Java
SQL
Java
Maven
Gradle
Databases
RabbitMQ
Apache Kafka
Mobile
JUnit
DevOps
Rest API
CI/CD
Git
AWS
Docker
Kubernetes
Management
Agile
Scrum
QA
TestNG
Swagger
Postman
Rest-Assured
Pact
Apply
$40k – $74k per year • In office • Secret • Internship • Bachelor's Degree • Rolling Meadows
Python
MATLAB
AI/ML
CUDA Toolkit
OpenCL
CUDA
DevOps
GitHub
Management
Agile
Apply
≈ $27k – $54k per year (Estimated) • Hybrid • 5+ years exp • Bengaluru
Python
SQL
Databases
Google BigQuery
Azure SQL Database
Microsoft Fabric
BigQuery
AI/ML
AutoGen
LangChain
Embeddings
Function Calling
AI Agents
Semantic Kernel
RAG
OpenAI
LLMOps
Human-in-the-Loop
Tool Use
DevOps
GCP
Azure
CI/CD
SLI/SLO/SLA
Analytics
Power BI
Azure Data Factory
Looker
Dimensional Modeling
Management
Power Automate
Power Apps
Apply
Data Engineer III 23 days ago
≈ $14k – $32k per year (Estimated) • Hybrid • Bengaluru
Python
SQL
Scala
Databases
MySQL
PostgreSQL
Databricks
Amazon Aurora
Azure SQL Database
DevOps
Azure
CI/CD
AWS
GitHub
Analytics
ETL/ELT
Azure Data Factory
Management
Agile
Apply
≈ $18k – $36k per year (Estimated) • In office • 3+ years exp • Bachelor's Degree • Chennai
Python
SQL
Assembly
Assembly
Keystone Engine
Databases
Snowflake
Databricks
DevOps
Rest API
Azure
AWS
Platform Engineering
Analytics
ETL/ELT
Collibra
Apply
≈ $18k – $41k per year (Estimated) • In office • 5+ years exp • Bengaluru
Python
AI/ML
Anomaly Detection
DevOps
Splunk
Terraform
GCP
Azure DevOps
GitHub Actions
Loki
New Relic
OpenTelemetry
OpenTofu
Datadog
Dynatrace
PagerDuty
Prometheus
GitLab CI
Azure
CI/CD
Jenkins
AWS
Kubernetes
Grafana
Platform Engineering
Configuration Management
AppDynamics
Opsgenie
AIOps
Incident Management
IAM
Management
Jira
ServiceNow
ITSM
Apply
≈ $12k – $30k per year (Estimated) • In office • 3+ years exp • Chennai
Python
AI/ML
Anomaly Detection
DevOps
Splunk
Terraform
GCP
Azure DevOps
GitHub Actions
Loki
New Relic
OpenTelemetry
OpenTofu
Datadog
Dynatrace
PagerDuty
Prometheus
GitLab CI
Azure
CI/CD
Jenkins
AWS
Kubernetes
Grafana
Platform Engineering
Configuration Management
AppDynamics
Opsgenie
AIOps
Incident Management
IAM
Management
Jira
ServiceNow
ITSM
Apply
Data Engineer 6 hours ago
≈ $13k – $30k per year (Estimated) • In office • 1+ year exp • Chennai
Python
SQL
Databases
PostgreSQL
AI/ML
NumPy
DevOps
Git
AWS
Analytics
ETL/ELT
QA
Selenium
Playwright
Apply
In office • Full-Time • 5+ years exp • Bengaluru • Pune • Chennai • Mumbai • Hyderabad
Python
SQL
AI/ML
LLM
RAG
DevOps
AWS
Amazon S3
Analytics
ETL/ELT
QA
Selenium
Pytest
Apply
Back Office Manager 1 hour ago
≈ $7.5k – $21k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Chennai
Management
Microsoft Office
Apply
≈ $16k – $35k per year (Estimated) • In office • Contractor • 3+ years exp • Bachelor's Degree • Chennai
Analytics
Microsoft Excel
Apply
≈ $16k – $35k per year (Estimated) • In office • Contractor • 8+ years exp • Bachelor's Degree • Chennai
Analytics
Microsoft Excel
Apply
See all jobs
This is one of many
969,314 more open roles from verified company boards, updated every day.