1,301,603open jobs
75,803companies
208,113added this week
Browse all
Salary
≈ $42k – $113k per year (Estimated)
Location
In office (Cairo)
Seniority
Middle · 4+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 7, 2026. First seen by Alion on Oct 6, 2026. Capgemini scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Capgemini is a French information technology services and consulting group founded in Grenoble in 1967 by Serge Kampf, and one of the largest IT service providers in the world. It designs, builds and runs systems for large enterprises and governments across some fifty countries, with a delivery workforce concentrated in India and a European client base heavier than most of its competitors. The group is organised into Capgemini Invent for strategy and design, Capgemini Engineering for research and development services acquired with Altran, and Sogeti for local technology services, and it has reoriented its portfolio around cloud migration, data platforms and generative AI.

Job Description

We are seeking a highly skilled and motivated Senior Data Engineer to join our Data & Analytics practice. The successful candidate will build and operate enterprise-scale data solutions using PySpark, Apache Spark and modern data engineering technologies, with a strong grounding in data warehousing, lakehouse architecture and large-scale data integration.

The target environment uses Cloudera Data Platform (CDP). Prior Cloudera experience is preferred but is not a mandatory requirement. Candidates with strong PySpark/Spark expertise and relevant experience on comparable enterprise data platforms will be considered, and Cloudera CDP upskilling will be provided to the selected candidate.

This role requires hands-on engineering capability, technical ownership and effective collaboration with architects, analysts, source-system teams and client stakeholders. Experience in banking and financial services is highly desirable.

Critical hiring priority: Strong, production-grade PySpark/Spark engineering skills and the ability to rapidly learn and work effectively within the Cloudera CDP environment.

Key Responsibilities

Data Engineering & Development

Design, develop, test and maintain scalable batch and streaming data pipelines using PySpark.

Build and optimize ETL/ELT processes for large-volume enterprise data workloads.

Develop reusable ingestion, transformation and validation frameworks.

Implement data quality controls, reconciliation checks, monitoring and operational logging.

Support dimensional, warehouse, data lake and lakehouse data modeling activities.

Apply Spark performance optimization techniques including partitioning, join optimization, caching and handling data skew.

Cloudera Platform Engineering

Develop and run Spark workloads within the Cloudera Data Platform (CDP) environment.

Work with components including Cloudera Data Engineering (CDE), Cloudera Data Warehouse (CDW), Hive, Impala, Ozone, Ranger, Atlas and NiFi.

Configure, execute, monitor and troubleshoot Spark workloads.

Performance-tune PySpark jobs, Spark configurations, SQL queries and Hive/Impala workloads.

Diagnose platform, ingestion, transformation, storage and workload execution issues.

Apply security, access-control, metadata and lineage practices using Ranger and Atlas.

Candidates without prior Cloudera experience will be expected to complete the required Cloudera enablement/upskilling and apply their existing Spark and data engineering knowledge to the platform.

Architecture, Governance & Delivery

Contribute to lakehouse implementations using modern table formats such as Apache Iceberg.

Apply engineering standards covering modular design, code quality, testing, version control and CI/CD.

Participate in requirements analysis, solution design, technical estimation and design reviews.

Support system integration testing, UAT, production deployment and operational readiness.

Produce clear technical documentation and deliver structured knowledge-transfer sessions.

Collaborate effectively with distributed, multicultural teams and technical and business stakeholders.

Candidate Profile

Education & Experience

Bachelor's degree in Computer Science, Information Systems, Engineering or a related discipline.

At least 4 years of professional experience in data engineering.

At least 3 years of hands-on PySpark / Apache Spark development experience.

Experience delivering large-scale enterprise data pipelines and ETL/ELT solutions.

Cloudera CDP experience is an advantage, but not mandatory. Strong candidates from comparable Spark-based data platforms are encouraged to apply.

Mandatory Technical Skills

PySpark

Apache Spark

Python

Advanced SQL and query optimization

ETL/ELT development

Data warehousing concepts

Linux

Git

Data modeling fundamentals

Experience working with large-scale enterprise data platforms

Preferred / Upskillable Skills

Cloudera CDP

Cloudera CDE and CDW

Hive and Impala

Ranger and Atlas

Apache Iceberg

Apache NiFi and Kafka

Apache Airflow

Denodo and Informatica

Oracle and Teradata

Azure data services

DevOps and CI/CD pipelines

Cloudera CDP knowledge is preferred, not mandatory. The selected candidate will be provided with Cloudera enablement/upskilling where required.

Preferred Domain Background

Banking and financial services

Digital banking platforms

Customer analytics

Regulatory reporting

Enterprise data warehousing

Data lakehouse implementation and migration programs

Behavioural Competencies

Strong analytical thinking and structured problem solving.

Ownership mindset with the ability to work independently and deliver reliably.

Clear written and verbal communication with technical and non-technical stakeholders.

Ability to mentor junior engineers and contribute to team capability building.

Ability and willingness to quickly learn new data platforms and technologies.

Comfort working in distributed and multicultural delivery teams.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,301,603 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
Cairo
≈ $80k – $180k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Cairo
Python
Java
SQL
Scala
Databases
Snowflake
Databricks
Apache Iceberg
Delta Lake
Cassandra
DynamoDB
LookML
MS SQL
Apache Kafka
Google BigQuery
Amazon Redshift
Azure Cosmos DB
Apache Hudi
Trino
Teradata
BigQuery
AI/ML
Cursor
LangGraph
AutoGen
LangChain
Spark
Airflow
Claude Code
LlamaIndex
Model Context Protocol
MLFlow
Vertex AI
dbt
AI Agents
AWS Bedrock
Great Expectations
CrewAI
LLM
Amazon SageMaker
OpenAI Codex
DevOps
GCP
Azure DevOps
GitHub Actions
GitLab CI
Azure
CI/CD
AWS
Docker
Kubernetes
AWS Lambda
Amazon S3
IAM
Amazon Kinesis
AWS Step Functions
Apache HTTP Server
Cybersecurity
HashiCorp Vault
Game Dev
Unity
Chips/EDA
PoC Library
Analytics
Tableau
Power BI
ETL/ELT
Informatica
Talend
SSIS
Azure Data Factory
AWS Glue
Looker
SSAS
Dimensional Modeling
Collibra
Management
n8n
Apply
Lead Data Engineer 1 day ago
≈ $69k – $155k per year (Estimated) • In office • Full-Time • 10+ years exp • Cairo
Python
SQL
Python
pySpark
Databases
Apache Iceberg
Apache Kafka
AI/ML
Spark
Airflow
Flink
DevOps
CI/CD
Analytics
ETL/ELT
Informatica
Apache NiFi
Apply
≈ $80k – $180k per year (Estimated) • In office • Full-Time • Bachelor's Degree • New Cairo
SQL
Databases
Snowflake
Databricks
Apache Kafka
Teradata
Microsoft Fabric
DevOps
CI/CD
Analytics
Tableau
Power BI
ETL/ELT
Informatica
Apache NiFi
Apply
In office • Bengaluru
Python
Java
SQL
C#
C++
DevOps
AWS
Apply
$5.2k – $13k per year • In office • Full-Time • Kochi
SQL
Databases
MS SQL
Apply
Lead Data Engineer 1 day ago
≈ $69k – $155k per year (Estimated) • In office • Full-Time • 10+ years exp • Cairo
Python
SQL
Python
pySpark
Databases
Apache Iceberg
Apache Kafka
AI/ML
Spark
Airflow
Flink
DevOps
CI/CD
Analytics
ETL/ELT
Informatica
Apache NiFi
Apply
≈ $17k – $35k per year (Estimated) • Hybrid • Full-Time • 7+ years exp • Bachelor's Degree • Pune
Python
SQL
Scala
Python
pySpark
Databases
Databricks
MS SQL
Apache Kafka
Azure SQL Database
AI/ML
Hadoop
Spark
Machine Learning
DevOps
Azure
CI/CD
Git
Analytics
Power BI
Azure Data Factory
Management
Agile
Scrum
Apply
≈ $20k – $56k per year (Estimated) • In office • Bengaluru
Python
Python
pySpark
AI/ML
LangGraph
LangChain
Spark
Model Context Protocol
MLFlow
Prompt Engineering
AI Agents
RAG
Google ADK
Agentic Workflows
Multi-Agent Systems
Management
Agile
Apply
In office • Welwyn Garden City
Python
Python
pySpark
AI/ML
Hadoop
Spark
Time Series Forecasting
DevOps
CI/CD
Apply
≈ $89k – $171k per year (Estimated) • In office • 3+ years exp • Dallas
Python
SQL
Databases
Snowflake
Databricks
Google BigQuery
BigQuery
AI/ML
Spark
Airflow
dbt
DevOps
GCP
Azure
CI/CD
Git
AWS
IAM
Analytics
Power BI
ETL/ELT
Dimensional Modeling
Management
Agile
Scrum
Apply
Lead Data Engineer 1 day ago
≈ $69k – $155k per year (Estimated) • In office • Full-Time • 10+ years exp • Cairo
Python
SQL
Python
pySpark
Databases
Apache Iceberg
Apache Kafka
AI/ML
Spark
Airflow
Flink
DevOps
CI/CD
Analytics
ETL/ELT
Informatica
Apache NiFi
Apply
≈ $69k – $137k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Birmingham
SQL
Apex
Apex
MuleSoft
Salesforce Flow
Databases
Snowflake
Databricks
AI/ML
Agentforce
DevOps
Rest API
GCP
Azure
AWS
SOAP
Analytics
Tableau
ETL/ELT
Apply
≈ $33k – $69k per year (Estimated) • Hybrid • Full-Time • Lisbon
SQL
Databases
MS SQL
DevOps
Windows Server
Linux
Apply
≈ $27k – $67k per year (Estimated) • In office • Full-Time • 1+ year exp • Bydgoszcz
Python
SQL
DevOps
IAM
Cybersecurity
Zero Trust
SIEM
Apply
$160k – $170k per year • In office • Full-Time • Atlanta
Python
SQL
Bash
Databases
Google BigQuery
BigQuery
DevOps
Terraform
GCP
Helm
GitHub Actions
Prometheus
CI/CD
Kubernetes
Grafana
Google GKE
Google Cloud Run
FinOps
SLI/SLO/SLA
GitLab
IAM
Linux
Apply
Lead Data Engineer 1 day ago
≈ $69k – $155k per year (Estimated) • In office • Full-Time • 10+ years exp • Cairo
Python
SQL
Python
pySpark
Databases
Apache Iceberg
Apache Kafka
AI/ML
Spark
Airflow
Flink
DevOps
CI/CD
Analytics
ETL/ELT
Informatica
Apache NiFi
Apply
≈ $57k – $145k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Cairo
SQL
PowerShell
C#
C#
.NET
DevOps
Rest API
IAM
SOAP
Cybersecurity
Microsoft Entra ID
Active Directory
LDAP
Apply
≈ $27k – $57k per year (Estimated) • In office • Full-Time • Cairo
Management
Outlook
Microsoft Office
Apply
≈ $41k – $100k per year (Estimated) • In office • Full-Time • Cairo
Design
Figma
Adobe XD
Axure RP
Balsamiq
Management
Confluence
Agile
UML
BPMN
Apply
≈ $25k – $52k per year (Estimated) • In office • Full-Time • 1+ year exp • Bachelor's Degree • Cairo
Apply
See all jobs
This is one of many
1,301,603 more open roles from verified company boards, updated every day.