953,927open jobs
57,674companies
156,942added this week
Browse all
Salary
≈ $31k – $55k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Principal · 12+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 30, 2026. First seen by Alion on Sep 7, 2026. Takeda Pharmaceutical scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Takeda Pharmaceutical Company Limited is a global biopharmaceutical enterprise specializing in the research, development, and commercialization of targeted human therapeutics. The company focuses on core therapeutic areas including oncology, rare diseases, neuroscience, gastroenterology, plasma-derived therapies, and vaccines. Headquartered in Tokyo, Japan, it maintains extensive research, manufacturing, and commercial operations across more than 80 countries in Asia-Pacific, North America, Europe, and Latin America.

By clicking the “Apply” button, I understand that my employment application process with Takeda will commence and that the information I provide in my application will be processed in line with Takeda’s Privacy Notice and Terms of Use. I further attest that all information I submit in my employment application is true to the best of my knowledge.

Job Description

Principal Data Engineer

About the Role

We are seeking a highly experienced Principal Data Engineer to lead the design, architecture, and implementation of enterprise-scale data platforms that power Takeda's R&D, Clinical, Regulatory, Commercial, and Enterprise Analytics ecosystems.

As a technical leader, you will drive the strategic direction of data engineering, define architecture standards, and deliver scalable, secure, and compliant data solutions using Databricks, AWS, PySpark, Delta Lake, and modern DataOps practices. You will partner with business stakeholders, solution architects, product owners, data scientists, and engineering teams to build reliable data products that accelerate innovation and data-driven decision making across Takeda.

This role requires deep expertise in modern cloud data platforms, distributed data processing, software engineering practices, and cross-functional leadership. You will mentor engineering teams, establish best practices, and ensure enterprise data solutions meet quality, scalability, security, and regulatory requirements.

Key Responsibilities

Data Architecture & Platform Leadership

  • Lead architecture, design, and implementation of enterprise data platforms on Databricks and AWS.
  • Define and govern enterprise data engineering standards, reference architectures, and reusable frameworks.
  • Design scalable Lakehouse architectures using Delta Lake, Unity Catalog, Databricks Workflows, and AWS cloud services.
  • Collaborate with enterprise architects and business stakeholders to translate business requirements into scalable technical solutions.
  • Drive platform modernization initiatives, cloud migration programs, and data transformation roadmaps.
  • Establish best practices for data modeling, metadata management, data lineage, and governance.
  • Evaluate emerging technologies and recommend improvements to Takeda's data ecosystem.

Data Engineering & Solution Delivery

  • Design and build high-performance batch and streaming (Optional) data pipelines using PySpark, Spark SQL, and Databricks.
  • Architect ingestion frameworks supporting structured, semi-structured, and unstructured data from internal and external systems.
  • Lead implementation of medallion architecture patterns (Bronze, Silver, Gold) to support trusted enterprise data products.
  • Optimize large-scale data processing workloads to improve performance, reliability, and cost efficiency.
  • Design and implement reusable ETL/ELT frameworks and accelerator components.
  • Establish data contracts and engineering standards to ensure consistency and reliability across platforms.
  • Ensure data solutions are scalable, maintainable, and aligned with enterprise architecture principles.

Cloud Engineering & AWS Platform Management

  • Architect and implement cloud-native data solutions using AWS services including:
    • S3
    • IAM
    • Glue
    • Lambda
    • ECS/EKS
    • Step Functions
    • EventBridge
    • CloudWatch
    • Secrets Manager
    • KMS
    • Redshift
  • Design secure multi-account architectures and governance models.
  • Establish infrastructure automation practices using Terraform, CloudFormation, or AWS CDK.
  • Drive optimization of cloud resources through cost management, workload tuning, and automation.
  • Partner with cloud platform teams to ensure operational excellence and security compliance.

DataOps, CI/CD & Engineering Excellence

  • Lead adoption of software engineering best practices across data engineering teams.
  • Design and implement CI/CD pipelines for data platforms using GitHub Actions, DevOps, GitLab CI, or Jenkins.
  • Establish automated deployment frameworks across Development, Test, Validation, and Production environments.
  • Implement unit testing, integration testing, regression testing, and automated quality gates.
  • Drive code quality initiatives including:
    • Peer reviews
    • Static code analysis
    • Test automation
    • Release management
    • Version control strategies
  • Standardize engineering practices to improve delivery velocity and platform reliability.
  • Promote Infrastructure as Code and automated environment provisioning.

Data Quality, Governance & Compliance

  • Establish enterprise data quality frameworks and monitoring capabilities.
  • Implement end-to-end data lineage, metadata management, and observability solutions.
  • Collaborate with governance, security, quality, and compliance teams to ensure adherence to corporate standards.
  • Support implementation of:
    • Data governance policies
    • Role-based access controls
    • Data retention policies
    • Auditability requirements
  • Ensure compliance with:
    • GxP requirements
    • HIPAA
  • Promote secure handling of sensitive healthcare and research data.

Performance Optimization & Reliability Engineering

  • Establish platform observability using monitoring, logging, and alerting solutions.
  • Define service-level objectives and operational metrics for critical data platforms.
  • Lead root-cause analysis and resolution of complex production issues.
  • Implement resiliency, disaster recovery, and business continuity strategies.
  • Continuously improve platform performance, stability, and operational efficiency.

Leadership, Mentoring & Cross-Functional Collaboration

  • Provide technical leadership and guidance to data engineers across multiple programs and delivery teams.
  • Lead architectural reviews, design discussions, and engineering governance forums.
  • Mentor engineers in Databricks, Spark, AWS, DataOps, and software engineering best practices.
  • Partner with:
    • Product Owners
    • Business Stakeholders
    • Data Scientists
    • Cloud Engineering Teams
    • Security Teams
    • Quality and Compliance Teams
  • Influence strategic data platform decisions and long-term technology roadmaps.
  • Drive delivery excellence through collaboration, coaching, and continuous improvement.

Required Qualifications

  • Bachelor’s or master’s degree in computer science, Engineering, Information Systems, or related field.
  • 12+ years of experience in Data Engineering, Big Data, Data Platforms, or Cloud Engineering.
  • 4 to 5 years of experience architecting enterprise-scale data solutions on Databricks / AWS.
  • Deep expertise in:
    • Databricks
    • Delta Lake
    • Unity Catalog
    • Spark
    • PySpark
    • SQL
  • Strong AWS experience across compute, storage, security, and monitoring services.
  • Advanced Python development experience and software engineering practices.
  • Proven experience building large-scale enterprise data pipelines and Lakehouse architectures.
  • Expertise in CI/CD implementation and Git-based development practices.
  • Experience implementing automated unit testing and quality assurance frameworks.
  • Strong understanding of distributed systems and performance optimization.
  • Experience with Infrastructure as Code using Terraform or CloudFormation, or AWS CDK.
  • Excellent communication, stakeholder management, and leadership skills.

Preferred Qualifications

  • Life Sciences, Pharmaceutical, Healthcare, or Clinical data domain experience.
  • Experience supporting GxP-regulated environments.
  • Knowledge of:
    • Clinical Trial Data
    • Regulatory Data
    • Pharmacovigilance
    • Real World Data (RWD/RWE)
    • Omics and Research Data
  • Experience with ML/AI data platforms and MLOps foundations.
  • Exposure to streaming technologies such as Kafka, Kinesis, or Event Hubs.
  • Databricks Certified Data Engineer Professional.
  • AWS Solutions Architect Professional.
  • AWS Data Analytics Specialty or equivalent certifications.

What Success Looks Like (First 12 Months)

  • Data delivery timelines are significantly reduced through reusable frameworks, automation, and CI/CD.
  • Data pipelines achieve high reliability, observability, and operational excellence.
  • Engineering teams consistently follow software engineering, testing, and deployment best practices.
  • Cloud infrastructure and Databricks environments are optimized for performance, scalability, and cost.
  • Data products meet quality, governance, and compliance expectations across R&D and enterprise functions.
  • Multiple teams successfully adopt reusable engineering accelerators developed under your leadership.
  • Stakeholders recognize the data platform as a strategic enabler for innovation and business outcomes.

Locations

IND - Bengaluru

Worker Type

Employee

Worker Sub-Type

Regular

Time Type

Full time
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
953,927 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
Bengaluru
Data Architect I 5 hours ago
≈ $34k – $68k per year (Estimated) • Hybrid • 7+ years exp • Master's Degree • Bengaluru
Python
SQL
Databases
Snowflake
AI/ML
LLM
DevOps
CI/CD
Analytics
Power BI
ETL/ELT
Apply
≈ $28k – $50k per year (Estimated) • In office • Full-Time • Bengaluru
Python
SQL
Databases
Snowflake
MS SQL
Apache Kafka
AI/ML
Spark
Embeddings
AI Agents
LLM
RAG
Semantic Search
Semantic Search
Knowledge Graph
Agentic Workflows
Machine Learning
DevOps
Kong
CI/CD
Git
AWS
Analytics
Dimensional Modeling
Apply
≈ $16k – $29k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Gurgaon
Databases
Snowflake
Databricks
Oracle
AI/ML
Spark
AI Agents
Edge AI
DevOps
Azure
Analytics
ETL/ELT
Management
Agile
Apply
≈ $28k – $50k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Databases
Snowflake
Apache Kafka
Amazon Redshift
AI/ML
Streamlit
DevOps
AWS
AWS Lambda
Amazon S3
Amazon Kinesis
Cybersecurity
PCI DSS
SOC 2
GDPR
HIPAA
Analytics
Tableau
Power BI
Looker
Data Vault
Apply
≈ $30k – $60k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
ABAP
Python
pySpark
ABAP
SAP BTP
Databases
Databricks
Delta Lake
Apache Kafka
AI/ML
Spark
Embeddings
AI Agents
LLM
RAG
Tokenization
Time Series Forecasting
Feature Store
Knowledge Graph
Machine Learning
DevOps
Azure
Analytics
Power BI
ETL/ELT
Management
Agile
Apply
≈ $23k – $69k per year (Estimated) • In office • Bachelor's Degree • Bari
Python
SQL
SAS
Python
pySpark
Databases
MySQL
PostgreSQL
Oracle
HBase
Cassandra
MS SQL
Azure Cosmos DB
AI/ML
Hadoop
Spark
Scikit-learn
TensorFlow
NumPy
Machine Learning
DevOps
AWS
Analytics
Tableau
Power BI
Seaborn
Matplotlib
ETL/ELT
Informatica
Talend
DataStage
MicroStrategy
SAP BusinessObjects
Apply
≈ $134k – $345k per year (Estimated) • Hybrid • Bromley
SQL
C#
DevOps
Rest API
GitHub Actions
CI/CD
AWS
API Gateway
SOAP
Cybersecurity
OWASP Top 10
Apply
Backend Engineer 4 days ago
≈ $44k – $120k per year (Estimated) • Equity • In office • Full-Time • 5+ years exp • Buenos Aires
JavaScript
TypeScript
SQL
Node JS
Node JS
Nest.JS
Frontend
tRPC
Mobile
Clean Architecture
DevOps
AWS
Docker
Kubernetes
Apply
≈ $25k – $62k per year (Estimated) • Equity • In office • 8+ years exp • Bachelor's Degree • Southfield
Python
Apply
≈ $14k – $35k per year (Estimated) • Hybrid • Cebu City
SQL
Apply
≈ $19k – $44k per year (Estimated) • In office • Full-Time • 3+ years exp • PhD • Bengaluru
Python
SQL
SAS
AI/ML
Machine Learning
DevOps
GCP
Apply
≈ $25k – $51k per year (Estimated) • In office • Full-Time • 7+ years exp • PhD • Bengaluru
Python
SQL
SAS
AI/ML
Machine Learning
DevOps
GCP
Apply
≈ $25k – $63k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • PhD • Bengaluru
Python
SQL
SAS
AI/ML
Machine Learning
DevOps
GCP
Unix
Apply
≈ $20k – $53k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Bengaluru
SQL
Databases
PostgreSQL
DynamoDB
Apache Kafka
DevOps
Terraform
GitHub Actions
CI/CD
Jenkins
Git
AWS
GitHub
GitLab
SOAP
Analytics
Dimensional Modeling
Master Data Management
Management
Confluence
Agile
Scrum
Apply
≈ $18k – $42k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
Apache Kafka
AI/ML
Spark
MLFlow
AWS Bedrock
LLM
RAG
Amazon SageMaker
Feature Store
LLM Guardrails
Machine Learning
DevOps
Terraform
GitHub Actions
CloudFormation
GitLab CI
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Platform Engineering
Amazon EKS
AWS Lambda
SLI/SLO/SLA
Amazon S3
IAM
Amazon ECS
Amazon CloudWatch
Amazon Kinesis
AWS Step Functions
Cybersecurity
GDPR
HIPAA
Apply
≈ $19k – $39k per year (Estimated) • Remote (India) • Full-Time • 4+ years exp • Bengaluru
Python
JavaScript
SQL
Databases
Snowflake
Databricks
AI/ML
Spark
dbt
Frontend
React.js
Apply
≈ $13k – $30k per year (Estimated) • Remote (India) • Full-Time • Bengaluru • Hyderabad • Pune • Greater Noida
SQL
Databases
PostgreSQL
Db2
Oracle
Cassandra
MS SQL
DevOps
Ansible
Dynatrace
Azure
Git
Configuration Management
Cybersecurity
Active Directory
LDAP
Management
ServiceNow
ITSM
Apply
Data Engineer 1 day ago
≈ $20k – $49k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
AI/ML
Airflow
Machine Learning
DevOps
GCP
CI/CD
Git
AWS
Analytics
ETL/ELT
Management
Agile
Apply
≈ $23k – $64k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Bengaluru
Python
SAS
SPSS
AI/ML
Machine Learning
Analytics
ETL/ELT
Informatica
Talend
Apply
In office • Full-Time • 3+ years exp • Bengaluru
ABAP
ABAP
SAP Fiori
SAP Gateway
Apply
See all jobs
This is one of many
953,927 more open roles from verified company boards, updated every day.