803,192open jobs
51,523companies
126,647added this week
Browse all
Salary
$35k – $42k per year
Location
In office (Noida)
Seniority
Staff · 10+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 26, 2026. First seen by Alion on Sep 24, 2026.

Overview
Company
Impact
Profile match
Algoworks is an artificial intelligence, engineering services and experience transformation firm that designs, builds and runs software, data and CRM platforms for enterprise clients. The company has offices across the United States, Europe, South America and India, has worked with Fortune 500 organizations for more than 20 years, and runs its main delivery centre in Noida with further teams in Gurugram and Hyderabad. Its current openings are mostly in Noida and cover ServiceNow, Salesforce and Microsoft Dynamics specialists, data and backend engineers in Python and Java, DevOps and mobile QA engineers, solution architects and delivery managers.

Role: Lead Data Engineer

Location: India, Remote

Experience: 10+ Years

Algoworks

www.algoworks.com

About the company

Algoworks is an award-winning artificial intelligence, engineering services and experience transformation firm with offices across the United States, Europe, South America and India. We bring together a global team of engineers, architects, designers, researchers and operators united by rigor, accountability and a commitment to delivering measurable results.

For over 20 years, Algoworks has partnered with Fortune 500 organizations across the Americas, Europe and Asia to define, build and run technology that drives meaningful business outcomes. Our work combines human-centered design, engineering excellence and AI-powered capabilities to solve complex challenges with clarity and precision. Innovation, particularly in the responsible application of AI, is embedded in how teams approach problem-solving and continuous improvement.

At Algoworks, growth is continuous and closely tied to impact. Teams collaborate across geographies and disciplines, strengthening outcomes through shared insight and collective expertise. The culture values transparency, open dialogue and an environment where every voice is heard and contribution is recognized.

Through collaboration, accountability and a focus on results, Algoworks operates at the intersection of technology and people, building not only advanced systems but strong global teams that elevate performance and create lasting impact.

Follow the video below to know about us! Clipchamp

Role overview

We are seeking a hands-on Senior Data Engineer with strong expertise in Azure Databricks and Azure Data Factory to build, optimize, and maintain scalable enterprise data pipelines.

The role will focus on high-performance ETL/ELT development, Delta Lake optimization, data processing across Bronze, Silver, and curated layers, and close collaboration with DWH and reporting teams for downstream consumption.

The ideal candidate will act as a senior technical contributor, ensuring reliability, performance, data quality, and maintainability across the data platform.

Key responsibilities:

1.Pipeline Development

  • Build and maintain scalable data pipelines using Azure Databricks and Azure Data Factory.
  • Implement ingestion and transformation logic across Bronze and Silver data layers.
  • Develop batch and incremental data-processing patterns.
  • Design reliable and reusable pipeline components for enterprise workloads.
  • Monitor and troubleshoot pipeline execution and data-processing issues.

2.Curated Layer & Delta Lake Development

  • Implement hydration, merge, and upsert logic using Delta Lake.
  • Build and maintain curated datasets aligned with data quality and business requirements.
  • Handle late-arriving data and incremental updates.
  • Implement reliable data transformation and reconciliation processes.
  • Ensure curated datasets are optimized for downstream consumption.

3.Performance & Storage Optimization

  • Optimize Delta Lake tables for performance and cost efficiency.
  • Select and tune appropriate storage formats such as Parquet and Delta.
  • Apply partitioning, compaction, and file-sizing strategies.
  • Tune Spark jobs for large-scale distributed data processing.
  • Identify and resolve performance bottlenecks across data pipelines and storage layers.

4.Downstream & DWH Collaboration

  • Work closely with DWH and reporting teams to support downstream data consumption.
  • Provide optimized datasets for reporting and analytical workloads.
  • Support data validation and reconciliation with Gold-layer outputs.
  • Collaborate with downstream teams to understand data requirements and optimize delivery.
  • Ensure consistency and reliability of data consumed by reporting and analytics platforms.

5.Engineering Best Practices

  • Implement basic CI/CD practices for data pipelines.
  • Follow coding standards, documentation, and version-control practices.
  • Maintain reusable, scalable, and maintainable pipeline code.
  • Support production troubleshooting and performance tuning.
  • Participate in Agile delivery processes and technical discussions.

6.Data Quality & Production Support

  • Implement data validation and quality checks across ingestion and transformation processes.
  • Investigate data discrepancies and pipeline failures.
  • Perform root-cause analysis and implement corrective actions.
  • Support production deployments and resolve data-processing issues.
  • Maintain reliability and consistency across enterprise data pipelines.

Required technical skills and competencies:

  • Strong hands-on experience in data engineering and enterprise data platforms.
  • Strong experience building data pipelines on Azure.
  • Advanced proficiency in PySpark.
  • Hands-on experience with Azure Databricks.
  • Strong experience with Azure Data Factory.
  • Deep knowledge of Delta Lake tuning and optimization.
  • Strong understanding of storage optimization using Parquet and Delta.
  • Strong SQL skills for data transformation, validation, and reconciliation.
  • Experience working with large datasets and distributed processing.
  • Experience implementing batch and incremental processing patterns.
  • Experience with hydration, merge, and upsert logic.
  • Experience with Git and basic CI/CD pipelines.
  • Familiarity with data quality and validation techniques.
  • Experience working in Agile delivery environments.

Must have skills:

  • Azure Databricks.
  • Azure Data Factory.
  • PySpark.
  • Delta Lake.
  • SQL.
  • Data Engineering.
  • ETL/ELT.
  • Data Pipeline Development.
  • Bronze/Silver/Curated Data Layers.
  • Delta Lake Performance Optimization.
  • Spark Performance Tuning.
  • Parquet.
  • Batch and Incremental Processing.
  • Git and Version Control.
  • Strong analytical and problem-solving skills.

Good to have skills:

  • Microsoft Fabric.
  • Streaming or near real-time data pipelines.
  • Data governance tools.
  • Metadata management tools.
  • Advanced CI/CD practices.
  • Experience with Gold-layer development.
  • Experience supporting enterprise reporting and DWH platforms.

Desired attributes:

  • Strong analytical and problem-solving capabilities.
  • Ability to independently design and develop complex data pipelines.
  • Strong focus on performance, scalability, reliability, and data quality.
  • Ability to troubleshoot complex production data issues.
  • Strong attention to detail and coding discipline.
  • Good communication and collaboration skills.
  • Ability to work effectively with DWH, reporting, and cross-functional teams.
  • Proactive approach to performance optimization and continuous improvement.
  • Strong ownership of data engineering deliverables.

Interview Process

2-3 rounds of discussion.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
803,192 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
Noida
≈ $36k – $78k per year (Estimated) • Remote (India) • Full-Time • Bengaluru • Hyderabad • Pune • Chennai • Greater Noida
DevOps
Incident Management
SLI/SLO/SLA
Cybersecurity
Least Privilege
Analytics
Power BI
Apply
Staff Data Architect 2 hours ago
≈ $32k – $75k per year (Estimated) • In office • Full-Time • 7+ years exp • Bengaluru
SQL
Databases
Google BigQuery
BigQuery
DevOps
GCP
CI/CD
AWS
Incident Management
Analytics
ETL/ELT
Collibra
Apply
≈ $40k – $93k per year (Estimated) • Remote (India) • 9+ years exp
Python
SQL
Databases
Snowflake
AI/ML
Claude
dbt
Apply
≈ $39k – $90k per year (Estimated) • Remote (India) • 9+ years exp
Python
SQL
Databases
Snowflake
AI/ML
Claude
dbt
Management
WhatsApp
Apply
≈ $36k – $78k per year (Estimated) • Hybrid • 11+ years exp • Bengaluru
Python
SQL
C++
Analytics
Tableau
Power BI
Cognos
Erwin
Apply
≈ $93k – $197k per year (Estimated) • Remote (likely India) • 5+ years exp
Python
SQL
Scala
Databases
Databricks
MS SQL
Apache Kafka
Microsoft Fabric
AI/ML
Spark
DevOps
Azure
Analytics
Power BI
Dimensional Modeling
Apply
≈ $66k – $140k per year (Estimated) • Remote (likely United States) • 10+ years exp • Bachelor's Degree
SQL
DevOps
Azure DevOps
Azure
Analytics
Power BI
Management
Power Automate
Power Apps
Apply
≈ $103k – $207k per year (Estimated) • Remote (likely United States) • 4+ years exp • Bachelor's Degree
SQL
DevOps
Azure DevOps
Azure
Analytics
Power BI
Management
Power Automate
Apply
Growth Lead - Jar 1 day ago
In office • 7+ years exp • Bachelor's Degree
SQL
Marketing
Amplitude
Mixpanel
Apply
$100k – $127k per year • In office • Public Trust • 3+ years exp
JavaScript
TypeScript
DevOps
Azure DevOps
Azure
CI/CD
Git
Management
Agile
QA
Playwright
Apply
$34k – $37k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Noida
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
Apache Kafka
AI/ML
Spark
DevOps
Terraform
Azure
CI/CD
Bicep
Analytics
Azure Data Factory
Apply
$31k – $39k per year • In office • Full-Time • 10+ years exp • Noida
Python
SQL
Databases
RabbitMQ
DevOps
Rest API
CI/CD
AWS
Apply
Data Engineer 1 month ago
$31k – $37k per year • In office • Full-Time • 6+ years exp • Noida
Python
SQL
Python
pySpark
Databases
Snowflake
AI/ML
Spark
Dagster
dbt
Frontend
GraphQL
DevOps
CI/CD
Apply
$10k – $31k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
Apache Kafka
AI/ML
Spark
DevOps
Terraform
Azure
CI/CD
Bicep
Analytics
Azure Data Factory
Apply
Senior Engineer 1 day ago
$19k – $26k per year • In office • Full-Time • 5+ years exp • Noida
Apex
DevOps
Datadog
CI/CD
Incident Management
SLI/SLO/SLA
Management
ITSM
Apply
≈ $28k – $59k per year (Estimated) • In office • Full-Time • Noida
AI/ML
Machine Learning
Cybersecurity
ISO 27001
SOC 2
GDPR
HIPAA
Apply
≈ $15k – $31k per year (Estimated) • Remote (India) • Full-Time • 10+ years exp • Bachelor's Degree • Noida
Python
Java
SQL
Scala
Databases
Snowflake
Databricks
Delta Lake
Apache Kafka
AI/ML
Spark
AI Agents
Flink
Edge AI
DevOps
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Management
Agile
Scrum
Apply
≈ $39k – $88k per year (Estimated) • In office • 10+ years exp • Noida • Bengaluru • Chennai • Thiruvananthapuram • Kochi
Python
SQL
Python
pySpark
Databases
Azure SQL Database
Microsoft Fabric
AI/ML
Spark
dbt
DevOps
Azure
Analytics
Power BI
Azure Data Factory
Apply
≈ $25k – $62k per year (Estimated) • In office • Noida
Python
Go
Java
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Cybersecurity
GDPR
HIPAA
Apply
In office • Full-Time • Master's Degree • Noida
Python
C++
AI/ML
Prompt Engineering
NLP
LLM
RAG
Semantic Search
OpenAI
Anthropic
Semantic Search
DevOps
CI/CD
Git
Docker
Kubernetes
Management
Agile
Apply
See all jobs
This is one of many
803,192 more open roles from verified company boards, updated every day.