430,878open jobs
14,717companies
59,947added this week
Browse all
Salary
$24k – $50k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Senior · 3+ years exp
Overview
Company
Impact
Profile match
LogixHealth is a hospital & healthcare organization that provides services related to health care using BI tools, specializing in full-service coding, billing and revenue cycle solutions.

We're looking for a Senior Data Engineer with 3+ years of hands-on experience in Apache Spark and Databricks to design and build scalable data pipelines and distributed data processing solutions on Azure. If you're strong in PySpark, Spark optimization, Delta Lake, Databricks, and modern data engineering, and enjoy working with large-scale data, this role is for you.

Responsibilities:

  • Design, develop, and maintain high-performance batch and streaming data pipelines using Databricks, Apache Spark, PySpark, SQL, Python, and Delta Lake.
  • Build scalable data solutions within the Azure ecosystem, leveraging services such as Azure Blob Storage, Azure Data Factory, and Event Hubs.
  • Implement Medallion Architecture (Bronze/Silver/Gold) and robust ETL/ELT frameworks.
  • Develop incremental processing and CDC pipelines for large-scale datasets.
  • Design and optimize Spark workloads handling TB+ datasets.
  • Apply Spark performance optimization techniques, including partitioning and repartitioning, efficient joins and broadcast joins, caching and persistence, adaptive query execution (AQE), data skew handling, Query and job optimization.
  • Build and manage Delta Lake solutions, including ACID transactions, schema evolution, optimization, and reliable data pipelines.
  • Develop Delta Live Tables (DLT) pipelines with strong data quality and pipeline design practices.
  • Use Databricks Jobs and Workflows for production-grade orchestration.
  • Implement data governance, access control, security, and lineage using Unity Catalog and RBAC.
  • Establish strong practices around data quality, testing, monitoring, maintainability, and reusability.
  • Integrate the data platform with Power BI, Tableau, APIs, and custom applications.
  • Contribute to CI/CD, infrastructure-as-code, automated testing, monitoring, and alerting using tools such as Git, Terraform, Azure DevOps, GitHub Actions, Jenkins, Azure Monitor, or Datadog.
  • Work closely with engineering, product, and business teams to deliver reliable data solutions that support reporting and analytics.
  • Contribute to the development of a self-service data platform that enables teams to discover, access, and use trusted data efficiently.

Requirements:

  • 3+ years of strong hands-on experience with Apache Spark and Databricks.
  • Expert-level knowledge of PySpark / Spark, with Python or Scala.
  • Strong understanding of Spark DataFrames, Spark SQL, Structured Streaming, Distributed Data Processing, and Large-Scale Data Engineering.
  • Deep Databricks experience across Delta Lake, Delta Live Tables (DLT), Unity Catalog, Databricks Jobs & Workflows, and Databricks Notebooks.
  • Proven experience designing and tuning high-performance Spark jobs.
  • Hands-on experience working with large-scale datasets, preferably TB+.
  • Strong understanding of batch and streaming architectures.
  • Experience implementing Medallion Architecture.
  • Practical experience with incremental processing and CDC patterns.
  • Strong SQL and data engineering fundamentals.
  • Experience working with Azure cloud services, particularly Azure Blob Storage.

Good to Have:

  • Experience integrating Databricks with Azure Data Factory.
  • Experience with Azure Event Hubs and streaming pipelines.
  • Experience with external orchestration tools such as Apache Airflow.
  • Experience with NoSQL databases.
  • Experience with MS SQL, PostgreSQL, or MySQL.
  • Knowledge of data governance, security, compliance, RBAC, and data lineage.
  • Experience with Terraform / Infrastructure as Code.
  • Experience with Azure DevOps, GitHub Actions, or Jenkins.
  • Experience with Azure Monitor or Datadog.
  • Exposure to Power BI, Tableau, APIs, or custom applications.

You'll be expected to:

  • Build for scale: Design Spark pipelines that perform reliably across TB+ datasets.
  • Build for reliability: Create pipelines with strong data quality, testing, monitoring, and failure-handling practices.
  • Build for change: Handle schema evolution, incremental loads, CDC, and continuously changing data sources.
  • Build for governance: Use Unity Catalog, RBAC, lineage, and security best practices to make data discoverable and trustworthy.
  • Build for reuse: Create maintainable frameworks and components that can be used across multiple data products.
  • Build for production: Follow CI/CD, Infrastructure as Code, deployment, observability, and operational best practices.

Technology Stack:

  • Cloud: Azure.
  • Data Platform: Databricks.
  • Processing: Apache Spark, PySpark, Spark SQL, Structured Streaming.
  • Storage: Delta Lake, Azure Blob Storage.
  • Pipelines: DLT, Databricks Jobs, and Workflows.
  • Governance: Unity Catalog, RBAC, Data Lineage.
  • Integration: Azure Data Factory, Event Hubs, Airflow.
  • Databases: SQL Server, PostgreSQL, MySQL, NoSQL.
  • DevOps: Git, Terraform, Azure DevOps, GitHub Actions, Jenkins.
  • Monitoring: Azure Monitor, Datadog.
  • Analytics: Power BI, Tableau.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
430,878 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
In office • Full-Time • Singapore
Python
SQL
AI/ML
Copilot
Claude
Analytics
Tableau
Power BI
Microsoft Excel
Apply
$40k – $106k per year (Estimated) • In office • Full-Time • Master's Degree • Taichung
Python
JavaScript
SQL
Python
pySpark
Databases
Snowflake
AI/ML
Spark
AI Agents
DevOps
GCP
Analytics
Tableau
Apply
$120k – $286k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Singapore
Python
JavaScript
TypeScript
SQL
Python
FastAPI
AI/ML
Quantization
Prompt Engineering
Function Calling
AI Agents
LLM
RAG
Streamlit
Anomaly Detection
Gaussian Splatting
LLMOps
Agentic Workflows
Tool Use
Frontend
Angular
React.js
DevOps
GCP
OpenShift
Azure
CI/CD
AWS
Docker
Kubernetes
Game Dev
Unity
Unreal Engine
Robotics
ROS
Gazebo
NVIDIA Omniverse
Isaac Sim
SLAM
Sim-to-Real
Imitation Learning
Digital Twin
Apply
AI Engineer, SMAI 1 hour ago
In office • Full-Time • 2+ years exp • Bachelor's Degree • Singapore
Python
JavaScript
TypeScript
SQL
Python
FastAPI
AI/ML
Prompt Engineering
AI Agents
LLM
RAG
Streamlit
Frontend
Angular
React.js
DevOps
GCP
OpenShift
GitHub Actions
Azure
CI/CD
AWS
Docker
Kubernetes
GitHub
Apply
$30k – $72k per year (Estimated) • Remote • Bucharest
Python
Java
SQL
PowerShell
C#
C#
.NET
Databases
MySQL
PostgreSQL
MS SQL
DevOps
GCP
Azure DevOps
Datadog
Prometheus
Azure
CI/CD
Jenkins
AWS
Docker
Grafana
Incident Management
Apply
$13k – $34k per year (Estimated) • In office • Full-Time • Bengaluru
DevOps
SLI/SLO/SLA
Apply
$27k – $63k per year (Estimated) • In office • Full-Time • 5+ years exp • Pune • Bengaluru • Coimbatore • Gurgaon • Noida
JavaScript
AI/ML
Copilot
Claude
Prompt Engineering
DevOps
CI/CD
GitHub
Apply
$40k – $87k per year (Estimated) • In office • Full-Time • 12+ years exp • Gurgaon • Bengaluru • Mumbai
Python
Java
AI/ML
LangGraph
LangChain
AI Agents
Agentic Workflows
Apply
In office • Full-Time • Bengaluru
Java
Java
Spring Boot
Databases
MySQL
PostgreSQL
DevOps
Git
Apply
$29k – $74k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
ABAP
Databases
SAP HANA
Apply
See all jobs
This is one of many
430,878 more open roles from verified company boards, updated every day.