Overview
Technical skills
Timeline
Roles

Overview

Data engineer with about 3 years of experience building and optimizing production-scale data pipelines and ETL/ELT solutions. Hands-on with Python, SQL, PySpark/Spark, Hadoop, Kafka, and AWS services, including Amazon S3 for large data processing and migration. Also works on dimensional modeling, data quality, and real-time streaming architectures.
Phone

Technical skills

Python
SQL
Java
Python
pySpark
DevOps
AWS
Amazon S3
Git
CI/CD
Jenkins
Rest API
AI/ML
Spark
Hadoop
LLM
Databases
Apache Iceberg
Apache Kafka
Analytics
ETL/ELT
Dimensional Modeling

Timeline

Data Engineer • Middle
Axis Bank • Full-Time
Nov 2023 to Present 2 Years 10 Months Mumbai In office
Develops and optimizes production data pipelines that process large-scale payment data from many heterogeneous sources for analytics and BI. Builds PySpark/Spark ETL/ELT workflows on Hadoop/cloud, improving pipeline turnaround time through performance tuning and data-skew handling. Designs dimensional warehouse models (star schema and SCD) and performs SQL optimization for validation and reporting. Implements Kafka-based streaming with Spark Structured Streaming and supports large data migrations with strict data integrity.
Python
SQL
pySpark
Spark
Hadoop
Apache Kafka
AWS
Amazon S3
ETL/ELT
Apache Iceberg
CI/CD
Git
Dimensional Modeling