Data Scientist
Python
SQL
Java
Active 4 days ago
Invite to interview
Download CVCV
Overview
Technical skills
Timeline
Roles
Overview
Data engineer with about 3 years of experience building and optimizing production-scale data pipelines and ETL/ELT solutions. Hands-on with Python, SQL, PySpark/Spark, Hadoop, Kafka, and AWS services, including Amazon S3 for large data processing and migration. Also works on dimensional modeling, data quality, and real-time streaming architectures.
Phone
Technical skills
Python
SQL
Java
Python
pySpark
DevOps
AWS
Amazon S3
Git
CI/CD
Jenkins
Rest API
AI/ML
Spark
Hadoop
LLM
Databases
Apache Iceberg
Apache Kafka
Analytics
ETL/ELT
Dimensional Modeling
Timeline
Data Engineer
•
Middle
Axis Bank
•
Full-Time
Develops and optimizes production data pipelines that process large-scale payment data from many heterogeneous sources for analytics and BI. Builds PySpark/Spark ETL/ELT workflows on Hadoop/cloud, improving pipeline turnaround time through performance tuning and data-skew handling. Designs dimensional warehouse models (star schema and SCD) and performs SQL optimization for validation and reporting. Implements Kafka-based streaming with Spark Structured Streaming and supports large data migrations with strict data integrity.
Python
SQL
pySpark
Spark
Hadoop
Apache Kafka
AWS
Amazon S3
ETL/ELT
Apache Iceberg
CI/CD
Git
Dimensional Modeling
