Data Scientist
7+ years exp
7+ years ML exp
Python
SQL
Active 14 days ago
Invite to interview
Message
Download CVCV
Overview
Technical skills
Timeline
Roles
Overview
Data engineer and Power BI developer with 7+ years delivering enterprise ETL/ELT, data warehousing/lakehouse, and analytics solutions across Azure and AWS. Builds Power BI semantic models and dashboards with security and performance optimization, and implements batch/streaming pipelines with orchestration and infrastructure automation. Experienced in data governance and compliance-focused engineering within regulated domains.
Technical skills
Python• Senior • 7y+
SQL• Junior • 4y+
Python
pySpark• 7y+
Databases
Databricks
Delta Lake
MS SQL
Oracle
PostgreSQL
Azure SQL Database• 7y+
Amazon Redshift• 4y+
Apache Kafka• 4y+
Snowflake• 4y+
Azure Cosmos DB
AI/ML
Spark• 7y+
Airflow• 4y+
DevOps
AWS Lambda
Azure
Rest API
SLI/SLO/SLA
Azure DevOps• 7y+
CI/CD• 7y+
Git• 7y+
AWS• 4y+
CloudFormation• 4y+
Docker• 4y+
Jenkins• 4y+
Terraform• 4y+
Cybersecurity
GDPR
HIPAA
Microsoft Defender
PCI DSS
Analytics
Power BI
Cryptography
Vault
Timeline
Senior Azure Data Engineer / Power BI Developer
•
Senior
FamilyWell Health
•
Full-Time
Collaborated with stakeholders to translate reporting needs into data models, ETL/ELT pipelines, and enterprise Power BI solutions in a lakehouse environment. Built Power BI semantic models and dashboards with security, incremental refresh, and performance optimization techniques. Created reusable dataflows and orchestration workflows for event-driven refresh and governed dataset publishing via CI/CD and workspace governance.
Azure SQL Database
Azure Cosmos DB
Azure DevOps
Git
CI/CD
Data Engineer
•
Middle
Bank of America
•
Full-Time
Built scalable cloud ETL/ELT and analytics solutions for financial and regulatory datasets using AWS batch and streaming services. Developed metadata-driven pipelines with PySpark and orchestration, and engineered distributed processing on EMR for performance and reliability. Optimized Redshift and Snowflake warehouses using tuning strategies and automated deployments with Terraform, CI/CD tooling, and infrastructure governance.
AWS
Airflow
Amazon Redshift
Snowflake
Terraform
CloudFormation
Jenkins
Docker
Git
Apache Kafka
pySpark
Python
Spark
SQL
Data Engineer
•
Middle
The Cigna Group
•
Full-Time
Designed ETL/ELT pipelines to ingest and transform clinical, claims, pharmacy, and provider data from multiple sources into Azure-based lakehouse and warehouse models. Implemented distributed PySpark transformations to improve data quality and reduce batch processing time, and set up near real-time ingestion using Azure services. Supported Power BI semantic modeling with security and incremental refresh, and delivered automated deployment and monitoring via Azure DevOps and CI/CD.
pySparksince 2019
Sparksince 2019
Pythonsince 2019
Azure SQL Databasesince 2019
Azure DevOpssince 2019
Gitsince 2019
CI/CDsince 2019
University of Dayton
Master's Degree •
Computer Science
