Data Scientist
8+ years exp
7+ years ML exp
Bash
PowerShell
JavaScript
Python
SQL
TypeScript
Active 14 days ago
+1 (919) 3356891 Invite to interview
Message
Download CVCV
Overview
Technical skills
Timeline
Roles
Overview
Senior Data Engineer with about 7 years of experience designing and operating cloud data pipelines across Azure, AWS, and GCP. Focused on ETL/ELT, SQL and Python-based transformations, orchestration, and data quality improvements, including governed analytics with Databricks and Delta Lake. Experienced in CI/CD automation and performance tuning for data platforms supporting enterprise reporting.
Technical skills
Bash
PowerShell
JavaScript
Python• Middle • 8y+
SQL• Senior • 8y+
TypeScript• Middle • 5y+
Python
pySpark• 5y+
Databases
Apache Kafka
Azure SQL Database
MySQL
PostgreSQL
Snowflake
MS SQL
Google BigQuery• 5y+
Amazon Redshift
Databricks
Delta Lake
AI/ML
Airflow
Dagster• 7y+
Prefect• 7y+
LLM• 5y+
Spark• 5y+
Prompt Engineering
DevOps
Amazon EC2
Amazon EKS
Ansible
Bicep
CloudFormation
GitHub Actions
Google GKE
Jenkins
New Relic
SLI/SLO/SLA
Terraform
Kubernetes
GCP• 5y+
AWS
AWS Lambda
Azure
Azure DevOps
CI/CD
Git
Cybersecurity
CIS Benchmarks
Microsoft Entra ID
NIST 800-53
PCI DSS
Analytics
Power BI
Tableau
Cryptography
Vault
Timeline
Sr. Data Engineer
•
Senior
T-Mobile
•
Full-Time
Designed and implemented ETL pipelines and made architecture decisions to improve ingestion speed and ensure reliable analytics delivery. Built Azure-based governed data pipelines using Azure Data Factory, Databricks, and data lake storage, and developed PySpark transformations with Delta Lake for higher data quality. Automated CI/CD deployments using Azure DevOps and Git and tuned Synapse SQL workloads to deliver trusted data marts for business dashboards.
Azure
Databricks
Delta Lake
pySpark
Spark
Azure DevOps
Git
CI/CD
Prompt Engineering
SQL
Data Engineer
•
Middle
McKesson
•
Full-Time
Led data quality initiatives to improve model outcomes and reduce production data issues in healthcare workflows. Built and optimized AWS-based data infrastructure for high-throughput processing and reliable availability for analytics consumption. Standardized ETL jobs using AWS services, developed Python/SQL workflows, and automated incremental loading with monitoring to reduce manual intervention. Enhanced reporting datasets by configuring Redshift schemas and validating data with quality checks.
AWS
TypeScript
Amazon Redshift
AWS Lambda
SQL
Python
Data Engineer
•
Middle
Accenture
•
Full-Time
Designed and implemented TypeScript-based LLM workflows on cloud infrastructure, improving reliability and reducing latency for high-volume processing. Improved ML operations by applying MLOps practices that shortened model deployment cycles and reduced production incidents. Built and validated BigQuery data pipelines using GCP streaming tooling, then transformed datasets with PySpark and SQL for curated assets. Modeled dimensional tables in BigQuery to enable consistent metrics and faster self-service analytics.
TypeScriptsince 2021
LLM
Google BigQuery
pySparksince 2021
SQL
GCP
Confluence
Sparksince 2021
Associate Data Engineer
•
Middle
Cognizant
•
Full-Time
Streamlined risk modeling data workflows to improve decision-making speed and predictive analytics accuracy. Modernized ETL orchestration by implementing Dagster and Prefect for improved pipeline efficiency and operational visibility. Developed SQL extraction logic and reusable Python validation scripts to detect defects earlier and improve delivery readiness. Documented data dictionaries and lineage in Confluence and supported Agile sprint execution to deliver tested ETL increments.
SQL
Python
Dagster
Prefect
Confluencesince 2019
Data Engineer Intern
•
Junior
Espire
•
Full-Time
Built observability tooling for product support to strengthen monitoring and improve uptime for critical services. Assisted in developing user-facing features with version control and contributed SQL query work to improve data extract accuracy. Created Python scripts for file processing and data cleansing to reduce repetitive tasks. Mapped source columns to target schemas and tested sample loads against data quality rules during internship cycles.
SQLsince 2018
Pythonsince 2018
University of Bridgeport
Master's Degree •
Computer Science
