1,287,534open jobs
74,513companies
206,839added this week
Browse all
Salary
≈ $41k – $120k per year (Estimated)
Location
In office (Medellín)
Seniority
Junior · 2+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Oct 5, 2026.

Overview
Company
Impact
Profile match
Nearshore engineering for regulated industries. 200+ engineers across the Americas — purpose-built for life sciences and healthcare.

We’re looking for a Data Engineer to join Source Meridian.

About Source Meridian

Source Meridian is a development software company that works to solve the industry’s most challenging problems in healthcare practices. We are laser focused on specific technologies in the healthcare and life science industries: Healthcare technology, artificial intelligence, and healthcare interoperability.

About the Role

We're looking for a Data Engineer to help build and operate an AWS-native data platform processing healthcare claims data and tokenized identifiers. You'll design and implement Spark-based pipelines that transform, intersect, and enrich tokenized datasets stored primarily as Parquet on S3, queried via Athena and related AWS services. This environment intentionally avoids managed lakehouse platforms (e.g., no Databricks and no Snowflake)-you'll be doing "real" data engineering directly on AWS.

What You’ll Do

  • Build and maintain Spark pipelines to process large-scale Parquet datasets on S3.

  • Implement tokenization workflows, including transit token → real token conversion and dataset intersection/join logic.

  • Process and deliver healthcare claims datasets for matched individuals, ensuring accurate identity mapping and data integrity.

  • Orchestrate data pipelines using Airflow and/or AWS-native orchestration tools when appropriate.

  • Develop reliable, testable, and observable ETL/ELT processes (retries, idempotency, monitoring, reprocessing).

  • Optimize performance and cost across Spark jobs, S3 partitioning/layout, and Athena query patterns.

  • Contribute to dbt models when applicable (transformations, documentation, data quality checks).

  • Collaborate with cross-functional stakeholders in a healthcare environment, with a strong focus on privacy and secure data handling.

Required Qualifications

  • 2+ years of professional experience in Data Engineering.

  • Strong experience with  Apache Spark  (PySpark or Scala), including joins, intersections, partitioning, and performance tuning.

  • Strong hands-on experience with the  AWS data stack, including:

    • Amazon S3 (Parquet datasets, partition strategies, data layout best practices)

    • Amazon Athena (SQL, query optimization, managing large datasets)

    • Familiarity with AWS-native data lake patterns (Glue Catalog, Lake Formation concepts are a plus)

  • Experience building and operating pipelines using  Airflow  (DAGs, scheduling, dependencies, backfills).

  • Excellent SQL  skills and solid data modeling fundamentals.

  • Advanced English level: able to lead technical discussions, write clear documentation, and work directly with US-based stakeholders.

Nice to Have

  • Experience with  dbt  (core, tests, documentation, exposures).

  • Familiarity with healthcare data (claims data, eligibility, member-level datasets).

  • Experience with tokenization, identity resolution, or privacy-preserving data workflows.

  • Knowledge of AWS security concepts such as  IAM, KMS, encryption, and secure data handling.

  • Experience running Spark on AWS (e.g., EMR) or Spark-on-containers architectures.

Tech Stack

  • AWS-native architecture

  • Amazon S3 + Parquet (core storage layer)

  • Amazon Athena (query engine)

  • Apache Spark (no Databricks)

  • Airflow (orchestration)

  • dbt (optional, as applicable)

Soft Skills

  • Strong and empathetic leadership.

  • Proven client-facing experience.

  • Excellent communication skills.

  • Strong expectation management  abilities.

  • Strategic mindset with a solution-oriented approach and strong decision-making skills.

What We Offer

Permanent contract

Learning and continuous growth environment

Benefits package focused on health and well-being

Competitive salary based on experience

 Apply only if you reside in Colombia or Ecuador  

At Source Meridian, you’ll be part of a high-impact tech-health  company, building products that truly make a difference.

If you meet the profile - or know someone who might be interested -  apply now!

We’d love to meet you

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,287,534 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
Medellín
≈ $61k – $154k per year (Estimated) • In office • Full-Time • 5+ years exp • Quito
Python
SQL
Scala
Python
pySpark
Databases
Snowflake
Databricks
AI/ML
Spark
Airflow
dbt
DevOps
AWS
Amazon S3
IAM
Analytics
ETL/ELT
AWS Glue
Apply
$65k – $98k per year • Remote (United States) • Full-Time • United States
Python
SQL
Python
Flask
Databases
Snowflake
Databricks
AI/ML
Vertex AI
Scikit-learn
TensorFlow
NumPy
PyTorch
Amazon SageMaker
Machine Learning
DevOps
Rest API
GCP
Azure
CI/CD
AWS
Cybersecurity
HIPAA
Analytics
A/B Testing
Apply
≈ $128k – $253k per year (Estimated) • In office • Full-Time • San Jose
Python
SQL
Apply
≈ $108k – $210k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Dallas
SQL
Databases
Oracle
Microsoft Fabric
Cybersecurity
SOC 2
Apply
Data Engineer 1 hour ago
$200k – $300k per year • Hybrid • Full-Time • 5+ years exp • San Francisco • Seattle
SQL
AI/ML
Time Series Forecasting
Analytics
Power BI
Looker
Dimensional Modeling
Apply
≈ $44k – $113k per year (Estimated) • Hybrid • Contractor • 1+ year exp • Medellín
Management
Stripe
Apply
≈ $89k – $239k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Medellín
JavaScript
Apex
Apex
Lightning Web Components
Gearset
Visualforce
DevOps
CI/CD
Management
Jira
Agile
Apply
≈ $73k – $188k per year (Estimated) • Remote (Colombia) • Full-Time • 6+ years exp • Medellín
DevOps
Rest API
AWS
Apply
≈ $32k – $81k per year (Estimated) • In office • Full-Time • 2+ years exp • Medellín
Management
ServiceNow
Apply
≈ $53k – $153k per year (Estimated) • Hybrid • Full-Time • Medellín
JavaScript
TypeScript
Frontend
React.js
DevOps
Azure
Management
Power Apps
Agile
Apply
See all jobs
This is one of many
1,287,534 more open roles from verified company boards, updated every day.