368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$180k – $220k per year
Location
Remote/Hybrid (San Francisco, United States)
Seniority
Senior · 6+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Metriport is a health data infrastructure company headquartered in San Francisco, California, and founded in 2022. The company provides an open source API that retrieves and normalises patient medical records from health information exchanges, electronic health record systems, and consumer wearables into a single format. It sells to digital health companies, providers, and payers that need patient history without building separate integrations for each data source.

Senior Data Engineer

San Francisco, CA

Hybrid

About us

Metriport is an open-source data intelligence platform that helps healthcare organizations access and exchange patient data in real-time: record retrieval that used to take days now takes seconds, and clinicians walk into every appointment with the full patient picture. We integrate with all major US healthcare IT systems and tap into comprehensive medical data for 300+ million individuals.

We've found product-market fit with multi-million ARR, 100+ customers (including Amazon One Medical, Strive Health, Circle Medical, and Brightside Health), backing from top VCs, and years of runway. We're ready to scale. We're a tight-knit, high-performing team of mostly former founders (including two YC alumni). We're engineering-heavy, operate with minimal bureaucracy and high autonomy, and hire based on competence, not prestige. We push hard-founders work six days a week from our SF office-but give everyone freedom to craft their schedule. We measure output and we're committed to sustainable intensity.

About you

We're looking for a data engineer who can operate at a Senior level:

  • You've built and scaled data pipelines and systems, with hands-on experience across the ecosystem - distributed processing, lakehouses, warehouses, streaming, orchestration. You know when each is (and isn't) the right tool for the job, and people usually come to you for guidance on how to move, transform, and serve data reliably.

  • You fully own your work - technical decisions, delivery, results - and you level up the engineers around you through code reviews, pairing, and guidance. Multiplying the team's output energizes you as much as shipping your own.

  • You're entrepreneurial-minded with an olympian-level work ethic (about half our engineering team are former founders).

  • You understand data reliability, quality, and governance as first-class parts of delivery.

  • You care about delivering value to customers, not about what frilly new tech is under the hood.

  • When someone scopes a project for 3 weeks, you ask "why can't it be done in 3 days?" - and you help others develop that same instinct.

  • You're a hacker at heart, with a good sense of which rules should, and shouldn't, be broken.

What you'll be doing

We ingest clinical data for millions of patients from external healthcare sources, with continuous updates for a growing subset of those patients. You'll be a catalyst to scale the data platform that powers our product - and ship it to customers fast.

Day to day, that looks like:

  • Raising the technical bar for our data platform: building on and improving our warehouse, data lake, and ETL/ELT architecture so it scales with patient and customer growth, and helping evaluate the right tools (batch and streaming processing, table formats, orchestration, query engines).

  • Driving data projects end-to-end: writing Design Documents, shipping v0's quickly, and iterating to v1 and beyond.

  • Supporting AI/ML efforts - making sure the AI Engineers have the data they need.

  • Multiplying the team: mentoring engineers on data fundamentals, reviewing designs and PRs, and judging when to invest in quality vs. ship fast.

  • Participating in bi-weekly sprint planning and retros, joining our daily 30-min remote stand-up at 7:30am PST (our only mandatory meeting), and taking part in the on-call rotation.

Example projects you could own:

  • Scaling our patient data consolidation pipeline (deduplication, normalization, hydration) to handle 100x today's volume without 100x the cost.

  • Building pipelines that deliver clinical data directly into customers' data warehouses, reliably and at scale.

  • Building the ingestion path for customers pushing large volumes of their own data into the platform.

  • Building document-processing pipelines that extract structured data from PDFs, images, and free text to feed ML models.

Requirements

  • 6+ years of engineering experience, with a heavy lean towards data engineering - building, maintaining, and scaling pipelines processing terabytes of data and millions of events a day.

  • Experience across the data stack - ingestion, storage, processing, warehousing, serving - and an understanding of the tradeoffs (cost, latency, correctness, operability) at each layer.

  • Experience with modern, cloud-native data stacks: e.g., Spark, open table formats (Parquet, Iceberg, Delta) on S3, and warehouses (Snowflake, BigQuery, Redshift).

  • Strong software engineering fundamentals - you write production code (we're a TypeScript shop, with Python in data/ML workflows), not just orchestration configs.

  • Experience mentoring or guiding other engineers - through code reviews, pairing, design feedback, or onboarding.

  • Located in San Francisco / Bay Area, or willing to relocate.

  • Bonus:

    • Experience with streaming systems (Kafka, Kinesis), dbt, or orchestration tooling (Airflow, Dagster).

    • Experience building or supporting ML/data science workflows (feature pipelines, model inputs/outputs, unstructured data extraction).

    • Healthcare standards/technologies: FHIR, HIE, IHE, EHR/EMR, NPI, TEFCA, ADT, HL7, HEDIS, RAF, SNOMED, LOINC, ICD-10, etc.

Benefits

  • Competitive equity + compensation package

  • Full family Platinum health insurance, dental, and vision coverage

  • 401(k) retirement plan + matching

  • Flexible work from home or in-office

  • Healthy lunches are complimentary when working in-office (and breakfast + dinners as needed)

  • Quarterly company off-sites with the team

  • MacBook provided by us

  • Unlimited PTO (we work hard, but trust you to take time you need to be at your best)

Our tech

Core business logic in Node.js and TypeScript, with Python in data and ML workflows. AWS across the board (ECS, Lambda, SQS, SNS, Batch, etc.), infrastructure as code with CDK. Data lives in S3, PostgreSQL/Aurora, DynamoDB, Snowflake, and our FHIR server - with Athena for querying S3 and SageMaker for ML. Our data platform is still early: you'll shape what we adopt next, picking the best tool for the job rather than the trendiest one.

Metriport provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, national origin, age, disability, genetics, sexual orientation, gender identity, or gender expression. We are committed to a diverse and inclusive workforce and welcome people from all backgrounds, experiences, perspectives, and abilities.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
Team Lead DevOps 1 day ago
$23k – $62k per year (Estimated) • Remote • 5+ years exp • Moscow
Bash
Python
Erlang
Erlang
EMQX
Databases
Apache Kafka
ClickHouse
PostgreSQL
RabbitMQ
Redis
Redpanda
Trino
DevOps
Ansible
AWS
AWX
FinOps
HAProxy
Hetzner
Kubernetes
SLI/SLO/SLA
Terraform
Yandex Cloud
Amazon S3
Apply
$123k – $251k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Dallas • Denver • Birmingham
Java
SQL
Java
Gradle
Hibernate
Maven
Spring Boot
Spring Framework
Databases
Apache Kafka
MySQL
Redis
DevOps
CI/CD
Dynatrace
Jenkins
Kubernetes
OpenShift
Cybersecurity
SonarQube
Apply
Platform Engineer 1 day ago
$87k – $140k per year • In office • Full-Time • 3+ years exp • Berlin
Databases
PostgreSQL
Redis
DevOps
AWS
Azure
Bicep
CI/CD
Docker
GCP
GitHub Actions
Kubernetes
OpenShift
Terraform
GitHub
Apply
Founding Engineer 1 day ago
$81k – $116k per year • In office • Full-Time • Bachelor's Degree • Munich
JavaScript
Python
TypeScript
Databases
MySQL
PostgreSQL
Frontend
Next.js
React.js
Tailwind CSS
DevOps
AWS
Azure
CI/CD
Docker
GCP
Grafana
Kubernetes
OpenTelemetry
Prometheus
Apply
Product Engineer 1 day ago
$105k – $128k per year • In office • Full-Time • Leipzig
TypeScript
Databases
PostgreSQL
AI/ML
AI Agents
DevOps
AWS
Apply
$160k – $190k per year • In office • Full-Time • 7+ years exp • San Francisco
DevOps
Rest API
Apply
Product Design 10 days ago
$140k – $180k per year • In office • Full-Time • San Francisco
Node JS
TypeScript
JavaScript
Databases
DynamoDB
PostgreSQL
Snowflake
Frontend
React.js
DevOps
AWS
AWS CDK
AWS Fargate
AWS Lambda
Amazon ECS
Amazon S3
Design
Figma
Apply
$180k – $220k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
Node JS
TypeScript
JavaScript
Databases
DynamoDB
PostgreSQL
Snowflake
AI/ML
AI Agents
Prompt Engineering
Function Calling
Frontend
React.js
DevOps
AWS
AWS CDK
AWS Fargate
AWS Lambda
CI/CD
GitHub Actions
Rest API
WebSockets
Amazon ECS
Amazon S3
GitHub
Cybersecurity
Least Privilege
Management
Slack
n8n
Zapier
Apply
$180k – $220k per year • Remote/Hybrid • Full-Time • 6+ years exp • San Francisco
Node JS
Python
TypeScript
JavaScript
Databases
Amazon Aurora
DynamoDB
PostgreSQL
Snowflake
AI/ML
Amazon SageMaker
DevOps
AWS
AWS Lambda
CI/CD
Amazon ECS
Amazon S3
Apply
$180k – $220k per year • Remote/Hybrid • Full-Time • 6+ years exp • San Francisco
Node JS
Python
TypeScript
JavaScript
Databases
Amazon Aurora
DynamoDB
PostgreSQL
Snowflake
AI/ML
Amazon SageMaker
DevOps
AWS
AWS Lambda
CI/CD
Amazon ECS
Amazon S3
Apply
$293k – $385k per year • In office • Full-Time • San Francisco
AI/ML
OpenAI
Apply
$180k – $260k per year • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco
Python
AI/ML
ChatGPT
OpenAI
OpenAI Codex
Cybersecurity
FedRAMP
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.