697,746open jobs
41,033companies
105,083added this week
Browse all
Salary
$168k – $230k per year
Location
Remote/Hybrid (San Francisco, United States)
Seniority
Staff · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
AI platform connecting patients to clinical trials. Activate providers, identify eligible patients, and navigate them through trials across health systems.

About the Role

We're hiring a Staff Data Engineer to own the pipelines that move patient data through our platform, from EMR systems to the workflows our teams use to identify eligible patients for clinical trials and help them enroll.

This is the center of a major architectural shift for us: moving from narrow, human-assisted data pulls to robust, automated integrations that ingest complete patient records. The clean, complete, and trustworthy data you’ll manage powers our AI-assisted patient-trial matching. If healthcare data at HIPAA-grade stakes, messy real-world inputs, and pipelines that genuinely change patient outcomes sound like the right problem space, please keep reading.

Our engineering culture values direct communication, strong ownership, and low-ego collaboration. We laugh a lot and use a ton of Slack emojis. We make decisions quickly, give open feedback, and have a strong bias toward action.

How We Build with AI

AI is deeply embedded in how we work, in our product, and in our SDLC. Our embrace of it didn't come from a mandate but from our own desire as engineers to do more, safely. We use it for coding, testing, measuring, and iterating, and we treat it as a genuine competitive advantage. We're looking for someone who shares that instinct: excited to question long-standing processes and think ambitiously about what's possible when you pair strong engineering fundamentals with AI tooling. If you're already building that way, you'll fit right in.

Underneath it all is a modern platform built on AWS (Lambda, Fargate, SQS, RDS, Bedrock) with Pulumi-managed infrastructure. Our backend is primarily TypeScript, and our data layer centers on PostgreSQL and Drizzle. We care about pragmatic architecture, developer velocity, and systems that evolve as fast as our product and AI capabilities do.

Your Responsibilities

    You will own the ingestion and transformation pipelines that bring patient records into our platform and shape them into data our matching systems and human experts can trust.

    That means designing for reliability in compute-intensive, long-running workflows: the kind of problems where serial processing breaks under load, timeouts become production incidents, and the right architecture (async pipelines, message queues, container-based compute) separates a working feature from a failing one. You'll own data quality end-to-end: schema design, validation, transformation logic, and the monitoring that catches problems before a clinician does.

    You'll partner closely with our AI engineers on the data foundations for patient-trial matching and with our product engineers on how ingested data flows into the workflows our users depend on. You'll monitor production, triage issues quickly, and exercise pragmatic judgment by matching technology to business needs.

Your Qualifications

    We are looking for a data engineer with 8 or more years of experience, with at least a couple of years operating at staff scope or equivalent impact. Beyond tenure, what matters most is whether you have demonstrated the kind of high-impact ownership this role requires.

  • Strong pipeline engineering: ingestion, transformation, and orchestration of data at meaningful scale, with real attention to data quality and reliability.

  • Deep SQL and PostgreSQL fluency, including schema design and query performance

  • Solid Python and comfort working in a codebase with TypeScript.

  • Deep AWS experience (Lambda, Fargate, SQS, RDS, and the surrounding ecosystem) and the ability to choose the right service for the right job.

  • Proven track record of leveraging AI coding tools creatively and effectively to build production systems.

  • Demonstrated ability to take autonomous ownership, identifying and resolving systemic issues independently.

  • Startup experience, where you have built from scratch at an early-stage company and treated ambiguity as an opportunity rather than an obstacle.

  • Strong systems thinking, encompassing backend architecture, APIs, databases, and scalability under real-world constraints.

  • Clear communication and influence, with the ability to explain trade-offs to both engineers and non-engineers to earn trust through honesty.

  • Healthcare alignment with a genuine interest in improving clinical trial access and health equity; HIPAA experience a strong plus.

Nice to Have

    Familiarity with IaC tooling (Terraform, Pulumi, etc), clinical or healthcare data standards (EHR/EMR, HL7/FHIR), and API design patterns are all additive. Experience in a regulated industry is a plus.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
697,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
Remote/Hybrid • 2+ years exp • Bachelor's Degree
Python
SQL
Analytics
Tableau
Power BI
Microsoft Excel
Management
Outlook
Microsoft Office
Apply
Remote/Hybrid • 1+ year exp • Bachelor's Degree
Python
SQL
Analytics
Power BI
Microsoft Excel
Apply
$31k – $73k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Buenos Aires
Python
SQL
Databases
Google BigQuery
BigQuery
AI/ML
Claude
ChatGPT
Analytics
Tableau
Power BI
Looker
Apply
In office • Bachelor's Degree
SQL
Databases
Oracle
Teradata
DevOps
Red Hat
Azure
SLI/SLO/SLA
Linux
Windows
Analytics
Cognos
Management
ServiceNow
ITIL
ITSM
Apply
Equity • In office • Full-Time • 2+ years exp • Sunnyvale
Python
AI/ML
AI Agents
LLM
Agentic Workflows
DevOps
Azure DevOps
Azure
CI/CD
GitHub
Management
Confluence
Jira
Apply
$108k – $135k per year • Remote/Hybrid • Full-Time • San Francisco
Marketing
Salesforce
Apply
$58k – $78k per year • Remote • Contractor • 2+ years exp
Marketing
Salesforce
LinkedIn
Apply
$120k – $150k per year • Remote/Hybrid • Full-Time • 3+ years exp • San Francisco
Marketing
Salesforce
Apply
$168k – $205k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
Python
TypeScript
Databases
PostgreSQL
Amazon Aurora
AI/ML
AWS Bedrock
LLM
LLM Guardrails
DevOps
Terraform
GitHub Actions
CloudFormation
AWS
AWS Lambda
GitHub
Amazon S3
IAM
Amazon CloudWatch
VPN
Cybersecurity
SOC 2
HIPAA
NIST 800-53
Least Privilege
Threat Modeling
Apply
$120k – $150k per year • Remote/Hybrid • Full-Time • 10+ years exp • San Francisco
Marketing
Salesforce
Apply
Account Executive 1 hour ago
$150k – $250k per year • In office • Full-Time • 3+ years exp • San Francisco
Marketing
LinkedIn
Apply
$100k – $120k per year • In office • Full-Time • 3+ years exp • San Francisco
Marketing
LinkedIn
Apply
$55k – $120k per year • In office • Full-Time • San Francisco
Marketing
LinkedIn
Apply
$100k – $200k per year • Equity 0.1–2% • In office • Full-Time • San Francisco
Python
JavaScript
TypeScript
AI/ML
LLM
OpenAI
Anthropic
Frontend
React.js
Apply
$120k – $200k per year • Equity 0.5–2% • Remote/Hybrid • Full-Time • 6+ years exp • San Francisco
AI/ML
Prompt Engineering
Apply
See all jobs
This is one of many
697,746 more open roles from verified company boards, updated every day.