368,941open jobs
9,452companies
47,951added this week
Browse all
Salary
$102k – $207k per year (Estimated)
Location
Remote/Hybrid (United States)
Seniority
Principal · 12+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

H1

H1 is a healthcare data company headquartered in New York City and founded in 2017. The company maintains a global database of healthcare professionals, their publications, trial history, and affiliations, which pharmaceutical companies use for clinical trial site selection, medical affairs, and commercial targeting. It sells to life sciences organizations and combines licensed and public data sources into a single provider graph.

At H1, we believe access to the best healthcare information is a basic human right. Our mission is to provide a platform that can optimally inform every doctor interaction globally. This promotes health equity and builds needed trust in healthcare systems. To accomplish this our teams harness the power of data and AI-technology to unlock groundbreaking medical insights and convert those insights into action that result in optimal patient outcomes and accelerates an equitable and inclusive drug development lifecycle. Visit h1.co to learn more about us.

The Data & Research team is responsible for the end-to-end lifecycle of H1's data - from source acquisition and onboarding through normalization, entity resolution, and product delivery. Working closely with Engineering, Product, and client-facing teams, this group ensures that every data source H1 adds is evaluated rigorously, integrated accurately, and delivered in a way that creates real value for the healthcare and life sciences organizations that rely on our platform.

WHAT YOU'LL DO AT H1

As a Principal Analyst, Data Integration, you will own the end-to-end process of evaluating, scoping, and onboarding new data sources into H1's platform. This is a senior IC role at the intersection of data, engineering, and product - the connective tissue between raw data acquisition and what ultimately ships to clients. You will work across Data & Research, Engineering, and Product to define what a new source is, how it maps to H1's schemas, what it can realistically deliver, and what it can't. You will also work directly with client-facing teams to gather requirements before integration decisions are made, translating commercial needs into data specs and data constraints back into product expectations.

You will:

- Lead structured evaluation of new data sources from scratch - assessing schema, coverage, freshness, legal constraints, and fit against H1's product needs before any engineering work begins

- Own field mapping from source to H1's bronze/silver/gold layers, producing data dictionaries, entity definitions, and structural guidance for downstream teams

- Partner with engineering and Data Lake to define ingestion requirements, entity resolution rules, and refresh cadences for new sources

- Gather requirements from client-facing teams and translate them into integration specifications; serve as the authoritative voice on what a new source can and cannot deliver before product commitments are made

- Shepherd each source end-to-end: scoping → QA → entity matching → product launch, including product QA and communicating source capabilities and limitations to product and enablement partners

- Work with the Insights team to develop new taxonomies and QA mechanisms for novel data types

- Define acceptance criteria and lead QA validation including field-level fill rates, count comparisons, and cycle-over-cycle anomaly detection

- Investigate and resolve data quality issues post-integration, coordinating with DART and engineering as needed

- Hand off to the maintaining team with complete mapping documentation; you own onboarding, not ongoing maintenance

- Produce and maintain documentation other people actually use - across scoping assessments, field mapping specs, and post-mortems

ABOUT YOU

You are a senior data professional who has personally owned the full lifecycle of a data integration - not just contributed to one. You are comfortable working across ambiguous, novel data structures and can ramp quickly on new source types each quarter. You operate with strong judgment about where integration risk lives, and you communicate clearly to both technical partners and client-facing stakeholders about what data can and cannot do.

REQUIREMENTS

- 8-12+ years in data-focused roles at healthcare data companies, pharma/biotech data vendors, health IT firms, or equivalent

- Demonstrated end-to-end ownership of data integrations built from scratch - scoping, field mapping, QA, and handoff - with documentation to show for it

- Healthcare or life sciences domain context required; ability to ramp on new datasets and source types each quarter without needing deep subject matter expertise upfront

- Analytical fluency to assess data quality; hands-on experience with tools such as VBA, R, or SPSS; SQL a plus but not a primary requirement

- Familiarity with data lake architectures (bronze/silver/gold or equivalent) and how raw data moves through normalization and entity resolution to a product-ready state

- Experience gathering requirements from client-facing stakeholders and translating them into data or product specifications

- Experience at a B2B data company where you understood how external clients consumed your data and where client retention drove decisions

- AWS infrastructure familiarity (Athena, S3, Glue) at a query and inspection level preferred

- Comfort working in Jira or Monday in a ticket-based workflow

- Exceptional written communication - your documentation is legible, maintained, and actually used

COMPENSATION

This role pays $145,000 to $170,000 per year, based on experience, in addition to equity.

Anticipated role close date: 08/25/2026

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,941 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$98k – $195k per year (Estimated) • In office • Full-Time • 7+ years exp • Wellington
Java
Python
SQL
Java
Spring Boot
Databases
Apache Kafka
Databricks
Neo4j
AI/ML
Flink
Spark
Frontend
GraphQL
DevOps
Azure
CI/CD
Datadog
Dynatrace
Kibana
Kubernetes
OpenShift
Platform Engineering
Splunk
Amazon ECS
Apply
$84k – $178k per year (Estimated) • In office • Full-Time • 10+ years exp • Wellington
Java
Python
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
Helm
Kubernetes
Platform Engineering
Prometheus
Service Mesh
Terraform
GitLab
IAM
Apply
$123k – $251k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Dallas • Denver • Birmingham
Java
SQL
Java
Gradle
Hibernate
Maven
Spring Boot
Spring Framework
Databases
Apache Kafka
MySQL
Redis
DevOps
CI/CD
Dynatrace
Jenkins
Kubernetes
OpenShift
Cybersecurity
SonarQube
Apply
Platform Engineer 1 day ago
$87k – $140k per year • In office • Full-Time • 3+ years exp • Berlin
Databases
PostgreSQL
Redis
DevOps
AWS
Azure
Bicep
CI/CD
Docker
GCP
GitHub Actions
Kubernetes
OpenShift
Terraform
GitHub
Apply
$70k – $105k per year • In office • Full-Time • 3+ years exp
Python
SQL
TypeScript
AI/ML
LLM
RAG
Function Calling
LLM Guardrails
Cybersecurity
GDPR
Management
n8n
Apply
$190k – $220k per year • Equity • Remote/Hybrid • Full-Time • 10+ years exp • New York
SQL
AI/ML
Claude
dbt
OpenAI
DevOps
SLI/SLO/SLA
Apply
$28k – $68k per year (Estimated) • Remote • Full-Time • 6+ years exp
Go
Node JS
Python
TypeScript
JavaScript
Node JS
Nest.JS
Python
pySpark
Databases
Apache Kafka
ElasticSearch
PostgreSQL
AI/ML
LangChain
LLM
RAG
Spark
Anthropic
OpenAI
Frontend
React.js
DevOps
AWS
CI/CD
Helm
Kubernetes
Vercel
Apply
$130k – $160k per year • In office • Full-Time • 5+ years exp • New York
Apply
$27k – $67k per year (Estimated) • Remote • Full-Time
Node JS
Python
JavaScript
Databases
ElasticSearch
PostgreSQL
AI/ML
Airflow
Claude
Copilot
Knowledge Graph
DevOps
CI/CD
CircleCI
Docker
Git
Kubernetes
Terraform
Analytics
ETL/ELT
Management
Jira
Apply
$170k – $190k per year • Equity • Remote • 8+ years exp • Bachelor's Degree
C++
Go
Java
Python
Scala
Databases
Apache Kafka
AI/ML
Spark
DevOps
AWS
Apply
See all jobs
This is one of many
368,941 more open roles from verified company boards, updated every day.