701,226open jobs
41,584companies
98,076added this week
Browse all
Salary
$155k – $170k per year
Location
Remote (United States)
Seniority
Senior · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
We give innovators the data they need to build healthcare AI.

Our Team

Dandelion Health was founded in 2020 by experts in health tech, hospital systems, academia, and clinical AI. We are building the world’s largest AI training and clinical development platform. Today, we pride ourselves on our ability to make data access as easy as possible for AI developers, pharma, and medical devices, while raising the bar for patient safety and data quality. Tomorrow, we will be the place where any healthcare organization can go to build a responsible clinical AI product. Our culture is all about learning from data and improving, so we can help our clients improve health through AI. Meet the rest of our team here.

Our Data

We partner with health systems to safely and ethically make their de-identified patient data available to AI developers. Currently, the data is acquired from Sharp HealthCare, Sanford Health, and Texas Health Resources - with two additional U.S. health systems joining soon.

We have clinical data dating back to July 1, 2016. This data represents over 10 million patients and includes but is not limited to:

  • Structured data (e.g., 100% of the EMR, including some claims)

  • Unstructured text (e.g., clinical notes, radiology reports)

  • Images (e.g., DICOM, pathology)

  • Video

  • Waveforms

  • Continuous streaming monitoring data

Your Role

You are a healthcare data scientist who knows your way around clinical and electronic health record data. Your primary responsibility is to partner collaboratively with our clients to develop data science solutions and analyze AI-ready datasets to provide key insights using multimodal data. You will curate datasets by identifying patient subpopulations or disease cohorts, and pool multimodal data to drive rapid exploratory AI/ML and Real-World Evidence analyses, model experimentation, and/or model validation. You will lead analytics research from inception to closeout by having ownership over study design, data curation strategies, effort estimation, analytic design, and delivery of final results. You will use your data expertise, programming abilities, and critical thinking skills to support our clients by designing and delivering solutions to help them tackle a broad range of business challenges. Your team’s ultimate goal is to deliver the highest-quality data possible to our clients, who are building products that improve patient health, and demonstrate the value of these data. You will report to the Research Data Science Manager.

Responsibilities

Your day-to-day responsibilities will include the following:

  • Collaborate with clients to develop creative analytic solutions that create value and address challenges across critical areas of their business.

  • Design statistical analysis plans for evidence generation projects that include methodology, results structure, expected outputs and dataset requirements;

  • Query complex source systems in a range of health data sources (e.g., EMRs, ECG data, DICOM data) to identify and map data elements in order to create high-quality datasets for analytic and AI use cases;

  • Develop descriptive analyses and in-depth predictive and causal inference models to drive evidence generation on Dandelion data and enhance existing data products;

  • Summarize the methods and results from projects into clear explanations and documentation for internal and external audiences including for submission to conferences and peer-reviewed publications;

  • Support all phases of SQL/analytical programming, data management, quality control, and reporting for analytics projects;

  • Develop code and documentation to deliver high-quality and HIPAA-compliant data products on time to internal and external customers;

  • Create summarized findings and recommendations that are clearly presented and adapted for audiences that have a varying range of technical and clinical experience;

  • Identify and resolve problems using your knowledge, background, and troubleshooting skills;

  • Ensure accuracy, data integrity, and validity of data and analysis in all work;

  • Present to senior leadership as well external audiences

You are not afraid to dig into massive, confusing, disorganized new datasets and get them under control. You are excited to learn new environments, languages, and skills. This is a small, early stage company with enormous ambitions and everyone pitches in across the team.

Qualifications

  • M.S. or PhD degree in a relevant field, such as Biomedical Informatics, Data Science, Biostatistics, or B.S. with at least seven years of experience

  • Experience in client interaction, crafting project proposals, and leading teams.

  • Background in statistical modeling and real-world data analysis

  • Fluency in Python and SQL

  • 1+ years experience with extracting, curating, and analyzing data created within the HIT and healthcare delivery ecosystem (e.g., EMR, claims, registry); this may include knowledge of the roles of data exchange and content standards (e.g., FHIR, CDA, CQL) and clinical terminology standards (e.g., ICD, CPT, LOINC, SNOMED-CT, NDC, RxNorm)

  • Strong technical writing, editing, and communication skills along with a collaborative, client-first mindset

  • Excellent organizational skills with an ability to embrace change and effectively manage multiple projects and consistently plan work to meet deadlines

  • Experience working in or with startups is a plus

Technology Experiences and Skills

We don’t expect anyone to have all of the following skills or experiences, but we do seek candidates who are interested in growing their skill sets and working with healthcare data in all its glorious complexity. The Data Team works closely with our Engineering Team to put our work into production and meet client needs.

  • SQL

  • Python and/or R

  • Git and version control

  • Familiarity with encryption methods and writing regular expressions

  • Prior experience querying EDWs or databases and creating reports or analytics for healthcare data

  • Familiarity with the data aspects of electronic medical records, ex. Epic, Cerner, Allscripts

  • Prior experience working with insurance claims data

  • Familiarity with medical terminologies or controlled vocabularies such as ICD-10, SNOMED-CT, LOINC, CPT/HCPCS, NDC, and RxNorm

  • Any medical ontology experience

  • Any NLP experience

  • Any experience working with DICOM or other imaging modalities

  • Experience with OMOP common data model

  • Familiarity with Machine Learning concepts

  • Experience with AWS

Note that familiarity with machine learning model development and deployment is not required for this position, but familiarity with high-level ML concepts is a plus.

Nature of our work

Our work is fast paced and iterative. We are growing, and we want to support our team members to grow in their skills as well. We are building a team that approaches problems with a diversity of perspectives, values experimentation, and refining our approach based on that experimentation. We work with the full spectrum of healthcare data from tabular data, videos, images, waveforms, etc. If a health system collects it, we might work with it!

If this looks like a partial fit, please reach out, we would love to share more about the work we do for you to understand if it would be a good fit for you.

There is occasional travel for in-person company working days on roughly a quarterly basis.

Team Benefits

  • Remote work and flexible hours. Availability needed for meetings, which we try to keep to a healthy minimum

  • Complete wellness benefits including healthcare, dental, vision, PTO, sick days and more. Ask for details

  • Professional development days to build your skills

  • Collegial work environment

  • Academic bent towards inquiry and problem solving but start-up speed and flexibility

  • Great balance of focus time to work on projects but easy to access team members to discuss issues and work collaboratively

  • Dandelion is a mission-driven company that is focused on improving patient care

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
701,226 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$149k – $287k per year (Estimated) • Remote/Hybrid • 5+ years exp • Master's Degree • San Francisco
Python
SQL
AI/ML
Copilot
Claude
Fine-tuning
SciPy
Computer Vision
AI Agents
PyTorch
Statsmodels
Machine Learning
DevOps
AWS
Analytics
Seaborn
Matplotlib
Apply
Senior Data Engineer 7 hours ago
$87k – $212k per year (Estimated) • Remote • 5+ years exp • Bogotá
Python
SQL
Python
FastAPI
pySpark
Databases
Snowflake
Databricks
Google BigQuery
BigQuery
AI/ML
Cursor
LangGraph
LangChain
Spark
Claude Code
Dagster
dbt
Scikit-learn
PyTorch
Gemini
LLM
Structured Outputs
DevOps
Terraform
GCP
CloudFormation
Pulumi
CI/CD
AWS
Docker
Platform Engineering
Google Cloud Run
Apply
$335k – $385k per year • Remote/Hybrid • 10+ years exp • Bachelor's Degree • San Francisco
AI/ML
Multimodal AI
Anthropic
Interpretability
Machine Learning
Apply
$120k – $165k per year • Remote/Hybrid • Public Trust • 5+ years exp • Bachelor's Degree • Ashburn
Python
JavaScript
Java
TypeScript
COBOL
Java
Spring Boot
Hibernate
COBOL
IBM MQ
Databases
PostgreSQL
ActiveMQ
Apache Kafka
Frontend
Angular
Mobile
JUnit
DevOps
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Bamboo
Linux
Cybersecurity
SonarQube
Management
Jira
Agile
Apply
$170k – $250k per year • In office • Bachelor's Degree • Seattle
Python
Apply
$135k – $165k per year • Remote • Full-Time • 5+ years exp • New York
Python
SQL
AI/ML
Unstructured.io
Multimodal AI
NLP
TensorFlow
PyTorch
LLM
Hugging Face
Machine Learning
DevOps
Git
AWS
Cybersecurity
HIPAA
Apply
$140k – $160k per year • Remote • Full-Time • 5+ years exp • New York
AI/ML
Multimodal AI
Marketing
LinkedIn
Apply
$160k – $180k per year • Remote • Full-Time • 7+ years exp • Bachelor's Degree • New York
Apply
Analytics Engineer 2 months ago
$140k – $150k per year • Remote • Full-Time • 3+ years exp • Bachelor's Degree
Python
SQL
Databases
Snowflake
AI/ML
Unstructured.io
dbt
Multimodal AI
NLP
AWS Bedrock
LLM
RAG
OpenAI
DevOps
Prometheus
Azure
CI/CD
Git
AWS
Cortex
Analytics
ETL/ELT
Management
Linear
Jira
Agile
Apply
$135k – $150k per year • Remote • Full-Time • 4+ years exp
Management
Jira
Marketing
Salesforce
Apply
See all jobs
This is one of many
701,226 more open roles from verified company boards, updated every day.