368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$240k – $265k per year
Location
Remote (United States)
Seniority
Staff · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Hims & Hers is a telehealth and wellness company headquartered in San Francisco, California, and founded in 2017. The platform provides access to medical consultations and a range of prescription and over-the-counter products for hair loss, sexual health, skincare, mental health, and weight management. Operating primarily in the United States and the United Kingdom, the company is publicly traded on the New York Stock Exchange and focuses on making healthcare more accessible and affordable through a digital-first approach.

Hims & Hers is the leading health and wellness platform, on a mission to help the world feel great through the power of better health. We are redefining healthcare by putting the customer first and delivering access to care that is affordable, accessible, and personal, from diagnosis to treatment to delivery. No two people are the same, so we provide access to personalized care designed for results. By normalizing health & wellness challenges and innovating on their solutions, we’re making better health outcomes easier to achieve.

Hims & Hers is a public company, traded on the NYSE under the ticker symbol “HIMS.” To learn more about the brand and offerings, you can visit hims.com/about and hims.com/how-it-works . For information on the company’s outstanding benefits, culture, and its talent-first flexible/remote work approach, see below and visit www.hims.com/careers-professionals.

About the Role:

How do we make advanced AI/ML not just powerful but trustworthy enough to run in a regulated healthcare environment? We're looking for a Senior Staff engineer who can own that question end to end: the data pipelines that feed our models and evaluations, and the evaluation infrastructure - judges, scorers, statistical regression gates, red-team testing - that decides whether AI models are effective and safe to ship to patients.

This is a leadership role for someone who works comfortably across disciplines. You'll set technical direction for how we build, version, and trust the data and judgments that our AI products are evaluated against. Your scope will expand from raw data ingestion and feature/dataset pipelines, through evaluation methodology and statistical rigor, to the reporting surfaces that let clinical and product teams act on what we learn.

You'll spend most of your time on problems that don't have an existing playbook: ambiguous, cross-team, and genuinely hard to reason about. The job is to bring clarity to that ambiguity, chart a path the rest of the team and organization can follow, and see it through from idea to production, building the relationships and buy-in along the way to make it stick.

You Will:

Own the evaluation as a whole, not just a slice of it

  • Set the technical direction for our evaluation systems; metric, judge and scorer design, the statistical methodology behind regression decisions, and the infrastructure that tracks and categorizes failures over time.

  • Design and scale the data pipelines - ingestion, transformation, dataset versioning, labeling and calibration workflows - that both evaluation and downstream data science work depend on.

  • Proactively address challenges in scaling and complexity AI evaluation.

Lead projects that span teams and quarters

  • Define how we evaluate any new AI service from scratch. Drive multi-team initiatives like replacing manual, inconsistent review processes with statistically sound, automated gates.

  • Own our approach to adversarial and red-team evaluation as a risk-reduction program, designing the test suites and failure taxonomies that catch safety and edge-case issues before they reach patients.

  • Work through complex, cross-team technical disagreements and drive alignment across engineering, product and AI leaders.

Turn hard, ambiguous problems into solutions other teams can build on

  • Originate new approaches and methodology that becomes a reusable standard rather than a one-off fix.

  • Take vague, cross-team pain points ("we do this manually and it's inconsistent") all the way from a rough idea to a fully-specified, shipped system, without needing to hand off any part of the journey.

  • Lead major platform improvements; re-architecting core systems, removing brittle logic, modernizing how things run with impact that's felt org-wide, not just on your own team.

Grow the people and the network around you

  • Build real working relationships across ML engineering, data science, platform engineering, clinical, legal, and product, the kind of trust that gets you looped in early, before decisions are locked in.

  • Become a go-to voice on evaluation methodology and data pipeline design: share what you've learned in internal talks, write things up so other teams can use them, and expect your ideas to shape how others approach similar problems.

  • Mentor other engineers, including experienced ones, and help raise the technical and statistical bar of the teams you work with.

You Have:

  • 10+ years of experience in ML infrastructure, data engineering, or evaluation/testing systems, with a track record of impact that reaches beyond a single team or project.

  • Hands-on depth in evaluation systems: designing and calibrating LLM judges/scorers, building statistically sound regression-testing methodology (e.g., paired significance testing with proper correction for multiple comparisons), measuring agreement against human labels, and designing adversarial/red-team evaluation approaches.

  • Hands-on depth in data pipeline engineering: dataset versioning, feature and benchmark pipelines, labeling and calibration workflows, and high-throughput ingestion and transformation systems.

  • A history of building things that became the standard approach for others - not just solving your own problem, but changing how a broader group of people tackle a category of problem.

  • Experience leading multi-team projects to completion, including navigating and resolving genuine technical disagreement along the way.

  • A track record of mentoring other engineers, including senior ones, and visibly raising the bar for the teams around you.

  • Excellent communication - comfortable adapting the same idea for different audiences, and confident building support for it well before launch.

  • Strong Python, and enough statistical fluency to design and defend a testing framework that real production decisions ride on.

Nice to Have:

  • Experience with Databricks, MLflow, Unity Catalog, or similar data/eval platforms.

  • Experience building reporting tools for people without direct engineering access (e.g., automated Slack digests, spreadsheet reports for non-technical teams).

  • Prior experience in a regulated industry (healthcare, fintech, life sciences).

  • A track record of company-wide talks or write-ups that changed how other teams approached a problem.

Our Benefits (there are more but here are some highlights):

  • Competitive salary & equity compensation for full-time roles

  • Unlimited PTO, company holidays, and quarterly mental health days

  • Comprehensive health benefits including medical, dental & vision, and parental leave

  • Employee Stock Purchase Program (ESPP)

  • 401k benefits with employer matching contribution

  • Offsite team retreats

We are committed to building a workforce that reflects diverse perspectives and prioritizes ethics, wellness, and a strong sense of belonging. If you're excited about this role, we encourage you to apply-even if you're not sure if your background or experience is a perfect match.

Hims considers all qualified applicants for employment, including applicants with arrest or conviction records, in accordance with the San Francisco Fair Chance Ordinance, the Los Angeles County Fair Chance Ordinance, the California Fair Chance Act, and any similar state or local fair chance laws.

It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Hims & Hers is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures. If you need assistance or an accommodation due to a disability, please contact us at [email protected] and describe the needed accommodation. Your privacy is important to us, and any information you share will only be used for the legitimate purpose of considering your request for accommodation. Hims & Hers gives consideration to all qualified applicants without regard to any protected status, including disability. Please do not send resumes to this email address.

To learn more about how we collect, use, retain, and disclose Personal Information, please visit our Global Candidate Privacy Statement.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$98k – $195k per year (Estimated) • In office • Full-Time • 7+ years exp • Wellington
Java
Python
SQL
Java
Spring Boot
Databases
Apache Kafka
Databricks
Neo4j
AI/ML
Flink
Spark
Frontend
GraphQL
DevOps
Azure
CI/CD
Datadog
Dynatrace
Kibana
Kubernetes
OpenShift
Platform Engineering
Splunk
Amazon ECS
Apply
$84k – $178k per year (Estimated) • In office • Full-Time • 10+ years exp • Wellington
Java
Python
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
Helm
Kubernetes
Platform Engineering
Prometheus
Service Mesh
Terraform
GitLab
IAM
Apply
$70k – $105k per year • In office • Full-Time • 3+ years exp
Python
SQL
TypeScript
AI/ML
LLM
RAG
Function Calling
LLM Guardrails
Cybersecurity
GDPR
Management
n8n
Apply
Founding Engineer 1 day ago
$93k – $139k per year • In office • Full-Time • 3+ years exp • Munich
Python
Python
FastAPI
AI/ML
Fine-tuning
LLM
VLM
Apply
Staff Engineer 1 day ago
$105k – $232k per year • In office • Full-Time • 8+ years exp • Munich
Python
TypeScript
JavaScript
AI/ML
LLM
Frontend
React.js
DevOps
AWS
Azure
Docker
GCP
Terraform
Apply
$120k – $155k per year • Equity • Remote • Full-Time • 7+ years exp • Bachelor's Degree
Apex
AI/ML
Agentforce
DevOps
SLI/SLO/SLA
Marketing
Salesforce
Apply
$175k – $215k per year • Equity • Remote • Full-Time • 6+ years exp • Bachelor's Degree
Apply
$70k – $85k per year • Equity • Remote • Full-Time • 5+ years exp • Bachelor's Degree
JavaScript
Python
SQL
Analytics
Power BI
Apply
$130k – $150k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree
Design
SolidWorks
Apply
$75k – $145k per year (Estimated) • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree
Cybersecurity
CAPA
Management
Google Workspace
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.