368,910open jobs
9,449companies
47,822added this week
Browse all
Salary
$108k – $226k per year (Estimated)
Location
In office (New York)
Seniority
Middle · 4+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

About Minerva

Minerva builds AI for marketing leaders. Our platform allows marketers to focus on telling their brand's story, delegating operationally intensive to our AI agents which handle data management, analytics, campaign generation, measurement, and reporting.

Everything is built on Minerva's proprietary consumer graph, an identity and attribute layer covering 270M+ U.S. consumers across 1,000+ temporal attributes. We have two agentic systems built through an OpenAI research partnership: an Agentic Data Engineer that unifies and standardizes a brand's first party data in hours, and an Agentic Data Scientist that trains robust targeting models at scale. Together, these systems enhance the quality of first party data, increase campaign performance, and give marketing teams back their time.

Our clients include leading consumer brands across categories: the NBA, Ramp, Capital One, Hard Rock Stadium Group / Miami Dolphins, Wander, and Trust & Will. We have raised $20M from The General Partnership, 8VC, Lingotto, NBA Investments, Topology Ventures, Future Positive, Background Capital, and others.

About the Role

As a Data Scientist at Minerva, you build the models and features that power our consumer graph and the agents that run on top of it. You sit at the intersection of heavy data engineering and applied modeling: you architect feature engineering pipelines that are computed over terabytes of data, train and sharpen the models that drive targeting and prediction, and ensure the outputs are robust enough to be consumed autonomously by our Minerva Agents and our world-class modeled attributes (i.e. income / wealth).

This is a role that will be deploying constantly to production. The models you build are not handed off to be deployed by someone else, you own the path from raw data to a feature or model that an agent can call reliably at scale. As we grow, your work becomes the foundation other systems are built on.

What You'll Do

  • Create new features for models and agents, expanding the predictive surface area of our consumer data lake and building the pipelines that turn raw signal into trusted attributes.

  • Improve existing models through rigorous feature engineering, including our income/wealth, home buyer, and home seller models.

  • Play a pivotal role in the buildout of our world-class data lake, shaping how terabytes of consumer data are stored, transformed, and made queryable for both humans and agents.

  • Build feature engineering pipelines that run efficiently at terabyte scale, with the data engineering rigor to make them reliable in production. This is a 70/30 split DS/DE role.

  • Ensure model and feature outputs are reliable enough to be consumed agentically, writing the validations and guardrails that let our agents act on your work without a human in the loop.

Our Data Stack

  • Dagster for all things orchestration

  • dbt-core within Dagster as the primary data transformation surface

  • Spark, Iceberg, Trino, AWS Glue for Lakehouse workloads

  • Modal for ML eng

  • Frontier + OSS models & agent SDKs. We are heavy users of OpenAI/Anthropic batch APIs

Qualifications

  • 2-4+ years working as a data scientist, applied machine learning focused data engineer or software engineer in a data-heavy context. Simply put, you live and breathe data.

  • Highly proficient at Python and SQL.

  • You are driven by first-principles thinking and are a go-getter. You reason about what datasets and features are necessary to solve a modeling problem, and are scrappy and clever enough to bring that to life.

  • Strong intuition for data engineering principles, especially around data cleaning/ingestion and data modeling. We prefer these core skills to be second-nature, freeing up thinking for architecting and executing large-scale data initiatives, especially given the advancement of AI coding tools.

  • Strong engineering background. You are comfortable deploying complicated production pipelines and working within larger production systems, not just in sandboxed or research environments.

  • Willingness to work in office in NYC (we provide a relocation package).

  • Flexibility and openness to wearing several hats. We are lean and things are always changing.

  • Eagerness to learn and grow with the company and your coworkers.

Preferred

  • Experience building and training predictive models (e.g. lead scoring, LTV, propensity, lookalike modeling).

  • Experience with orchestration tools like Dagster, Airflow, Prefect and SQL transformation tools like dbt, SQLMesh.

  • Experience with both transactional databases (e.g. Postgres, MySQL) and analytical databases (e.g. Snowflake, Redshift), with a bias toward the latter.

  • Familiarity with a cloud resource provider (e.g. AWS, GCP).

  • Familiarity with backend and ML/AI engineering.

  • Experience with AI coding tools (e.g. Cursor, Claude Code, OpenCode) as a force multiplier.

  • Prior work at an early-stage startup.

You don't need to tick every box. If you're strong on the engineering side and hungry to build models that matter, we want to hear from you.

Compensation

Base salary: $200,000 to $225,000, commensurate with experience. Competitive equity and a marquee benefits package.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,910 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
$21k per year • In office • Contractor • Yekaterinburg
Python
SQL
Python
FastAPI
Flask
AI/ML
Claude
Claude Code
Embeddings
Function Calling
LLM
RAG
OpenAI
OpenAI Codex
Structured Outputs
DevOps
Docker
Git
Apply
ML Engineer 1 day ago
Remote/Hybrid • 3+ years exp
Python
SQL
AI/ML
AI Agents
Computer Vision
MLFlow
PyTorch
RAG
A2A
Model Context Protocol
DevOps
AWS
Azure
CI/CD
Docker
Docker Compose
GCP
Kubernetes
Analytics
ETL/ELT
Apply
DevOps Engineer 2 1 day ago
$18k – $52k per year (Estimated) • In office • 3+ years exp • Bengaluru
Java
Python
Databases
Amazon Redshift
Apache Kafka
ElasticSearch
Redis
DevOps
Amazon EC2
Amazon EKS
ArgoCD
AWS
CI/CD
Docker
GCP
Grafana
Jenkins
Kubernetes
New Relic
Opsgenie
Prometheus
Terraform
Ubuntu
Amazon S3
IAM
Cybersecurity
Threat Modeling
Apply
$68k – $85k per year • In office • Full-Time • Master's Degree • San Jose
Python
AI/ML
AI Agents
DevOps
Amazon EC2
AWS
AWS Lambda
Bitbucket
CI/CD
CloudFormation
Docker
Git
Kubernetes
Terraform
Amazon S3
IAM
HPC
Cybersecurity
Least Privilege
Apply
$36k – $78k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Pune
PowerShell
Python
SQL
DevOps
AppDynamics
AWS
Azure
GCP
Grafana
Prometheus
Splunk
Apply
$200k – $225k per year • In office • Full-Time • 5+ years exp • New York
Python
SQL
Databases
Amazon Redshift
ElasticSearch
Google BigQuery
MySQL
PostgreSQL
Snowflake
Trino
AI/ML
AI Agents
Claude
Claude Code
Cursor
Dagster
dbt
Embeddings
Prefect
Spark
Anthropic
LLM Guardrails
OpenAI
DevOps
AWS
Apply
Staff Data Engineer 3 months ago
$200k – $250k per year • In office • Full-Time • 5+ years exp • New York
Python
SQL
Databases
Amazon Redshift
Apache Kafka
ElasticSearch
Google BigQuery
MySQL
PostgreSQL
Snowflake
AI/ML
AI Agents
Claude
Claude Code
Cursor
dbt
Flink
LLM
Spark
LLM Evaluation
LLM Guardrails
Model Context Protocol
OpenAI
OpenAI Agents SDK
DevOps
Vector
Analytics
ETL/ELT
Marketing
HubSpot
Meta Ads
Apply
General Application 3 months ago
$131k – $290k per year (Estimated) • In office • Full-Time • New York
AI/ML
AI Agents
OpenAI
Apply
Chief of Staff 3 months ago
$99k – $222k per year (Estimated) • In office • Full-Time • 2+ years exp • New York
AI/ML
AI Agents
OpenAI
Apply
Founding Product Lead 3 months ago
$200k – $250k per year • In office • Full-Time • 5+ years exp • New York
Databases
Snowflake
AI/ML
AI Agents
LLM
Model Context Protocol
OpenAI
Marketing
HubSpot
Salesforce
Apply
$197k – $374k per year (Estimated) • In office • Full-Time • 8+ years exp • PhD • New York
AI/ML
AI Agents
Claude
LangChain
OpenAI
Vertex AI
Management
n8n
Zapier
Apply
Senior Data Analyst 3 hours ago
$86k – $171k per year (Estimated) • Equity • Remote • Full-Time • 5+ years exp • Bachelor's Degree • New York
Python
SQL
Databases
Snowflake
AI/ML
Anomaly Detection
Claude
Copilot
Cursor
dbt
Edge AI
DevOps
AWS
Analytics
Tableau
Marketing
Salesforce
Apply
$170k – $230k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • New York • San Francisco
Go
Java
Kotlin
Python
Scala
Databases
Apache Kafka
Databricks
Delta Lake
Snowflake
AI/ML
Dagster
dbt
Flink
Spark
Knowledge Graph
Apply
$96k – $134k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • New York
JavaScript
Swift
TypeScript
Java
Java
Spring Framework
Databases
Apache Kafka
PostgreSQL
AI/ML
AI Agents
Claude
Copilot
Fine-tuning
Flink
LangChain
LangGraph
Llama
LlamaIndex
Prompt Engineering
PyTorch
RAG
TensorFlow
Transformers
Devin
Hugging Face
OpenAI
Frontend
Angular
React.js
Mobile
MVC
DevOps
AWS
CI/CD
Docker
Kubernetes
OpenShift
Splunk
Vector
GitHub
Analytics
Tableau
Apply
$150k – $180k per year • In office • Full-Time • PhD • New York
Python
AI/ML
Anthropic
Anthropic SDK
Computer Vision
Fine-tuning
LangChain
LlamaIndex
LLM
OpenAI
OpenAI SDK
RAG
DevOps
AWS
Azure
GCP
Apply
See all jobs
This is one of many
368,910 more open roles from verified company boards, updated every day.