489,618open jobs
16,287companies
72,295added this week
Browse all
Salary
$247k – $441k per year (Estimated)
Location
Remote/Hybrid (New York, United States)
Seniority
Staff · 5+ years exp
Overview
Company
Impact
Profile match
Snorkel AI was founded in 2019 by Stanford researchers who developed programmatic labelling, where subject matter experts write rules that generate training data instead of annotating examples by hand. Its platform now covers data curation, model fine-tuning and expert evaluation for enterprise and frontier model builders. Banks, insurers and government agencies use it to specialise models on proprietary knowledge.

About Snorkel

At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data.

We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before. Excited to help us redefine how AI is built? Apply to be the newest Snorkeler!

About The Team

Snorkel's AI Platform organization builds the infrastructure and systems that power AI development at scale - synthetic data generation, evaluation, agentic workflows, simulation environments, LLM infrastructure, and distributed compute. Our platform enables engineering and research teams to rapidly experiment with models and agents, measure their behavior, and turn successful experiments into reliable production systems.

We're a small team operating at the intersection of distributed systems and applied AI, and we're in the middle of a foundational shift toward agent-first workflows where models interact with tools, environments, data, and other agents over long-running trajectories. The systems we build need to make these inherently non-deterministic workloads observable, reproducible, measurable, and scalable. You will help define how we do that.

About The Role

We're looking for AI Engineers who combine strong software and distributed systems fundamentals with experience operating AI systems in production. You'll build the infrastructure that lets teams create, experiment with, evaluate, and operate LLM and agentic workloads at significant scale - from synthetic data and evaluation pipelines to simulation environments, orchestration systems, and LLM infrastructure.

You'll work on systems where correctness is not defined by a single deterministic output. Instead, you'll build the infrastructure needed to understand behavior across models, prompts, tools, environments, and multi-step trajectories, and to continuously improve those systems through experimentation and evaluation.

We are looking to grow our team of AI Engineers, and are hiring at multiple levels.

What You'll Do

  • Design and build infrastructure for running large-scale agentic workloads, including multi-step agents interacting with tools, external services, sandboxes, and simulated environments
  • Build scalable synthetic data generation and automated labeling systems that allow teams to create, refine, and evaluate high-quality training and evaluation datasets
  • Design evaluation infrastructure for measuring AI system behavior across models, prompts, tools, environments, and multi-step trajectories - including reproducible experiments, benchmark execution, regression detection, and continuous evaluation
  • Build orchestration and distributed compute systems for running thousands to millions of AI experiments and simulations reliably across heterogeneous compute environments
  • Develop infrastructure for agent simulation environments, including environment provisioning, isolation, lifecycle management, and scalable execution
  • Build and operate LLM infrastructure for routing, rate limiting, retries, caching, provider failover, cost attribution, and efficient execution across multiple model providers
  • Instrument agent and model workloads so failures are observable and debuggable - capturing traces, model interactions, tool calls, environment state, evaluation results, latency, reliability, and cost
  • Design systems that make non-deterministic workloads reproducible and measurable, allowing engineers to compare experiments, diagnose behavioral regressions, and understand why an agent succeeded or failed
  • Improve the developer experience for AI experimentation by building APIs, SDKs, workflow abstractions, and tooling that make it easy to move workloads from local development to large-scale production execution
  • Collaborate with research, product, and engineering teams to turn experimental AI workflows into reliable, reusable platform capabilities

What You'll Bring

  • 5+ years building production software systems, with experience in AI/ML infrastructure, ML platforms, distributed systems, data platforms, or backend infrastructure
  • Experience operating non-deterministic AI or ML workloads in production or at significant scale - you are comfortable reasoning about behavior across models, tools, environments, and multi-step execution
  • Experience building infrastructure for experimentation, evaluation, model development, synthetic data, agentic workflows, training, inference, or production ML systems
  • Strong proficiency in Python and experience building production-quality APIs, services, and developer tooling
  • Strong background in distributed systems and cloud platforms (AWS preferred), including compute orchestration, storage, networking, isolation, and failure handling
  • Experience with workflow or distributed execution frameworks such as Prefect, Airflow, Dagster, Ray, Kubernetes, or similar systems
  • Strong understanding of production system fundamentals - observability, telemetry, reliability, performance, debugging, incident response, and cost management
  • Ability to reason about AI system quality beyond traditional service metrics, including evaluation design, experiment reproducibility, behavioral regressions, and model or agent variability
  • Track record of leading complex engineering initiatives, influencing stakeholders, and delivering measurable impact
  • Ability to work in a fast-paced environment with strong technical communication skills
  • Fluency with modern AI and developer tooling and a willingness to rapidly evaluate and adopt new models, frameworks, infrastructure, and techniques as the ecosystem evolves

Nice to Have

  • Experience building or operating LLM or agent infrastructure, including model gateways, agent runtimes, tool execution, tracing, or multi-agent systems
  • Experience building evaluation or experimentation platforms for LLMs, agents, or other probabilistic systems
  • Experience with synthetic data generation, automated labeling, data refinement, or dataset quality systems
  • Experience building reinforcement learning environments, agent simulations, benchmarks, or other environment-based evaluation systems
  • Experience running large-scale distributed AI workloads across containers, Kubernetes, serverless compute, sandboxes, or heterogeneous compute environments
  • Experience with LLM observability, tracing, prompt/version management, token and cost attribution, rate limiting, caching, or multi-provider routing
  • Experience designing isolation and sandboxing infrastructure for executing model-generated code or tool calls safely
  • Experience building shared AI platform libraries or SDKs consumed by multiple teams, including versioning, backwards compatibility, and migration support
  • Experience in hyper-growth startup environments or scaling engineering organizations
  • Prior experience as a Tech Lead, Team Lead, or hands-on Engineering Manager

Why This Role

You'll have meaningful ownership over the infrastructure that determines how quickly Snorkel can experiment with, evaluate, and productionize new AI systems. This isn't a role focused on training a single model or maintaining traditional ML pipelines - you'll build the platform that makes large-scale AI experimentation possible.

You'll work on some of the hardest emerging infrastructure problems in AI: operating non-deterministic systems reliably, reproducing agent behavior across environments, evaluating long-running trajectories, scaling simulations and experiments, and turning rapidly evolving research workflows into robust production systems.

The architecture decisions made by this team will define how Snorkel builds and operates AI systems for years to come.

Snorkel is proud to be an Equal Employment Opportunity employer and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. Snorkel embraces diversity and provides equal employment opportunities to all employees and applicants for employment.

Actual compensation will be determined based on factors including skills, qualifications, experience, and geographic location.

Salary range(s) for this role

 -  

Be Your Best at Snorkel

Joining Snorkel AI means becoming part of a company that has market proven solutions, robust funding, and is scaling rapidly-offering a unique combination of stability and the excitement of high growth. As a member of our team, you’ll have meaningful opportunities to shape priorities and initiatives, influence key strategic decisions, and directly impact our ongoing success. Whether you’re looking to deepen your technical expertise, explore leadership opportunities, or learn new skills across multiple functions, you’re fully supported in building your career in an environment designed for growth, learning, and shared success.

Snorkel AI is proud to be an Equal Employment Opportunity employer and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. Snorkel AI embraces diversity and provides equal employment opportunities to all employees and applicants for employment. Snorkel AI prohibits discrimination and harassment of any type on the basis of race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local law. All employment is decided on the basis of qualifications, performance, merit, and business need.

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.

Be Your Best at Snorkel

Joining Snorkel AI means becoming part of a company that has market proven solutions, robust funding, and is scaling rapidly-offering a unique combination of stability and the excitement of high growth. As a member of our team, you’ll have meaningful opportunities to shape priorities and initiatives, influence key strategic decisions, and directly impact our ongoing success. Whether you’re looking to deepen your technical expertise, explore leadership opportunities, or learn new skills across multiple functions, you’re fully supported in building your career in an environment designed for growth, learning, and shared success.

Snorkel AI is proud to be an Equal Employment Opportunity employer and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. Snorkel AI embraces diversity and provides equal employment opportunities to all employees and applicants for employment. Snorkel AI prohibits discrimination and harassment of any type on the basis of race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local law. All employment is decided on the basis of qualifications, performance, merit, and business need.

We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
489,618 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
Software Developer 1 day ago
$37k – $98k per year (Estimated) • In office • Full-Time • Tokyo
Python
AI/ML
AI Agents
DevOps
Docker
Kubernetes
Apply
$36k – $95k per year (Estimated) • Remote • Full-Time
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Framer Motion
DevOps
GitHub
Apply
Remote • Full-Time
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Framer Motion
DevOps
GitHub
Apply
$73k – $194k per year (Estimated) • Remote • Full-Time
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Framer Motion
DevOps
GitHub
Apply
$76k – $197k per year (Estimated) • Remote • Full-Time
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Framer Motion
DevOps
GitHub
Apply
$110k – $150k per year • Equity • Remote • 2+ years exp • New York
AI/ML
RAG
Snorkel
Management
Jira
Apply
$220k – $300k per year • Remote/Hybrid • New York
Python
Databases
ClickHouse
AI/ML
Dagster
dbt
Prefect
Ray
Snorkel
DevOps
Rest API
Terraform
OpenTelemetry
CI/CD
AWS
Kubernetes
Amazon EKS
Amazon S3
IAM
Amazon EventBridge
Cybersecurity
SOC 2
HIPAA
FedRAMP
Analytics
ETL/ELT
Apply
$150k – $220k per year • Remote/Hybrid • 2+ years exp • New York
AI/ML
Snorkel
DevOps
Terraform
GCP
Azure
CI/CD
AWS
IAM
Cybersecurity
SOC 2
FedRAMP
Zero Trust
Management
Jira
Apply
$200k – $320k per year • Remote/Hybrid • 5+ years exp • New York
AI/ML
Snorkel
Apply
$200k – $350k per year • Remote/Hybrid • PhD • New York
AI/ML
NLP
LLM
Snorkel
Apply
$101k – $214k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Malvern • New York • Washington
Apply
$83k – $120k per year • Equity • Remote/Hybrid • Full-Time • 8+ years exp • Chicago • New York
Apply
$82k – $110k per year • Remote/Hybrid • Full-Time • New York
Apply
$40k – $50k per year • In office • Part-Time • 1+ year exp • High School Diploma • New York
Apply
Product Manager 2 hours ago
$200k – $250k per year • In office • Full-Time • New York
Apply
See all jobs
This is one of many
489,618 more open roles from verified company boards, updated every day.