654,929open jobs
38,112companies
91,902added this week
Browse all
Salary
$150k – $350k per year
Location
In office (San Francisco)
Seniority
Staff · 2+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Sieve is a video artificial intelligence infrastructure company headquartered in San Francisco, California, and founded in 2021. The company provides APIs and a serverless runtime for video understanding and generation tasks such as transcription, dubbing, object tracking, background removal, and highlight extraction. It sells to developers building video products who would otherwise assemble and host several separate models themselves.

About Us

Sieve is a multi-modal lab curating the world's highest-quality training datasets - spanning video, audio, images, text, and 3D. We combine exabyte-scale data infrastructure and novel multimodal understanding techniques that push the frontier of foundation models. Video alone makes up 80% of internet traffic, and across modalities, data has become the enabling medium powering creativity, communication, gaming, AR/VR, and robotics. Sieve exists to solve the biggest bottleneck in the growth of these applications: high-quality training data.

We partner with top AI labs and did $XXM last quarter alone, as a team of ~30 people. We also raised our Series A from Tier 1 firms such asMatrix Partners,Swift Ventures,Y Combinator, andAI Grant.

Why Now

Sieve combines access to diverse multimodal data, infrastructure to process it at scale, and close relationships with the teams building frontier models. This gives us a unique opportunity to study what makes training data effective-and turn those findings into better models and datasets.

You’ll join a small team building our research capabilities, with ownership over experiments, training systems, and the decisions those results inform.

About the Role

As a Member of Technical Staff, Applied Research at Sieve, you’ll train and evaluate multimodal models to understand how data shapes their capabilities. Your work will span video generation and audiovisual understanding, connecting advances in data curation with measurable improvements in model performance.

You’ll own the research loop end-to-end: identify a model weakness, form a hypothesis about the data or training approach that could address it, build the experiment, and evaluate the results. This includes fine-tuning and post-training models, developing reproducible training and evaluation pipelines, and running controlled experiments on data quality, composition, and supervision.

You’re likely a good fit if you enjoy moving between research and engineering: reading a paper, implementing a method, debugging a training run, and figuring out whether an apparent improvement holds up. You care about building reliable systems and producing findings that change how we source, curate, and use training data.

What You’ll Work On

  • Train and post-train models for video generation and multimodal understanding.

  • Design controlled experiments to measure how data selection, mixtures, and supervision affect model capabilities.

  • Build evaluations that reveal specific model weaknesses, using quantitative metrics and human judgment.

  • Develop reliable training infrastructure, including distributed training, efficient data loading, checkpointing, and experiment tracking.

  • Turn research findings into improvements in our data curation pipelines and products.

  • Collaborate with research and engineering teams internally and at partner labs to define meaningful problems and communicate results.

Requirements

  • 2+ years of experience in machine learning research or engineering, with hands-on experience training or fine-tuning deep learning models.

  • Strong Python and PyTorch skills, including the ability to implement, debug, and modify model training code.

  • Experience designing experiments, establishing baselines, and evaluating results critically.

  • Familiarity with modern generative or multimodal architectures, such as diffusion models or transformers.

  • Comfortable working with large datasets and diagnosing training bottlenecks, instability, and data quality issues.

  • Able to turn ambiguous research questions into concrete experiments and maintainable systems.

  • Strong communication skills and the ability to explain findings, tradeoffs, and uncertainty clearly.

Bonus

  • Experience training video, image, or audio generation models.

  • Experience with supervised fine-tuning, preference optimization, or reinforcement learning.

  • Experience with distributed training and GPU performance optimization.

  • Research publications, open-source contributions, or substantial independent ML projects.

  • Experience as an early hire at a startup.

Benefits

  • 401k + Full Health Insurance

  • Breakfast, Lunch, and Dinner covered and your choice of snacks

  • Ubers covered home

*all roles at Sieve require you to be onsite in San Francisco 5 days per week

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
654,929 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$115k – $175k per year • In office • Full-Time • Bachelor's Degree • San Jose
Python
SQL
Analytics
Tableau
Management
Agile
Apply
$195k – $361k per year • Equity • Remote • Full-Time • 7+ years exp • PhD • United States
Python
SQL
Python
pySpark
Databases
Snowflake
Databricks
Google BigQuery
BigQuery
AI/ML
Spark
dbt
Prefect
Feast
Anomaly Detection
Time Series Forecasting
Amazon SageMaker
Feature Store
Edge AI
DevOps
CI/CD
Self-Healing
Cybersecurity
GDPR
HIPAA
Apply
$36k – $92k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Athens
Python
SQL
C#
C#
.NET
Databases
Databricks
Microsoft Fabric
Mobile
Clean Architecture
DevOps
Rest API
CI/CD
Git
Analytics
ETL/ELT
Azure Data Factory
Management
Power Automate
Power Apps
Apply
$33k – $61k per year (Estimated) • Remote/Hybrid • Internship • London
Python
Apply
$146k – $319k per year • Remote/Hybrid • TS/SCI • Full-Time • 19+ years exp • Bachelor's Degree • Park
Python
Python
FastAPI
Databases
PostgreSQL
AI/ML
TensorFlow
Pandas
NumPy
PyTorch
Anomaly Detection
DevOps
Git
Docker
Kubernetes
Analytics
Matplotlib
Apply
$150k – $250k per year • In office • Full-Time • 2+ years exp • San Francisco
AI/ML
Multimodal AI
Apply
Business Development 15 days ago
$175k – $275k per year • In office • Full-Time • San Francisco
AI/ML
Reinforcement Learning
Multimodal AI
Post-training
Red Teaming
Robotics
Reinforcement Learning
Apply
$150k – $350k per year • In office • Full-Time • San Francisco
Python
AI/ML
Claude
Fine-tuning
Multimodal AI
VLM
PyTorch
Apply
$150k – $350k per year • In office • Full-Time • San Francisco
Python
AI/ML
Multimodal AI
PyTorch
Apply
$191k – $353k per year (Estimated) • In office • Full-Time • 1+ year exp • Bachelor's Degree • San Francisco
AI/ML
Multimodal AI
Human-in-the-Loop
Apply
Sr. Data Engineer 1 day ago
$120k – $160k per year • Equity • Remote/Hybrid • Full-Time • San Francisco
Python
Python
pySpark
AI/ML
Unstructured.io
Spark
Embeddings
Scikit-learn
NLP
TensorFlow
NumPy
LLM
RAG
Anomaly Detection
Context Engineering
LLM Evaluation
LLM Guardrails
Analytics
Matplotlib
ETL/ELT
Apply
$200k – $240k per year • Remote • 8+ years exp • San Francisco
Apply
AI Researcher 1 day ago
$184k – $275k per year • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Milpitas • San Francisco • Washington
Python
SQL
YARA
Databases
Snowflake
AI/ML
Cursor
Claude Code
Model Context Protocol
AI Agents
OpenAI Codex
Red Teaming
LLM Guardrails
Agentic Workflows
Tool Use
Cybersecurity
YARA
OWASP Top 10
Management
Agile
Apply
$204k – $255k per year • Equity • Remote • 12+ years exp • Bachelor's Degree • San Francisco
Python
SQL
AI/ML
LLM
Analytics
Tableau
Microsoft Excel
Management
Google Sheets
Apply
$110k – $170k per year • In office • 5+ years exp • Bachelor's Degree • San Francisco
PowerShell
DevOps
Splunk
Management
Agile
Apply
See all jobs
This is one of many
654,929 more open roles from verified company boards, updated every day.