368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$170k – $260k per year
Location
In office (San Francisco)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Sesame AI is a company founded in 2022 that develops conversational speech models with unusually natural timing and prosody. Its demonstration voices drew wide attention for sounding present in a conversation rather than reading text aloud, and it released a base speech model openly. The company is also building lightweight eyewear intended to keep an assistant available throughout the day.

About Sesame

Sesame believes in a future where computers are lifelike - with the ability to see, hear, and collaborate with us in ways that feel natural and human. With this vision, we're designing a new kind of computer, focused on making voice agents part of our daily lives. Our team brings together founders from Oculus and Ubiquity6, alongside proven leaders from Meta, Google, and Apple, with deep expertise spanning hardware and software. Join us in shaping a future where computers truly come alive.

About the Role

We're looking for a Data Engineer to build and maintain the data pipelines that feed Sesame's AI models. You'll collaborate directly with machine learning engineers and researchers - your job is to make sure they have the right data, in the right shape, at the right time to train, evaluate, and ship models.

Sesame's data is rich and complex: conversations, voice, sensor signals, and product telemetry. You'll design the systems that take raw, unstructured, multimodal data and turn it into clean, versioned, well-documented datasets that ML teams can trust and build on confidently.

This is a deeply technical, infrastructure-focused role - closer to ML engineering than traditional data analytics. You'll be deeply embedded with ML teams, understanding their workflows and building infrastructure that accelerates the full model development lifecycle - from data collection and labeling through training and evaluation.

Responsibilities:

  • Design and build production data pipelines that prepare conversational, voice, and multimodal data for model training and evaluation.

  • Partner directly with ML engineers to understand data requirements for new models and experiments, and deliver datasets that meet those needs.

  • Build and maintain infrastructure for dataset versioning, lineage tracking, and reproducibility - so any training run can be traced back to its exact data.

  • Develop data quality frameworks that catch issues before they become model quality issues: schema validation, drift detection, and coverage monitoring.

  • Optimise large-scale data processing for cost and performance across Sesame's cloud infrastructure.

  • Build tooling that makes it easy for ML engineers and researchers to discover, explore, and request data independently.

  • Define and enforce data governance and privacy standards, particularly around sensitive conversational and voice data.

  • Contribute to architecture decisions around Sesame's broader data platform as the team and data volume grow.

Required Qualifications:

  • 5+ years in data engineering, with meaningful experience supporting ML or AI teams specifically.

  • Strong SQL and Python skills - you'll use both daily.

  • Experience building and operating ETL/ELT pipelines at scale using modern data platforms and tooling.

  • Experience with workflow orchestration systems such as Airflow, Dagster, or Prefect.

  • Hands-on experience with ML data workflows: training data pipelines, dataset versioning, data labeling pipelines, or model evaluation data.

  • A solid understanding of how ML teams work - you don't need to train models; what matters is understanding what makes a good training dataset and why data quality directly affects model performance.

  • Comfort working with unstructured and semi-structured data - audio, text, JSON logs - not just clean relational tables.

  • Strong communication skills. You'll be embedded with ML engineers and need to bridge data systems and model requirements effectively.

Preferred Qualifications:

  • Vector databases, embedding storage, or feature stores.

  • Data from hardware or embedded systems: telemetry, sensors, real-time streams.

  • Distributed compute frameworks for large-scale data processing such as Ray or Spark.

  • Kubernetes and managed Kubernetes environments such as GKE or EKS.

  • Data privacy frameworks, especially around voice or conversational data.

  • Building internal tooling or self-serve data platforms.

Sesame is committed to a workplace where everyone feels valued, respected, and empowered. We welcome all qualified applicants, embracing diversity in race, gender, identity, orientation, ability, and more. We provide reasonable accommodations for applicants with disabilities. Contact [email protected] for assistance.

Full-time Employee Benefits:

  • 401 (k) max employer match: 3.5% of compensation

  • 100% employer-paid health, vision, and dental benefits for you and your dependents

  • Unlimited PTO and sick time

  • Flexible spending account with employer matching up to $1,650/year (medical FSA)

  • Guardian Employee Assistance Program (EAP)

  • Opportunity to share in the company's success with competitive stock options

Benefits do not apply to contingent/contract workers.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$20k – $49k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
Ruby
SQL
Databases
Amazon Neptune
Neo4j
AI/ML
Hallucination
LangChain
LangGraph
LLM
Model Context Protocol
Spark
AI Agents
LLM Guardrails
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
GitHub Actions
GitLab CI
Jenkins
Kubernetes
Rest API
Terraform
GitHub
GitLab
QA
Playwright
Postman
Selenium
Swagger
Apply
$68k – $147k per year (Estimated) • Remote • Full-Time • 5+ years exp • Bachelor's Degree
Python
SQL
Databases
Google BigQuery
Snowflake
Analytics
Tableau
Apply
$14k – $30k per year (Estimated) • Remote • Full-Time • 2+ years exp
SQL
Management
Confluence
Apply
Remote • Full-Time • Bachelor's Degree
Python
Apply
$27k – $53k per year (Estimated) • Remote • Full-Time • 6+ years exp
SQL
Databases
Oracle
Analytics
Power BI
Tableau
Marketing
Salesforce
Apply
$175k – $280k per year • Equity • In office • Full-Time • San Francisco • New York • Bellevue
Python
AI/ML
AI Agents
DevOps
GCP
gRPC
Kubernetes
WebSockets
Apply
$175k – $280k per year • Equity • In office • Full-Time • 7+ years exp • San Francisco • New York • Bellevue
Mobile
Espresso
Firebase
Maestro
XCTest
React Native
DevOps
CI/CD
Datadog
QA
Appium
Detox
TestRail
XCUITest
Apply
Research Scientist 1 month ago
$190k – $320k per year • Equity • In office • Full-Time • Bachelor's Degree • San Francisco • New York • Bellevue
AI/ML
Computer Vision
NLP
Apply
$150k – $260k per year • Equity • In office • Full-Time • San Francisco • Bellevue
Python
AI/ML
Synthetic Data
DevOps
CI/CD
Apply
Research Engineer 2 months ago
$190k – $320k per year • Equity • In office • Full-Time • San Francisco • New York • Bellevue
AI/ML
Fine-tuning
LLM
Multimodal AI
PyTorch
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$83k – $188k per year (Estimated) • In office • 2+ years exp • San Francisco
Python
AI/ML
AI Agents
LLM Guardrails
Model Context Protocol
DevOps
Terraform
Cybersecurity
Crowdstrike
GDPR
Least Privilege
Okta
SentinelOne
Management
Google Workspace
Slack
Apply
$171k – $273k per year • In office • Full-Time • 8+ years exp • PhD • San Francisco • Washington
AI/ML
A2A
Agentforce
AI Agents
Model Context Protocol
DevOps
AWS
GCP
Marketing
Salesforce
Apply
Security GRC Analyst 2 hours ago
$119k – $268k per year (Estimated) • Remote/Hybrid • 4+ years exp • Bachelor's Degree • San Francisco
AI/ML
Ignite
PyTorch
Cybersecurity
ISO 27001
NIST CSF
SOC 2
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.