368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$200k – $220k per year
Location
In office
Employment
Full-Time
Overview
Company
Impact
Profile match
Cantina is a social artificial intelligence technology company based in San Francisco, California, and founded in 2023. The company provides a platform where users can create, interact with, and share multimodal AI characters that feature unique personalities and the ability to generate video content. It operates primarily through a mobile application and web interface, focusing on the intersection of generative AI and social media for a global audience of creators and consumers.

About Cantina:

Cantina Labs is a social AI company, developing a suite of advanced real-time models that push the boundaries of expression, personality, and realism. We bring characters to life, transforming how people tell stories, connect, and create. We build and power ecosystems. Cantina, our flagship social AI platform, is just the beginning.

If you're excited about the potential AI has to shape human creativity and social interactions, join us in building the future!

About the Role:

We are seeking an experienced Machine Learning Engineer (MLE) to focus on audio model evaluation, specifically for speech generation and recognition models.

This role involves designing and developing comprehensive model evaluation pipelines for both development and production environments, as well as creating automated dashboards for reporting evaluation results.

As the founding member of our evaluation team, the ideal candidate is expected to leverage their experience to lead our evaluation efforts and play a key role in the future growth of the evaluation team.

What You’ll Do:

  • Designing model evaluation pipelines for models in development and production

  • Designing user studies for subjective model evaluations.

  • Converting requirements into measurable metrics.

  • Designing and developing automated evaluation dashboard to see model performances and compare results.

  • Training new models to capture new and different evaluation metrics.

  • Communicating with the model team to help design better models based on the evaluation results.

  • Communicating with the data team to help decide the type of data necessary to improve model performance.

  • Communication with the product-manager to make sure product requirements are correctly measured.

  • Help grow the evaluation team as the founding member.

  • Lead the evaluation team in the future.

What You’ll Bring:

  • Strong experience and intuition for designing metrics that capture model performance.

  • Strong experience with designing user studies on Mechanical Turk or similar platforms. .

  • Strong experience with model training and fine-tuning for model evaluation.

  • Strong statistical knowledge and experience to statistically compare evaluation results and take decisions.

  • Very strong engineering and programming skills.

  • Experience with training ASR, TTS models.

  • Experience at ML teams working on large-scale machine learning problems. (>3B models with >1m hours of data)

Compensation:

The anticipated annual base salary range for this role is between $200,000-$220,000 (€170,000-€190,000). When determining compensation, a number of factors will be considered, including skills, experience, job scope, location, and competitive compensation market data.

Benefits for U.S.-based roles:

  • Competitive salary and generous company equity

  • Medical, dental, and vision insurance - 99.99% of premiums covered by Cantina

  • 42 days of paid time off, including:

    • 15 PTO days

    • 10 sick days

    • 15 company holidays

    • 2 floating holidays

  • Generous parental leave & fertility support

  • 401(k) retirement savings plan

  • Lifestyle spending account - $500/month to use however you’d like

  • Complimentary lunch and snacks for in-office employees

  • One Medical membership, and more!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$140k – $225k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Seattle
TypeScript
Python
Python
FastAPI
Pydantic
Databases
pgvector
PostgreSQL
Redis
AI/ML
Fine-tuning
LangGraph
LLM
RAG
LangChain
AutoGen
CrewAI
Knowledge Distillation
NLP
Reranking
Human-in-the-Loop
LLM Evaluation
LLM Guardrails
OCR
Structured Outputs
AI Agents
Function Calling
Speech Recognition
DevOps
Docker
Terraform
AWS
Azure
Bicep
Robotics
Digital Twin
Marketing
Salesforce
Apply
$300k – $475k per year • Remote/Hybrid • Full-Time • 4+ years exp • San Francisco
Python
Rust
TypeScript
JavaScript
AI/ML
Fine-tuning
Frontend
React.js
Apply
Founding Engineer 1 day ago
$93k – $140k per year • In office • Full-Time • 3+ years exp • Munich
Python
Python
FastAPI
AI/ML
Fine-tuning
LLM
VLM
Apply
$120k – $250k per year • Equity 0.2–1% • In office • Full-Time • Master's Degree • Seattle
Python
AI/ML
Diffusion Models
Fine-tuning
PyTorch
Self-Supervised Learning
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$175k – $250k per year • In office • Full-Time • 6+ years exp • San Francisco
SQL
Mobile
Braze
Deep Linking
Analytics
A/B Testing
Marketing
Amplitude
Iterable
Mixpanel
Apply
$180k – $270k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • Sunnyvale • San Francisco
C++
Go
Node JS
JavaScript
AI/ML
Speech Recognition
DevOps
WebRTC
Apply
$200k – $220k per year • In office • Full-Time
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
DeepSpeed
Knowledge Distillation
Multimodal AI
PyTorch
Synthetic Data
Tokenization
DPO
FSDP
GRPO
LLM Guardrails
Text-to-Speech
Speech Recognition
Apply
$200k – $240k per year • In office • Full-Time • 8+ years exp
Go
DevOps
AWS
Terraform
Apply
$200k – $220k per year • In office • Full-Time
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
DeepSpeed
Knowledge Distillation
Multimodal AI
PyTorch
Quantization
Synthetic Data
Tokenization
DPO
FSDP
GRPO
LLM Guardrails
Text-to-Speech
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.