368,746open jobs
9,444companies
47,506added this week
Browse all
Salary
$220k – $350k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
Cartesia is an artificial intelligence research and technology company that builds real-time, low-latency generative voice models for interactive applications. Founded by Stanford researchers, the company utilizes State Space Models (SSMs) to power its flagship text-to-speech platform, Sonic, which enables sub-second voice synthesis, instant voice cloning, and emotional expression across dozens of languages. By providing real-time speech generation and voice agent infrastructure, Cartesia allows developers and enterprises to power conversational AI systems, virtual assistants, and customer service platforms.

About Cartesia

Our mission is to architect AI that learns from and interacts with the world like humans do.

We're pioneering the model architectures that will make this possible. Our founding team met as PhDs at the Stanford AI Lab, where we invented State Space Models or SSMs, a new primitive for training efficient, large-scale foundation models. Our team combines deep expertise in model innovation and systems engineering paired with a design-minded product engineering team to build and ship cutting edge models and experiences.

We're funded by leading investors at Index Ventures and Lightspeed Venture Partners, along with Factory, Conviction, A Star, General Catalyst, SV Angel, Databricks and others. We're fortunate to have the support of many amazing advisors, and 90+ angels across many industries, including the world's foremost experts in AI.

About the Role

The New Horizons Evaluations team is reimagining how we measure progress in interactive machine intelligence. As Evaluations Lead, you will design evaluation frameworks that capture not just what models know - but how they reason, remember, and interact over time. You’ll work at the intersection of research, product, and infrastructure to develop metrics, systems, and studies that define what “intelligence” means in the next generation of AI. This role is ideal for someone who combines scientific rigor with technical execution, and who’s deeply curious about how people use - and want to use - intelligent systems. Your work will shape how Cartesia builds and evaluates frontier models, ensuring that progress isn’t measured solely by static benchmarks, but by deeper qualities like understanding, naturalness, and adaptability in real-world interaction.

Your Impact

  • Identify and define key model capabilities and behaviors that matter for next-generation model evals

  • Develop and implement new evaluation pipelines with robust statistical analysis and clear reporting

  • Partner closely with model training and research teams to embed evaluation systems directly into model development loops

  • Prototype new user studies and behavioral experiments to ground evaluations in real-world use.

What You Bring

  • Experience designing or implementing evaluation frameworks for generative models (audio, text, or multimodal)

  • Strong technical and analytical skills, ability to take open-ended research ideas and translate them into production-ready systems

  • Creativity in defining novel quantitative metrics for subjective or behavioral qualities

  • Excitement for building evaluation systems that bridge research and real-world use

  • Curiosity and rigor in equal measure; motivation driven by discovering how to measure meaningful progress in intelligent behavior

Nice-To-Haves

  • Understanding of model alignment concepts and evaluation approaches

Note: Cartesia participates in E-Verify and will provide the federal government with Form I-9 information to confirm employment eligibility after hire.

More Details

In-office policy: We’re an in-person team based out of offices in San Francisco, London and Bangalore. We love being in the office, hanging out together, and learning from each other every day.

Visa sponsorship: We provide visa sponsorship support and assess each circumstance on a case-by-case basis. However, visa sponsorship is dependent on many factors, including the role you are applying for, and the location you are going to be based, and so we can't always guarantee success. Your Recruiter will work with you to understand your visa sponsorship needs from the first call.

We ship fast. All of our work is novel and cutting edge, and execution speed is paramount. We have a high bar, and we don’t sacrifice quality or design along the way.

We support each other. We have an open & inclusive culture that’s focused on giving everyone the resources they need to succeed.

Our Benefits (US Employees Only)

Compensation Competitive base salary alongside attractive equity package.

Health Insurance Fully covered medical insurance along with dental and vision for you and your family.

Parental Leave 9 weeks paternity & 12 weeks maternity leave

401(k)

Commuter Allowance A monthly stipend to help you get to and from the office.

Flexible PTO Take as much time as you need to recharge your batteries.

Meals & Snacks Lunch, dinner and plenty of snacks, provided daily.

Your own personal Yoshi

Our Commitment to Equal Opportunity

Cartesia is an equal opportunity employer. We consider qualified applicants without regard to race, color, religion, sex, national origin, age, disability, veteran status, genetic information, or any other legally protected status.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
Remote/Hybrid • Hyderabad
Python
Python
FastAPI
Databases
Databricks
FAISS
Milvus
Pinecone
Weaviate
AI/ML
AI Agents
Computer Vision
Embeddings
Fine-tuning
Function Calling
Gemini
Hallucination
LangChain
LangGraph
Llama
LlamaIndex
LLM
LoRA
Mistral
MLFlow
Multimodal AI
Prompt Engineering
PyTorch
QLoRA
RAG
Reranking
PEFT
Anthropic
Hugging Face
LLM Guardrails
LLMOps
OpenAI
Model Context Protocol
DevOps
Azure
CI/CD
Docker
Git
Kubernetes
Vector
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$26k – $66k per year (Estimated) • In office • Full-Time • 7+ years exp • Master's Degree • Hyderabad
Python
SQL
TypeScript
Databases
OpenSearch
Snowflake
AI/ML
AI Agents
AWS Bedrock
Claude
Claude Code
Fine-tuning
Hallucination
LangChain
LLM
Model Context Protocol
Multimodal AI
Prompt Engineering
RAG
Synthetic Data
A2A
Amazon SageMaker
DevOps
AWS
CI/CD
Docker
Vector
Analytics
A/B Testing
Apply
$24k – $61k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Databases
Databricks
Microsoft Fabric
AI/ML
MLFlow
NLP
Sentiment Analysis
Time Series Forecasting
AI Agents
DevOps
CI/CD
Platform Engineering
Analytics
Power BI
ETL/ELT
Apply
$24k – $60k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Databases
Databricks
Microsoft Fabric
AI/ML
MLFlow
NLP
Sentiment Analysis
Time Series Forecasting
AI Agents
DevOps
CI/CD
Platform Engineering
Analytics
Power BI
ETL/ELT
Apply
$110k – $249k per year (Estimated) • Equity • In office • Full-Time • London
Databases
Databricks
AI/ML
Claude
Claude Code
Cursor
Apply
$180k – $350k per year • Equity • In office • Full-Time • San Francisco
Databases
Databricks
AI/ML
Fine-tuning
Multimodal AI
RLHF
Human-in-the-Loop
Post-training
Apply
$200k – $350k per year • Equity • In office • Full-Time • San Francisco
Databases
Databricks
AI/ML
Fine-tuning
Multimodal AI
Reinforcement Learning
Synthetic Data
Post-training
SFT
Apply
$160k – $210k per year • Equity • In office • Full-Time • San Francisco
Databases
Databricks
AI/ML
Structured Outputs
Text-to-Speech
DevOps
SLI/SLO/SLA
Apply
Software Engineer 25 days ago
$27k – $75k per year (Estimated) • In office • 3+ years exp • Bengaluru
Go
Python
JavaScript
Python
Django
AI/ML
Multimodal AI
Edge AI
Frontend
Next.js
React.js
DevOps
CI/CD
Platform Engineering
Apply
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 2 hours ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.