372,956open jobs
9,661companies
50,233added this week
Browse all
Salary
$192k – $389k per year (Estimated)
Location
Remote/Hybrid (San Francisco, Boston, United States)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Liquid AI is an artificial intelligence company headquartered in Boston, Massachusetts, and founded in 2023 as a spin-off from the MIT Computer Science and Artificial Intelligence Laboratory. The company builds Liquid Foundation Models, an architecture derived from liquid neural networks that aims to match transformer quality at a fraction of the memory and compute. It targets on-device and edge deployment where models must run on phones, vehicles, and embedded hardware rather than in a data center.

About Liquid AI

Spun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal memory usage, privacy, and reliability. We partner with enterprises across consumer electronics, automotive, life sciences, and financial services. We are scaling rapidly and need exceptional people to help us get there.

The Opportunity

LFM2.5-Audio is Liquid's end-to-end multimodal speech and text language model. At 1.5B parameters, it handles speech-to-speech conversation, ASR, and TTS without requiring separate components, making it uniquely suited for real-time, on-device deployment.

We're now bringing this model to enterprise customers. The core challenge: teaching audio models to understand user intents and translate them into structured tool calls. Think voice-driven function calling, where a spoken request triggers the right API, extracts the right parameters, and confirms back to the user in natural speech.

This role sits at the intersection of frontier audio models and real-world deployment. You'll own the applied post-training work that adapts LFM2.5-Audio for customer use cases end-to-end, from data generation through delivery. Unlike most roles that force a trade-off between customer impact and foundational work, this one gives you both: deep ownership over how audio models are adapted, evaluated, and shipped, and a direct line into the evolution of Liquid's post-training and audio stacks.

If you care about data quality, evaluation, and making models actually work in production, this is a chance to shape how applied audio AI is done at a foundation model company.

What We’re Looking For

We need someone who:

  • Takes ownership: Owns customer post-training projects end-to-end for audio workloads, from requirements through delivery and evaluation.

  • Thinks end-to-end: Can reason across audio data pipelines, speech-text alignment, model adaptation, and evaluation as a connected system.

  • Is pragmatic: Optimizes for model quality and customer outcomes over publications or theory.

  • Thrives under constraints: On-device, low-latency, memory-limited audio systems excite you. You see constraints as design parameters, not blockers.

The Work

  • Act as the technical owner for enterprise audio post-training engagements.

  • Translate customer requirements into concrete post-training specifications and workflows for LFM2.5-Audio and future audio models.

  • Design and build function calling capabilities for audio models: training models to map spoken user intents to structured tool calls (API invocations, parameter extraction, confirmation flows).

  • Design and execute data generation pipelines for speech-to-speech and text-to-text training, including synthetic dialogue, function calling examples, and intent-action pairs.

  • Run supervised fine-tuning, preference alignment, and reinforcement learning workflows on audio language models.

  • Design task-specific evaluations for audio function calling (intent recognition accuracy, parameter extraction, end-to-end task completion) and feed learnings back into core post-training pipelines.

Desired Experience

Must-have:

  • Hands-on experience with post-training for language models (SFT, preference alignment, and/or RL).

  • Experience with data generation and evaluation pipelines for LLM or audio model training.

  • Strong intuition for data quality and evaluation design.

  • Familiarity with function calling, tool use, or structured output training for language models.

Nice-to-have:

  • Experience with speech or audio language models (speech-to-speech, ASR, TTS, or multimodal audio-text systems).

  • Prior exposure to customer-facing or applied ML delivery environments.

  • Experience with alignment or RL techniques beyond basic supervised fine-tuning.

  • Familiarity with on-device or low-latency inference constraints.

What Success Looks Like (Year One)

  • Independently owns and delivers enterprise audio post-training projects with minimal oversight.

  • Has built and shipped function calling capabilities that reliably translate spoken user intents into tool calls for production use cases.

  • Is trusted by customers as the technical owner, demonstrating strong judgment and delivery quality.

  • Has made durable contributions to Liquid's general-purpose post-training and audio pipelines by feeding applied learnings back into baseline model development.

What We Offer

  • Real ML work: You will fine-tune audio and speech models, build audio data pipelines, and ship solutions to enterprise customers under real-time on-device constraints.

  • Compensation: Competitive base salary with equity in a unicorn-stage company

  • Health: We pay 100% of medical, dental, and vision premiums for employees and dependents

  • Financial: 401(k) matching up to 4% of base pay

  • Time Off: Unlimited PTO plus company-wide Refill Days throughout the year

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
372,956 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$54k – $128k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Mexico City
C#
JavaScript
Python
TypeScript
C#
ASP.NET Core
Python
FastAPI
Databases
PostgreSQL
AI/ML
AI Agents
Anthropic
Embeddings
Function Calling
LLM
LLM Guardrails
Model Context Protocol
OpenAI
Semantic Search
Semantic Search
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
GitHub
GitHub Actions
Vector
Vercel
Apply
$150k – $180k per year • In office • Full-Time • PhD • New York
Python
AI/ML
Anthropic
Anthropic SDK
Computer Vision
Fine-tuning
LangChain
LlamaIndex
LLM
OpenAI
OpenAI SDK
RAG
DevOps
AWS
Azure
GCP
Apply
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
In office • Internship • Master's Degree • Austin
C++
Python
C++
PyTorch C++
AI/ML
AI Agents
CUDA
CUDA Toolkit
LLM
NCCL
PyTorch
TensorRT
TensorRT-LLM
Triton
vLLM
Apply
$18k – $45k per year (Estimated) • Remote/Hybrid • Contractor • Moscow
AI/ML
CUDA
CUDA Toolkit
cuDNN
LLM
NVLink
Ollama
PyTorch
TensorRT
vLLM
DevOps
Docker
Windows Server
Apply
$194k – $394k per year (Estimated) • Remote/Hybrid • Full-Time • Master's Degree • San Francisco
Python
AI/ML
Computer Vision
DeepSpeed
Multimodal AI
Reinforcement Learning
VLM
FSDP
Hugging Face
Megatron-LM
Post-training
SFT
DevOps
GitHub
Apply
$156k – $316k per year (Estimated) • Remote/Hybrid • Full-Time • Boston
C++
Python
AI/ML
llama.cpp
MLX ML
ONNX
Quantization
Edge AI
Apply
$67k – $146k per year (Estimated) • Remote/Hybrid • Full-Time • Tokyo
AI/ML
Fine-tuning
Knowledge Distillation
Multimodal AI
Reinforcement Learning
Synthetic Data
Post-training
SFT
Apply
$66k – $144k per year (Estimated) • Remote/Hybrid • Full-Time • Tokyo
AI/ML
Fine-tuning
llama.cpp
LLM
MLX ML
Multimodal AI
ONNX
Quantization
SGLang
vLLM
Post-training
SFT
Apply
$174k – $324k per year (Estimated) • Remote • Internship • 2+ years exp • Bachelor's Degree
AI/ML
Computer Vision
Fine-tuning
Edge AI
Function Calling
Apply
$90k – $150k per year • In office • Full-Time • 3+ years exp • Dallas • New York • San Francisco
Apply
$196k – $351k per year (Estimated) • Equity • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
AI/ML
Claude
Claude Code
Cursor
Design
Figma
Apply
$405k per year • In office • 5+ years exp • Bachelor's Degree • San Francisco
AI/ML
Anthropic
Claude
LLM
Multimodal AI
DevOps
Kubernetes
Apply
$94k – $266k per year • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Boston • Milwaukee • Dallas • Columbus • Kirkland
Python
AI/ML
Claude
Claude Code
CoreWeave
CUDA
CUDA Toolkit
NCCL
DevOps
Ansible
Configuration Management
HPC
Incident Management
Kubernetes
Rest API
SLURM
Terraform
Apply
$94k – $266k per year • Remote/Hybrid • Full-Time • 5+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
Databases
Databricks
Google BigQuery
SAP HANA
Snowflake
AI/ML
Knowledge Graph
DevOps
Azure
Apply
See all jobs
This is one of many
372,956 more open roles from verified company boards, updated every day.