699,757open jobs
41,095companies
105,247added this week
Browse all
Salary
$157k – $317k per year (Estimated)
Location
Remote (United States)
Employment
Full-Time
Overview
Company
Impact
Profile match
Personified Application Layers that see, hear, and respond face-to-face in real time. Build and deploy with our API, no-code PAL Maker, or enterprise Solutions. Interactive AI avatars and conversational video reimagined

About Us

Tavus is a research lab pioneering human computing. We’re building AI Humans: a new interface that closes the gap between people and machines, free from the friction of today’s systems. Our real-time human simulation models let machines see, hear, respond, and even look real-enabling meaningful, face-to-face conversations. AI Humans combine the emotional intelligence of humans with the reach and reliability of machines, making them capable, trusted agents available 24/7, in every language, on our terms.

Imagine a therapist anyone can afford. A personal trainer that adapts to your schedule. A fleet of medical assistants that can give every patient the attention they need. With Tavus, individuals, enterprises, and developers can all build AI Humans to connect, understand, and act with empathy at scale.

We’re a Series B company backed by world-class investors including Sequoia Capital, Y Combinator, and Scale Venture Partners.

Be part of shaping a future where humans and machines truly understand each other.

The Role

We're hiring a Senior Software Engineer (Infrastructure) to own the systems behind CVI, our real-time conversational product. Every live conversation between a person and a PAL runs on infrastructure your team owns. You'll take goals like uptime, latency, and cost and chase them wherever they lead, including into backend services and product code.

What you'll own

  • CVI's inference deployments. The GPU infrastructure serving live conversations across multiple providers and regions. You'll join as an early senior member of a growing infra team, working on projects like tuning the newest GPU generations and cutting cold-start and model load times so users wait less.

  • Expanding our GPU footprint. You'll bring on new providers and regions, stand up clusters on EKS, and build the routing, scheduling, and throughput needed for fast weight loading.

  • Uptime. You'll be one of the people pushing our uptime bar higher, along with the security and SOC2 work that keeps our infrastructure trustworthy.

  • Fix what you find. When you see a problem, you have the trust and the mandate to fix it or flag it. Reworking our deploy pipeline so shipping is fast and boring is exactly the kind of thing you'd take on.

What this role has shipped

  • Multi-provider, multi-region inference infrastructure: routes live conversations across GPU providers and regions, so one provider's outage never becomes a user's problem

  • CUDA optimizations for Phoenix, our video rendering model: doubled the frame rate by tracing and optimizing hot paths with our researchers

  • Parallel conversations on a single GPU: several live conversations sharing one card, multiplying what the fleet can serve

Who you are

  • You own outcomes. You don't stop where "infrastructure" ends. If the fix lives in backend code or the CVI stack, you dive in, and you don't wait for a ticket to do it.

  • You're energized by unfamiliar problems. If the next thing that matters is standing up a training deployment you've never touched, you jump in and learn on the fly.

  • You adapt as priorities evolve. In a space moving this fast, the most important thing to build can change as we learn. When it does, you adjust course without losing momentum.

  • You care about this problem. Keeping large-scale, real-time systems fast and reliable is something you think about unprompted.

Requirements

  • Hands-on GPU inference experience. You've deployed and optimized inference workloads on GPUs and know what it takes to build reliable systems on top of GPU cloud providers.

  • Kubernetes and EKS depth, including routing and scheduling. You're comfortable designing how work gets placed across a fleet, and writing the services that make it happen.

  • Deep AWS experience. You're at home spinning up new services and turning them into simple, repeatable processes others can build on.

  • A senior track record of ownership. You've set technical direction, made decisions others built on, and carried ambiguous work over the finish line. You explain complex ideas clearly, to engineers and non-engineers alike.

Nice to have

  • Experience with GCP

  • Experience with video streaming infrastructure

  • Experience with training infrastructure or LLM serving

  • Experience with SOC2 or security compliance

If you don't check every box but this sounds like the work you want to be doing, apply anyway.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
699,757 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$21k – $57k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Bengaluru
Python
JavaScript
Node JS
Node JS
Commander.js
Databases
PostgreSQL
AI/ML
AI Agents
LLM
DevOps
Terraform
GitHub Actions
OpenTelemetry
PagerDuty
Prometheus
CI/CD
Git
AWS
Kubernetes
Grafana
Platform Engineering
Self-Healing
Amazon EKS
AWS Fargate
AWS Lambda
Amazon EC2
Incident Management
SLI/SLO/SLA
Amazon S3
IAM
Amazon ECS
Amazon CloudWatch
DNS
Cybersecurity
ISO 27001
SOC 2
Least Privilege
Apply
$35k – $76k per year (Estimated) • In office • Full-Time • Bengaluru
Python
DevOps
Terraform
OpenShift
Helm
OpenTelemetry
Prometheus
AWS
Docker
Kubernetes
Grafana
Thanos
Amazon EKS
Incident Management
TCP/IP
DNS
Cybersecurity
ISO 27001
PCI DSS
SOC 2
HIPAA
Apply
In office • 6+ years exp
Python
Databases
Snowflake
Databricks
Google BigQuery
Amazon Redshift
Microsoft Fabric
BigQuery
AI/ML
dbt
Machine Learning
DevOps
Prometheus
CI/CD
AWS
AWS Lambda
Amazon S3
IAM
Analytics
ETL/ELT
Fivetran
Apply
$15k – $41k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Cape Town
PHP
PHP
Laravel
Symfony
WordPress
Databases
MySQL
PostgreSQL
DevOps
Rest API
Git
AWS
Ubuntu
Linux
SOAP
Apply
AI Specialist 3 hours ago
$18k – $53k per year (Estimated) • In office • Full-Time • Master's Degree • Johannesburg
Python
Java
SQL
AI/ML
NLP
TensorFlow
PyTorch
LLM
Hugging Face
Text-to-Speech
Machine Learning
Frontend
GraphQL
DevOps
Rest API
GCP
Azure
CI/CD
Git
AWS
Docker
Analytics
ETL/ELT
Apply
$162k – $292k per year (Estimated) • In office • Full-Time • San Francisco
Python
Python
Asyncio
AI/ML
Multimodal AI
AI Agents
DevOps
WebRTC
Apply
Vibe Growth Marketer 2 months ago
$101k – $254k per year (Estimated) • In office • Full-Time • San Francisco
Python
JavaScript
SQL
AI/ML
Cursor
Claude
AI Agents
LLM
Lovable
DevOps
Vercel
Management
Notion
Airtable
n8n
Zapier
Marketing
HubSpot
Amplitude
Apply
$66k – $144k per year (Estimated) • In office • Full-Time • 1+ year exp • San Francisco
AI/ML
Agentic Workflows
Robotics
Apollo
Marketing
HubSpot
LinkedIn
Apply
Product Designer 2 months ago
$136k – $249k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • San Francisco
Apply
$131k – $303k per year (Estimated) • Remote/Hybrid • Full-Time • San Francisco
AI/ML
Edge AI
Apply
$59k – $126k per year (Estimated) • Remote/Hybrid • Contractor • 1+ year exp • Bachelor's Degree • San Francisco
Management
Microsoft Office
Apply
$97k – $188k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • New York • Boston • San Francisco • Washington • Seattle
Management
Outlook
Microsoft Office
Apply
$133k – $338k per year • Remote • Full-Time • 12+ years exp • Associate's Degree • New York • Milwaukee • Dallas • Columbus • Kirkland
Python
Databases
Neo4j
Amazon Neptune
AI/ML
Airflow
Prompt Engineering
Multimodal AI
AI Agents
NLP
TensorFlow
PyTorch
LLM
Knowledge Graph
Machine Learning
DevOps
GCP
Azure
AWS
Analytics
ETL/ELT
Apache NiFi
Apply
$123k – $317k per year • In office • Full-Time • 10+ years exp • Master's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
Management
Agile
Marketing
Salesforce
Apply
$99k – $205k per year (Estimated) • In office • Full-Time • 3+ years exp • High School Diploma • San Francisco
DevOps
SLI/SLO/SLA
Windows
Management
ITIL
Service Desk
Apply
See all jobs
This is one of many
699,757 more open roles from verified company boards, updated every day.