371,660open jobs
9,621companies
49,388added this week
Browse all
Salary
$21k – $58k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Middle · 4+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Sarvam AI is a leading Indian artificial intelligence company focused on building full-stack sovereign generative AI infrastructure, foundational large language models (LLMs), and speech technologies tailored for India’s diverse languages and enterprise requirements.

About Sarvam

Sarvam is building the bedrock of Sovereign AI for India. The company is developing India’s full-stack sovereign AI platform, building across research, models, infrastructure and applications with a singular focus on making AI genuinely work for India. Sarvam works with leading enterprises and public institutions and is backed by Lightspeed, Peak XV, and Khosla Ventures. Sarvam partners with India’s leading brands, including Tata Capital, SBI Life, CRED, IDFC, and LIC.

About the Role

We are hiring a Backend Engineer to work across Sarvam’s Studio media platform - spanning AI dubbing, live translation, and the shared service foundation that powers all Studio products (voice cloning, stem separation, lip sync, music generation, and more). You will build and maintain production services, ML pipeline libraries, and platform SDKs that together enable multilingual media processing at scale for enterprise customers and Sarvam Studio users.

The work cuts across multiple codebases: a core ML pipeline library (ASR, translation, TTS, audio processing), production services for dubbing and live translation, and a shared platform SDK that provides common capabilities to every Studio service.

What You’ll Do

Service & Infrastructure

  • Design and optimize production FastAPI services for dubbing and live translation - multi-stage task orchestration, rate-limited scheduling, and backpressure controls for concurrent workloads

  • Build and maintain distributed worker architectures with independent scaling per pipeline stage and automatic recovery of stuck or failed tasks

  • Own the data layer - async ORM models, schema migrations, and query optimization on PostgreSQL

  • Implement real-time features - WebSocket-based job tracking for dubbing and streaming audio pipelines for live translation

  • Manage Kubernetes deployments - Helm charts, secrets management, ingress configuration, and multi-role container images

ML Pipeline & Library

  • Extend and maintain the core dubbing library across all pipeline stages: audio extraction, VAD, speech recognition, translation, QC, TTS, and final video stitching

  • Integrate and optimize ML model serving - remote inference server clients and local model inference for audio analysis and vocal separation

  • Build and improve QC orchestration - automated scoring, tempo analysis, guided normalization, and pronunciation verification

  • Design async-first pipelines with efficient concurrency patterns for CPU-bound audio processing

  • Maintain and evolve LLM integration layers for translation, QC, and pre-processing across multiple provider backends

Platform SDK & Shared Services

  • Build and maintain the shared Studio service SDK - reusable FastAPI middleware and routers for authentication, billing, workspace isolation, and input validation

  • Design media storage abstractions - upload, signed URL generation, retention policies, and cloud blob storage integration

  • Implement cross-cutting concerns: rate limiting, metering, audit trails, and request history across all Studio services

  • Build observability foundations - OpenTelemetry instrumentation, structured logging, and metrics collection shared across services

Quality & Operations

  • Maintain high test coverage with strong CI gates, parallel test execution, and thorough mocking of external services

  • nstrument services with custom metrics and structured error tracking for production observability

  • Manage CI/CD pipelines - automated testing, linting, container builds, artifact publishing, and version management

What We’re Looking For

  • 4-6 years of experience in backend engineering, with a focus on building and operating production services at scale

  • Strong proficiency in Python with hands-on experience building production FastAPI or similar async web services (non-negotiable)

  • Deep understanding of async programming - asyncio, concurrent execution patterns, and designing for high-throughput workloads

  • Experience with distributed task systems: task queues (Celery or similar), message brokers, and designing fault-tolerant job orchestration

  • Hands-on with PostgreSQL and an async ORM (SQLAlchemy preferred) - comfortable with query optimization, schema design, and migrations

  • Familiarity with audio/media processing: FFmpeg, common audio formats, and processing libraries like soundfile or librosa

  • Experience integrating ML models into production - API-based (REST/gRPC) or local inference (PyTorch, ONNX Runtime)

  • Experience building reusable libraries or SDKs - designing clean APIs, managing backward compatibility, and publishing packages for internal consumers

  • Proficiency with Docker, Kubernetes, Helm charts, and at least one major cloud platform (Azure/GCP/AWS)

  • Strong testing discipline - writing thorough tests, mocking external services, and maintaining CI/CD pipelines

Bonus Points

  • Prior experience with speech/NLP systems: ASR, TTS, or machine translation

  • Experience with ML model serving infrastructure (Triton, TorchServe, or similar)

  • Familiarity with LLM orchestration for structured output and multi-step agent workflows

  • Experience with real-time streaming - WebSocket, WebRTC, or Server-Sent Events in production

  • Familiarity with Indic languages and the nuances of multilingual content (code-mixing, transliteration, regional dialects)

  • Experience designing platform middleware - auth, billing, rate limiting, or multi-tenant isolation

  • Experience with observability tooling - Prometheus, Grafana, OpenTelemetry, or similar stacks

  • Familiarity with video processing pipelines and media localization workflows

  • Contributions to open-source backend, audio, or NLP projects

Why Sarvam?

Sarvam is a fast-moving, high talent-density team building full-stack AI for India, working on problems that push the frontiers of AI with real population-scale impact.

  • Work alongside researchers, engineers, builders, and business leaders who move fast and hold each other to a very high bar

  • High ownership and high impact, from day one

  • Everything we do is AI-first, from the way we build and ship to the way we think about problems

  • You can work on problems that could change how an entire country learns, works, and communicates

If you want to work on problems at the frontier of AI in India, Sarvam is the place to be.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
371,660 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$121k – $147k per year • Remote • Full-Time • 5+ years exp • PhD • Columbus
Java
Java
Spring Boot
Databases
Apache Kafka
DevOps
Azure
Azure AKS
Azure DevOps
Bicep
CI/CD
Dynatrace
GitLab
Kubernetes
Splunk
Terraform
Management
Jira
Apply
$87k – $130k per year • In office • Full-Time • 6+ years exp • Murray
Python
TypeScript
JavaScript
Python
Alembic
FastAPI
Pydantic
SQLAlchemy
Databases
PostgreSQL
Redis
AI/ML
Embeddings
LLM
LLM Guardrails
Ollama
RAG
Frontend
React Query
React Router
React.js
Vite
DevOps
AWS
CI/CD
Docker
Docker Compose
IAM
Terraform
Cybersecurity
FedRAMP
NIST 800-53
QA
Pytest
Apply
$122k – $200k per year • In office • Full-Time • 10+ years exp • Charlotte • New York
AI/ML
AI Agents
Context Engineering
Copilot
Hallucination
Human-in-the-Loop
Knowledge Graph
LLM
LLM Guardrails
LLMOps
Prompt Engineering
RAG
DevOps
Azure
CI/CD
GitHub
Kubernetes
Vector
Analytics
A/B Testing
Apply
HLS specialist 1 day ago
In office • Full-Time • Israel
DevOps
AWS
Azure
Apply
$60k – $108k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • United States
PowerShell
Python
DevOps
AWS
Azure
IAM
Splunk
Cybersecurity
Crowdstrike
ISO 27001
Microsoft Defender
Microsoft Entra ID
Microsoft Sentinel
NIST CSF
Qualys Cloud Platform
Apply
DevOps Engineer 4 days ago
$18k – $81k per year (Estimated) • In office • Full-Time • Bengaluru
Python
DevOps
Amazon EC2
Amazon EKS
ArgoCD
AWS
Azure
Blue-Green Deployment
CI/CD
Crossplane
GitHub Actions
GitLab CI
Grafana
Helm
Kubernetes
Kustomize
Loki
Prometheus
Terraform
GitHub
GitLab
IAM
Apply
$29k – $66k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Python
SQL
AI/ML
Fine-tuning
LLM
Multimodal AI
Speech Recognition
Text-to-Speech
Apply
Visual Designer 7 days ago
$16k – $48k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Design
Adobe After Effects
Adobe Photoshop
Blender
Figma
Apply
Motion Designer 7 days ago
$16k – $49k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
JavaScript
Frontend
Three.JS
Mobile
Lottie
Game Dev
GLSL
Houdini
Design
Adobe After Effects
Adobe Photoshop
Blender
Cinema 4D
Figma
Apply
$26k – $60k per year (Estimated) • In office • Full-Time • 4+ years exp • Delhi
AI/ML
Fine-tuning
Apply
Lead Data Analyst 18 min ago
$28k – $59k per year (Estimated) • In office • Full-Time • Bengaluru
SQL
Databases
Google BigQuery
AI/ML
AI Agents
DevOps
GCP
Analytics
A/B Testing
Apply
Quality Engineer 1 hour ago
$15k – $47k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
JavaScript
QA
Playwright
Selenium
Apply
$49k – $107k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Java
Java
Spring Boot
DevOps
AWS
Apply
$26k – $70k per year (Estimated) • In office • Full-Time • 5+ years exp • Chennai • Bengaluru
Python
Scala
Databases
Databricks
Analytics
ETL/ELT
Apply
$35k – $87k per year (Estimated) • In office • Full-Time • 13+ years exp • Bengaluru
Python
SQL
AI/ML
AI Agents
Function Calling
Human-in-the-Loop
LLM
LLM Guardrails
RAG
Apply
See all jobs
This is one of many
371,660 more open roles from verified company boards, updated every day.