582,867open jobs
25,570companies
81,102added this week
Browse all
Salary
$34k – $80k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Staff · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

Lead the end-to-end technical development of speech models (ASR, TTS, Speech-LLM) - from architecture, training strategy, and evaluation to production deployment.You’ll act as an individual contributor and mentor, guiding a small team working on model training, synthetic data generation, active learning, and inference optimization for healthcare applications. As a Tech Lead specializing in ASR, TTS, and Speech LLM, you will spearhead the technical development of speech models. This involves everything from architectural design and training strategies to evaluation and production deployment.

This role is a blend of individual contribution and mentorship. You will guide a small team focused on model training, synthetic data generation, active learning, and inference optimization, all within the context of healthcare applications.

What You’ll Do

    * Prepare and maintain synthetic and real training datasets.

    * STT/TTS/Speech LLM and LLM training: model selection → fine-tuning → evaluation → deployment.

    * Build evaluation for clinical applications (RPM, triage, inbound/outbound).

    * Build scripts for data selection, augmentation (noise, codec, jitter), and corpus curation.

    * Fine-tune speech models using CTC/RNN-T or adapter-based recipes on multi-GPU systems.

    * Fine-tune LLMs using PEFT (LoRA/QLoRA/adapters) and preference methods such as DPO/RLHF where needed.

    * Implement evaluation pipelines to measure WER/sWER, entity F1, safety/quality metrics, and latency; automate MLflow logging.

    * Experiment with bias-aware training and context list conditioning.

    * Collaborate with backend and DevOps teams to integrate trained models into inference stacks.

    * Support creation of context biasing APIs and LM rescoring paths.

    * Assist in maintaining benchmarks versus commercial baselines (Deepgram, Elevenlabs, Cartesia, Whisper, etc.).

    * Optimize inference latency/cost for speech and LLM serving (batching, kv-cache, quantization, caching, autoscaling).

Desired Skills

    * Strong programming in Python (PyTorch, Hugging Face, NeMo, ESPnet).

    * Practical experience in audio data processing, augmentation, and ASR fine-tuning.

    * Training: SpecAugment, speed perturb, noise/RIRs, codec+PLC+jitter sims for PSTN/WebRTC.

    * Streaming ASR: Transducer/zipformer with chunked attention, frame-sync beam search, endpointing (VAD-EOU) tuning.

    * Context biasing: WFST boosts + neural re-scoring; patient/name dictionaries; session-aware bias refresh.

    * Familiarity with LoRA/QLoRA/adapters, distributed training, mixed precision.

    * Experience with LLM alignment and evaluation (SFT, DPO/RLHF, tool calling reliability, hallucination/safety checks).

    * Proficiency with evaluation frameworks: WER/sWER, Entity-F1, DER/JER, MOSNet/BVCC (TTS), PESQ/STOI (telephony), RTF/latency at P95/P99, and MLflow logging.

    * Inference/serving familiarity: vLLM/Triton, quantization, kv-cache, batching, and performance tuning.

    * Frameworks: ESPnet, SpeechBrain, NeMo, Kaldi/K2, Livekit, Pipecat, Dify.

    * Understanding of telephony speech characteristics, accents, and distortions.

    * Collaborative mindset for cross-functional work with ML-ops and QA.

Qualifications

  • M.S. / Ph.D. in Computer Science, Speech Processing, or related field.
  • 7-10 years of experience in applied ML, at least 3 in speech or multimodal AI.
  • Track record of shipping production ASR/TTS models or inference systems at scale.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
582,867 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$72k – $143k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Montreal
Python
JavaScript
Java
TypeScript
Node JS
Databases
PostgreSQL
Weaviate
pgvector
Pinecone
FAISS
AI/ML
LangGraph
Weights & Biases
LangChain
LlamaIndex
LoRA
vLLM
Vertex AI
Fine-tuning
Embeddings
Quantization
Scikit-learn
Prompt Engineering
Knowledge Distillation
Function Calling
AI Agents
Arize Phoenix
LangSmith
TensorRT
PEFT
Semantic Kernel
TensorRT-LLM
Accelerate
AWS Bedrock
Transformers
TensorFlow
PyTorch
LLM
RAG
Tokenization
Hallucination
Triton
OpenAI
Anthropic
Hugging Face
Feature Store
Human-in-the-Loop
Structured Outputs
LLM Evaluation
LLM Guardrails
Agentic Workflows
Tool Use
Model Distillation
Frontend
GraphQL
Angular
React.js
DevOps
GCP
OpenTelemetry
Azure
CI/CD
AWS
Vector
Analytics
A/B Testing
Management
ServiceNow
Marketing
Salesforce
Apply
$110k – $120k per year • Remote/Hybrid • 12+ years exp
Python
SQL
Databases
Weaviate
Databricks
Pinecone
AI/ML
Spark
MLFlow
Vertex AI
AI Agents
NLP
Kubeflow
RAG
DevOps
GCP
Apply
$142k – $298k per year (Estimated) • In office • Contractor • 8+ years exp • Durham
Python
Python
Flask
FastAPI
Django
AI/ML
Fine-tuning
AI Agents
AgentOps
DevOps
Rest API
Apply
$189k – $261k per year • Remote/Hybrid • Full-Time • 10+ years exp • Master's Degree • Louisville • New York • Dallas
Python
AI/ML
Fine-tuning
AI Agents
TensorFlow
PyTorch
LLM
RAG
DevOps
GCP
Azure
CI/CD
AWS
Apply
AI Engineer 3 hours ago
$136k – $207k per year • Remote/Hybrid • Full-Time • 4+ years exp • Master's Degree • Chicago • New York
Python
Python
FastAPI
AI/ML
LangChain
LlamaIndex
Fine-tuning
Prompt Engineering
Chain-of-Thought
AI Agents
NLP
Arize Phoenix
Cohere SDK
LangSmith
Semantic Kernel
LLM
RAG
In-Context Learning
OpenAI
Anthropic
Structured Outputs
DevOps
OpenTelemetry
Apply
$24k – $64k per year (Estimated) • In office • Full-Time • 4+ years exp • PhD • Bengaluru
Python
AI/ML
MLFlow
Triton Inference Server
Quantization
Speech Recognition
TensorRT
LLM
Triton
Text-to-Speech
Edge AI
DevOps
GCP
Prometheus
Azure
CI/CD
AWS
Docker
Kubernetes
Grafana
Apply
$37k – $93k per year (Estimated) • In office • Full-Time • 2+ years exp • Master's Degree • Bengaluru
Python
AI/ML
LoRA
MLFlow
Fine-tuning
Speech Recognition
PEFT
Kaldi
PyTorch
LLM
Whisper
Hugging Face
NVIDIA NeMo
Text-to-Speech
Deepgram
LiveKit
DevOps
WebRTC
Apply
$16k – $46k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru
Python
DevOps
Terraform
Ansible
GCP
Azure
AWS
Kubernetes
Service Mesh
OpenStack
Web3
Layer 2
Apply
$19k – $41k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Bengaluru
Management
QuickBooks
Apply
$32k – $65k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bengaluru • Chennai
Apply
Software Engineer I 4 hours ago
$8.5k – $23k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Bachelor's Degree • Mumbai • Bengaluru
JavaScript
TypeScript
C#
C#
.NET
AI/ML
Copilot
Claude Code
DevOps
Terraform
Azure DevOps
Azure
CI/CD
Git
Docker
Kubernetes
Management
Agile
QA
Playwright
Apply
Software Engineer II 4 hours ago
$13k – $32k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Mumbai • Bengaluru
JavaScript
TypeScript
C#
C#
.NET
AI/ML
Copilot
Claude Code
DevOps
Terraform
Azure DevOps
Azure
CI/CD
Docker
Kubernetes
Management
Agile
QA
Playwright
Apply
See all jobs
This is one of many
582,867 more open roles from verified company boards, updated every day.