We are looking for a Technical Lead to spearhead our AI engineering efforts and drive product innovation. We are pioneering the AI revolution as India's first scaled AI services company. We deliver cutting-edge AI and LLM solutions tailored for the dynamic needs of the modern world. Backed by robust funding, we are a vibrant, young team set on redefining technological boundaries. Join us in shaping the future. Join us in building the future, where every interaction is smarter, faster, and more impactful. About the Role: As a machine learning engineer at Nurix AI, you will play a pivotal role in solving some of the hardest challenges. You'll work on building robust agentic systems, designing evals, and managing model deployment at scale. This role is ideal for engineers who want to push the boundaries of speech, LLMs, and agentic AI in production at scale.
The core responsibilities for the job include the following:
Speech and Voice Systems:
- Own evaluation, selection, and integration of ASR/LLM/TTS/S2S model providers.
- Benchmark accuracy, latency, and noise robustness across language and code-switched settings, including noise cancellation and audio preprocessing variants.
- Build and optimize real-time streaming voice pipelines that handle interruptions, endpointing, natural pauses, and real-world audio variability without breakdowns.
Model Quality and Inference:
- Analyze conversational quality of voice agents against human benchmarks and translate findings into prioritized model improvement roadmaps (fine-tuning, prompting, and harness engineering).
- Design SFT datasets and lightweight eval frameworks for voice-agent quality at scale.
- Optimize LLM/ASR/TTS inference for in-house model deployments including memory/KV-cache planning, GPU selection and sizing, latency-throughput trade-offs for real-time voice workloads.
- Harden prompt, harness, and tool-calling reliability for production agents.
- Work on self-learning agents capable of continuous improvement by observing human workflows and feedback.
Engineering and Deployment:
- Collaborate with MLOps and Inference Providers to bring models into real-time, low-latency production environments.
- Work with DataOps to create, refine, and build high-quality real-world datasets for in-house data needs.
- Ensure scalability, security, and reliability of deployed ML systems.
Collaboration and Research:
- Partner closely with product and engineering teams to deliver production-ready features.
- Stay on top of the latest advances in ASR, TTS, LLMs, speech systems, and real-time inference frameworks, and apply them to Nurix's roadmap.
Requirements:
- Bachelor's degree in computer science, AI/ML, or a related field.
- 4+ years of hands-on ML engineering experience, with a focus on speech/NLP systems.
- Strong expertise in ASR, TTS, NLP, LLM inference and fine-tuning, and conversational voice systems.
- Track record of training/fine-tuning models and deploying them in production.
- Worked closely with data operations teams for training/testing, data annotation, and post-release sanity checks.
- Proficiency in Python backends and ML frameworks (PyTorch, vLLM, SGLang, Unsloth, Axolotl).
- Experience in designing low-latency, real-time ML applications. Strong understanding of ML lifecycle, evaluation, and deployment practices.
- Contributions to open-source ML projects or research publications. Prior work on multilingual or Indian-context ASR/TTS systems.

