1,271,664open jobs
73,728companies
209,529added this week
Browse all
Salary
≈ $22k – $55k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Middle · 3+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Sep 4, 2026.

Overview
Company
Impact
Profile match
ZenteiQ is a deep technology company headquartered in Bengaluru, India, and founded in 2022 out of research at the Indian Institute of Science. The company works on scientific machine learning, combining numerical simulation with neural networks for engineering and industrial modelling problems, and builds training programmes around those methods. It sells to industrial and research customers in India while running education initiatives that push scientific computing skills into the wider engineering workforce.

Distributed ML Infrastructure Engineer (GPU/TPU)

Full-Time (FTE) | Distributed ML Infrastructure Engineer (GPU/TPU) | Bangalore |

Experience: 3+ Years

About ZenteiQ

ZenteiQ is a deep-tech company born out of IISc Bangalore, building Scientific Intelligence

Infrastructure: physics-native AI for engineering, manufacturing, energy, mobility and national

systems. Rather than wrapping general-purpose language models, we train foundation models

(BrahmAI) from scratch to reason over thermal, electromagnetic, structural and materials

domains, and put them to work through industrial platforms (KogneX) and a talent OS for

engineers and researchers (AhamX). We're backed by the IndiaAI Mission (MeitY) and work

closely with IISc, ARTPARK and a national AI Hub Network - building the sovereign AI

infrastructure that India's engineering and industrial systems will run on.

About the Role

We are looking for a Distributed ML Infrastructure Engineer to build and optimize the large-scale

distributed training systems behind our foundation models. You will own the performance,

scalability and reliability of training infrastructure spanning multi-node GPU and TPU clusters.

This is a hands-on systems role at the core of how BrahmAI gets trained, working closely with

our research teams to turn raw compute into efficient, dependable training throughput.

What You'll Do

  • Build and maintain distributed training infrastructure for large-scale AI workloads.
  • Optimize training performance through profiling, memory, communication and compute

optimization.

  • Implement distributed training strategies including data, tensor, pipeline and sequence

parallelism.

  • Improve training throughput, scalability and reliability across multi-node clusters.
  • Develop automation, monitoring, profiling and debugging tools for production training

systems.

  • Collaborate with research and engineering teams to deliver efficient and scalable training

infrastructure.

What We're Looking For

  • 3+ years of experience in Distributed Systems, HPC or ML Infrastructure.
  • Strong proficiency in Python, Linux and shell scripting.
  • Experience with multi-node GPU/TPU clusters, CUDA, PyTorch Distributed, NCCL,

MPI/OpenMPI and distributed communication.

  • Strong understanding of GPU/TPU architecture, distributed computing, profiling and

performance optimization.

  • Experience with Docker, Kubernetes, Git and containerized development.
  • Strong debugging, profiling and performance analysis skills.

Good to Have / Bonus Points

  • JAX, XLA, MaxText, XPK, PJRT and TPU training infrastructure.
  • Triton, Pallas or custom kernel optimization.
  • Slurm, Ray, InfiniBand/RDMA and distributed storage systems.
  • Experience with TensorBoard Profiler, Nsight Systems/Compute or XProf.
  • Experience supporting large-scale LLM pretraining and distributed training at scale.

Why ZenteiQ

  • Build real physics-native foundation models from scratch, not another LLM wrapper -

deep, defensible technical work.

  • Be part of a nationally recognised mission: one of 8 startups selected under the

government's IndiaAI Mission to build a sovereign foundation model.

  • Work alongside IISc-trained scientists and researchers, in a company founded by an IISc

professor.

  • See your work land in real industry pilots across automotive, mobility, defence and

industrial R&D, with measurable impact.

  • Join a lean, high-caliber team at an early, high-ownership stage.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,271,664 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Bengaluru
≈ $17k – $35k per year (Estimated) • In office • 5+ years exp • Bengaluru
Apply
Senior AI/ML Engineer 11 hours ago
≈ $20k – $40k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Gurgaon
Python
TypeScript
Databases
PostgreSQL
pgvector
Pinecone
FAISS
AI/ML
LoRA
Fine-tuning
Embeddings
Quantization
Prompt Engineering
PEFT
Transformers
LLM
RAG
Hallucination
OpenAI
LLM Guardrails
Interpretability
DevOps
GCP
Azure
AWS
Apply
≈ $21k – $52k per year (Estimated) • In office • 3+ years exp • Bengaluru
Python
SQL
AI/ML
LangGraph
AutoGen
LangChain
LlamaIndex
Embeddings
Function Calling
AI Agents
Semantic Kernel
CrewAI
LLM
RAG
Hallucination
Reranking
Structured Outputs
DevOps
Azure
CI/CD
AWS
Cybersecurity
Least Privilege
Management
Agile
Apply
≈ $20k – $40k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Mumbai
Python
SQL
AI/ML
Embeddings
AI Agents
LLM
RAG
Hallucination
Agentic Workflows
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Platform Engineering
Apply
≈ $24k – $59k per year (Estimated) • In office • Full-Time • 4+ years exp • Bachelor's Degree • Bengaluru
Python
C++
C++
PyTorch C++
AI/ML
llama.cpp
CUDA Toolkit
Scikit-learn
PyTorch
CUDA
Triton
Edge AI
Machine Learning
DevOps
CI/CD
Apply
$100k – $200k per year • Equity 0.1–1% • In office • Full-Time • Los Angeles
Python
C++
AI/ML
Reinforcement Learning
DevOps
Git
Linux
Robotics
ROS
Inverse Kinematics
Teleoperation
Imitation Learning
Reinforcement Learning
Apply
$21k – $36k per year (gross) • Remote (likely India) • 3+ years exp • Gurgaon
Python
Go
Bash
AI/ML
Machine Learning
DevOps
Terraform
GCP
GitHub Actions
Istio
Terragrunt
Linkerd
CI/CD
ArgoCD
AWS
Docker
Kubernetes
Immutable Infrastructure
Platform Engineering
Atlantis
GitHub
IAM
Apply
$72k – $85k per year • Equity • In office • Bachelor's Degree • San Leandro
Python
Apply
Remote (Greece) • Part-Time • Athens
Python
JavaScript
SQL
Python
Django
Databases
PostgreSQL
DevOps
GCP
DigitalOcean
Heroku
Azure
AWS
Apply
≈ $34k – $81k per year (Estimated) • Hybrid • Moscow
Python
Java
SQL
AI/ML
Machine Learning
DevOps
Rest API
Apply
≈ $21k – $58k per year (Estimated) • In office • Full-Time • Bengaluru
Python
Java
Rust
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Weights & Biases
MLFlow
Vertex AI
JAX
TensorFlow
PyTorch
LLM
Ray
TPU
XLA
DevOps
Terraform
GCP
CI/CD
Kubernetes
Google GKE
Apply
AI/ML Engineers 1 month ago
≈ $14k – $35k per year (Estimated) • In office • Full-Time • 2+ years exp • Bengaluru
Python
Python
FastAPI
Databases
Redis
AI/ML
LangGraph
LangChain
Embeddings
Function Calling
AI Agents
LLM
RAG
Reranking
Structured Outputs
Knowledge Graph
Agentic Workflows
Tool Use
DevOps
Docker
Kubernetes
Apply
In office • Bengaluru
Python
AI/ML
vLLM
SGLang
TensorRT-LLM
TPU
Speculative Decoding
KV Cache
Apply
≈ $22k – $54k per year (Estimated) • In office • 3+ years exp • Bengaluru
Python
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
JAX
PyTorch
LLM
TensorBoard
Ray
TPU
NCCL
InfiniBand
DevOps
SLURM
Git
Docker
Kubernetes
HPC
Linux
Apply
Product Owner/Manager 26 days ago
In office • Full-Time • Bengaluru
Apply
AI Engineer 1 day ago
≈ $23k – $47k per year (Estimated) • In office • 8+ years exp • Bengaluru
Python
JavaScript
TypeScript
SQL
Databases
Snowflake
Weaviate
Databricks
Milvus
Pinecone
FAISS
Apache Kafka
AI/ML
LangGraph
LangChain
Spark
LlamaIndex
AI Agents
NLP
Gemini
LLM
RAG
OpenAI
Anthropic
Hugging Face
Multi-Agent Systems
Machine Learning
Frontend
React.js
DevOps
Rest API
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Apply
In office • Full-Time • 10+ years exp • Bachelor's Degree • Bengaluru
Python
Apply
≈ $37k – $84k per year (Estimated) • Hybrid • Full-Time • 18+ years exp • Master's Degree • Bengaluru
AI/ML
AI Agents
Agentforce
Management
Agile
Scrum
Waterfall
Marketing
Salesforce
Apply
PMO Lead 1 day ago
≈ $14k – $31k per year (Estimated) • In office • Full-Time • 15+ years exp • Bachelor's Degree • Bengaluru • New Delhi
Apply
≈ $20k – $45k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Bengaluru
Python
Ruby
DevOps
GCP
Azure
CI/CD
AWS
Platform Engineering
Management
Agile
Scrum
Kanban
Apply
See all jobs
This is one of many
1,271,664 more open roles from verified company boards, updated every day.