1,265,434open jobs
73,099companies
210,448added this week
Browse all
Salary
≈ $21k – $58k per year (Estimated)
Location
In office (Bengaluru)
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 6, 2026. First seen by Alion on Sep 4, 2026.

Overview
Company
Impact
Profile match
ZenteiQ is a deep technology company headquartered in Bengaluru, India, and founded in 2022 out of research at the Indian Institute of Science. The company works on scientific machine learning, combining numerical simulation with neural networks for engineering and industrial modelling problems, and builds training programmes around those methods. It sells to industrial and research customers in India while running education initiatives that push scientific computing skills into the wider engineering workforce.

About ZenteiQ

ZenteiQ is a deep-tech company born out of IISc Bangalore, building Scientific Intelligence

Infrastructure: physics-native AI for engineering, manufacturing, energy, mobility and national

systems. Rather than wrapping general-purpose language models, we train foundation models

(BrahmAI) from scratch to reason over thermal, electromagnetic, structural and materials

domains, and put them to work through industrial platforms (KogneX) and a talent OS for

engineers and researchers (AhamX). We're backed by the IndiaAI Mission (MeitY) and work

closely with IISc, ARTPARK and a national AI Hub Network - building the sovereign AI

infrastructure that India's engineering and industrial systems will run on.

About the Role

You will build the systems that let our researchers train, evaluate, and serve large foundation

models reliably at scale. This role sits at the intersection of model research and infrastructure,

with a focus on accelerator-based (TPU) training and inference, performance, reproducibility,

and researcher velocity. You will own the paved path that turns expensive, long-running model

runs into a repeatable, observable, and cost-efficient process.

What You'll Do

  • Build and improve the paved path for distributed training, evaluation, experiment tracking,

checkpoint management, model release, and inference on Cloud TPUs.

  • Operate long-running ML workloads with strong observability, failure detection, automated

recovery, and practical operational tooling.

  • Profile and remove bottlenecks across compute, HBM and host memory, data input,

networking and collectives, XLA compilation, checkpointing, and serving.

  • Build validation, representative-scale testing, CI/CD, reproducibility, and lineage systems

that catch problems before expensive model runs or production releases.

  • Create reusable APIs, abstractions, and self-service tooling that help researchers move

quickly while preserving useful low-level controls.

  • Own accelerator capacity workflows, including TPU provisioning, quotas, reservations,

priorities, topology-aware placement, utilization, and cost efficiency.

  • Partner with research teams to debug model and systems failures, translate recurring

pain points into durable platform improvements and define operational standards.

What We're Looking For

  • Strong software engineering skills and experience owning systems used by researchers

or engineers in production or research-critical environments.

  • Hands-on experience with ML training or inference infrastructure at meaningful scale,

including large accelerator or distributed compute workloads.

  • Strong distributed-systems fundamentals and experience with Kubernetes, GKE, or

comparable orchestration for long-running compute jobs.

  • Ability to debug across model code, data pipelines, runtimes, cluster services, storage,

networking, and accelerator behaviour.

  • Strong observability, reliability, and performance-engineering fundamentals, including

incident response and root-cause analysis.

  • Production-quality Python and the ability to build durable platform abstractions rather than

one-off scripts.

  • High ownership, pragmatic judgment, and clear collaboration with fast-moving research

teams.

Good to Have / Bonus Points

  • Experience operating large Cloud TPU clusters, including TPU VMs or Pods, multislice

jobs, topology, quotas, scheduling, and failure recovery.

  • Experience with JAX, PyTorch/XLA, TensorFlow, XLA/HLO, PJRT, MaxText, Pallas, or

another large-scale training stack.

  • Experience serving LLMs on TPUs, including batching, partitioning, compilation caching,

autoscaling, latency, and throughput optimization.

  • Familiarity with GCP/GKE, Terraform, Vertex AI, XPK, Ray, Argo, Airflow, MLflow, Weights

& Biases, or comparable internal platforms.

  • Experience with model registries, distributed checkpointing, data and artifact lineage, and

large artifact distribution.

  • Proficiency in a systems language such as Go, Rust, C++, or Java.

Why ZenteiQ

  • Build real physics-native foundation models from scratch, not another LLM wrapper -

deep, defensible technical work.

  • Be part of a nationally recognised mission: one of 8 startups selected under the

government's IndiaAI Mission to build a sovereign foundation model.

  • Work alongside IISc-trained scientists and researchers, in a company founded by an IISc

professor.

  • See your work land in real industry pilots across automotive, mobility, defence and

industrial R&D, with measurable impact.

  • Join a lean, high-caliber team at an early, high-ownership stage.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,265,434 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Bengaluru
AI/ML Engineer 1 day ago
≈ $21k – $42k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Python
pySpark
Databases
Databricks
Apache Kafka
AI/ML
Spark
MLFlow
Vertex AI
Embeddings
AI Agents
NLP
Kubeflow
TensorFlow
PyTorch
RAG
Semantic Search
Anomaly Detection
Amazon SageMaker
Semantic Search
Recommender Systems
Machine Learning
DevOps
Azure
Apply
≈ $27k – $55k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Bengaluru
Python
Java
C++
Cython
C++
PyTorch C++
Cython
Bottleneck
AI/ML
ONNX
PyTorch
Recommender Systems
Machine Learning
DevOps
QEMU
QA
Rest-Assured
Apply
≈ $30k – $72k per year (Estimated) • Hybrid • 10+ years exp • India
Python
Go
Java
AI/ML
AI Agents
LLM Guardrails
Agentic Workflows
DevOps
GCP
CI/CD
AWS
Kubernetes
Platform Engineering
Incident Management
Management
Gmail
Telegram
WhatsApp
Apply
≈ $23k – $47k per year (Estimated) • In office • Full-Time • 4+ years exp • PhD • Bengaluru
Python
SQL
AI/ML
MLFlow
SHAP
Vertex AI
Embeddings
Prompt Engineering
NLP
Kubeflow
LLM
Tokenization
Anomaly Detection
Amazon SageMaker
LLMOps
Machine Learning
DevOps
GCP
GitHub Actions
Azure
CI/CD
Git
AWS
Docker
Apply
Agentic AI Engineer 4 hours ago
In office • Full-Time • Bachelor's Degree • Bengaluru
Python
C#
C#
.NET
AI/ML
Copilot
Claude Code
Model Context Protocol
AI Agents
RAG
OpenAI Codex
Copilot Studio
DevOps
Azure DevOps
OpenTelemetry
Azure
CI/CD
GitHub
Cybersecurity
Least Privilege
Management
Agile
Apply
≈ $12k – $27k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Chennai • Pune
Python
Java
SQL
C#
Mobile
JUnit
DevOps
GCP
Azure DevOps
GitLab CI
Azure
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Management
Agile
QA
TestNG
Selenium
Cucumber
JMeter
Swagger
Postman
Pytest
Apply
$135k – $162k per year • Hybrid • TS/SCI • 4+ years exp • Master's Degree • Adelphi
Python
SQL
PowerShell
Bash
Databases
ElasticSearch
Apache Kafka
DevOps
Rest API
Kibana
Logstash
Git
AWS
Docker
Kubernetes
Linux
Windows
Cybersecurity
Elastic SIEM
Zero Trust
SIEM
Analytics
ETL/ELT
Management
Slack
Discord
Apply
$155k – $170k per year • In office • Secret • 12+ years exp • Bachelor's Degree
PowerShell
DevOps
Terraform
Ansible
OpenShift
CloudFormation
Azure
CI/CD
AWS
Docker
Kubernetes
Bicep
IAM
Cybersecurity
Okta
CyberArk
FedRAMP
Zero Trust
Microsoft Entra ID
Active Directory
Management
Slack
Confluence
Jira
Discord
Apply
In office • Full-Time • Bachelor's Degree
Python
AI/ML
Machine Learning
IoT
MQTT
OPC UA
Apply
$150k – $174k per year • In office • TS/SCI • 7+ years exp • Arlington
Python
AI/ML
Machine Learning
DevOps
Splunk
VMWare
Azure
AWS
AIOps
Incident Management
Management
Slack
ServiceNow
Discord
ITSM
Apply
≈ $22k – $55k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Python
AI/ML
CUDA Toolkit
JAX
PyTorch
LLM
TensorBoard
Ray
CUDA
Triton
TPU
NCCL
InfiniBand
XLA
DevOps
SLURM
Git
Docker
Kubernetes
HPC
Linux
Apply
AI/ML Engineers 1 month ago
≈ $14k – $35k per year (Estimated) • In office • Full-Time • 2+ years exp • Bengaluru
Python
Python
FastAPI
Databases
Redis
AI/ML
LangGraph
LangChain
Embeddings
Function Calling
AI Agents
LLM
RAG
Reranking
Structured Outputs
Knowledge Graph
Agentic Workflows
Tool Use
DevOps
Docker
Kubernetes
Apply
In office • Bengaluru
Python
AI/ML
vLLM
SGLang
TensorRT-LLM
TPU
Speculative Decoding
KV Cache
Apply
≈ $22k – $54k per year (Estimated) • In office • 3+ years exp • Bengaluru
Python
C++
C++
PyTorch C++
AI/ML
CUDA Toolkit
JAX
PyTorch
LLM
TensorBoard
Ray
TPU
NCCL
InfiniBand
DevOps
SLURM
Git
Docker
Kubernetes
HPC
Linux
Apply
Product Owner/Manager 26 days ago
In office • Full-Time • Bengaluru
Apply
AI Engineer 1 day ago
≈ $23k – $47k per year (Estimated) • In office • 8+ years exp • Bengaluru
Python
JavaScript
TypeScript
SQL
Databases
Snowflake
Weaviate
Databricks
Milvus
Pinecone
FAISS
Apache Kafka
AI/ML
LangGraph
LangChain
Spark
LlamaIndex
AI Agents
NLP
Gemini
LLM
RAG
OpenAI
Anthropic
Hugging Face
Multi-Agent Systems
Machine Learning
Frontend
React.js
DevOps
Rest API
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Apply
In office • Full-Time • 10+ years exp • Bachelor's Degree • Bengaluru
Python
Apply
≈ $37k – $84k per year (Estimated) • Hybrid • Full-Time • 18+ years exp • Master's Degree • Bengaluru
AI/ML
AI Agents
Agentforce
Management
Agile
Scrum
Waterfall
Marketing
Salesforce
Apply
PMO Lead 12 hours ago
≈ $14k – $31k per year (Estimated) • In office • Full-Time • 15+ years exp • Bachelor's Degree • Bengaluru • New Delhi
Apply
Senior DevOps Engineer 12 hours ago
≈ $20k – $45k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Bengaluru
Python
Ruby
DevOps
GCP
Azure
CI/CD
AWS
Platform Engineering
Management
Agile
Scrum
Kanban
Apply
See all jobs
This is one of many
1,265,434 more open roles from verified company boards, updated every day.