368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$25k – $63k per year (Estimated)
Location
Remote/Hybrid (Bengaluru, India)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
SatSure builds decision analytics from satellite imagery and other geospatial data. Its products serve agricultural lending, insurance, infrastructure and climate risk. Banks and government bodies use its crop and land intelligence.

About SatSure

SatSure is a deep-tech decision intelligence company operating at the nexus of agriculture, infrastructure, and climate action. We turn earth observation data into actionable insights for governments, financial institutions, and enterprises across the developing world - at scale, with reliability.

Our platform team owns the infrastructure backbone that powers SatSure's AI/ML products: multi-cloud Kubernetes clusters, LLM inference pipelines, geospatial data platforms, and the internal developer tooling used by every engineering team. If you care about infrastructure quality and want your work to have real-world impact, this is the role.

About the Role

We are looking for a Senior DevOps & MLOps Engineer to join our Platform & DevOps team. You will design, build, and operate cloud-native infrastructure that supports ML model serving, data pipelines, and developer platforms across AWS, GCP, and Azure. You will work closely with data science, product engineering, and security teams - and be expected to own large surface areas end-to-end.

This is a hands-on senior IC role. You will architect systems, write Terraform and Helm, debug production incidents, define SLOs, and contribute to platform standards adopted org-wide.

Roles & Responsibilities

ML Platform & LLM Infrastructure

  • Own and operate Kubernetes-based ML platform on EKS - supporting LLM inference (KServe), distributed compute (Dask/Ray), and workflow orchestration (Apache Airflow).
  • Partner with data science and ML teams to design, deploy, and scale ML workloads - including GPU scheduling, autoscaling, resource isolation, and SLO-driven reliability.
  • Architect, deploy, and optimize Ray clusters on Kubernetes for distributed ML workloads - enabling scalable training, batch inference, and low-latency serving with efficient CPU/GPU utilization.

Multi-Cloud Platform & Infrastructure

  • Design, build, and maintain cloud-native infrastructure across AWS (primary), GCP, and Azure - using Kubernetes (EKS / GKE / AKS), Terraform, Helm, and ArgoCD.
  • Drive GitOps adoption and platform standardization - define reusable infrastructure patterns, Helm charts, and deployment workflows used across all product teams.
  • Manage Kubernetes platform operations - cluster lifecycle, Karpenter-based autoscaling, multi-tenancy, and workload isolation for data science and engineering teams.
  • Implement and maintain service mesh (Istio) - mTLS enforcement, traffic policies, and observability for inter-service communication.
  • Maintain and improve the internal developer platform (Backstage IDP) - enabling self-service environments, service catalog, and onboarding workflows for engineering teams.

Observability & Reliability Engineering

  • Build and maintain full-stack observability infrastructure - metrics (Prometheus / Mimir), logs (Loki), traces (Tempo), and dashboards (Grafana) integrated with OpenTelemetry instrumentation.
  • Define SLIs, SLOs, and error budget policies for production ML and platform services; lead incident response and post-mortem reviews.
  • Proactively identify reliability risks and drive engineering improvements to maintain 99.9%+ uptime targets.

FinOps & Cost Engineering

  • Implement Kubernetes cost attribution and chargeback using Kubecost / OpenCost - driving per-team visibility and FinOps decision-making for AI infrastructure.
  • Continuously optimize cloud spend through workload right-sizing, spot/preemptible usage, and resource scheduling strategies.

Platform Security & Governance

  • Manage AWS multi-account governance using Control Tower, SCPs, GuardDuty, and IAM Identity Center - ensuring security posture across all environments.
  • Own OIDC identity and SSO infrastructure integrated across internal tooling - Backstage, Airflow, and platform services.
  • Support compliance and audit processes - ISO 27001, CIS Benchmarks, Well-Architected Reviews, and VAPT assessments.

Requirements

Must Have

  • 5+ years of hands-on platform, DevOps, or SRE experience in production environments.
  • Strong Kubernetes expertise - cluster operations, Helm, RBAC, autoscaling (Karpenter / Cluster Autoscaler), multi-tenancy; EKS experience preferred.
  • Infrastructure as Code - Terraform (advanced), Ansible; experience managing large, multi-environment IaC codebases.
  • AWS expertise - EC2, EKS, S3, RDS, IAM, VPC, CloudWatch, Control Tower, GuardDuty; GCP or Azure exposure is a plus.
  • GitOps & CI/CD - ArgoCD, Bitbucket Pipelines / Jenkins, GitOps workflows at team scale.
  • Observability - hands-on with Prometheus, Grafana, and at least one of: Loki, Tempo, OpenTelemetry, Datadog, or ELK.
  • Scripting & automation - Python and Bash for tooling, automation, and platform integrations.
  • Strong understanding of networking, security, and cloud cost management in Kubernetes environments.

Nice to Have

  • Experience with ML serving infrastructure - KServe, vLLM, Ray Serve, or similar model serving frameworks.
  • Experience with Apache Airflow, Dask, or other data/ML pipeline orchestration at scale.
  • Familiarity with Backstage or similar internal developer platforms (IDP).
  • Istio or Envoy service mesh experience.
  • FinOps tooling - Kubecost, OpenCost, or cloud provider cost management tools.
  • OIDC / identity provider experience (Zitadel, Keycloak, or similar).
  • AWS Certified Solutions Architect or equivalent cloud certification.
  • Exposure to geospatial data workloads or satellite imagery pipelines.

Minimum Qualification

  • Bachelor's degree in Computer Science, Information Technology, or a related engineering discipline.

Our Stack

Kubernetes (EKS / GKE / AKS) · AWS · GCP · Azure · Terraform · Helm · ArgoCD · Istio · KServe · Apache Airflow · Dask · Backstage IDP · Prometheus · Grafana · Loki · Tempo · OpenTelemetry · Kubecost · Python · Bash

Why SatSure

  • Real Production Scale: LLM inference, geospatial data pipelines, and multi-cloud Kubernetes - not toy projects.
  • High Ownership: You architect systems end-to-end. No tickets-only culture, no hand-holding required.
  • Meaningful Impact: Your infrastructure powers products used by governments and institutions across the developing world.
  • Growth & Benefits: Learning allowances, broadband, medical insurance, best-in-class leave policy, and hybrid work from Bengaluru.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$71k – $149k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Jacksonville • King of Prussia
Go
Python
Java
Java
Flyway
Liquibase
DevOps
Amazon EKS
ArgoCD
CI/CD
Git
GitHub Actions
GitLab CI
Helm
Jenkins
JFrog Artifactory
Kubernetes
Platform Engineering
SRE
Terraform
AWS
GitHub
GitLab
Cybersecurity
SonarQube
Apply
$25k – $66k per year (Estimated) • In office • 8+ years exp • Bengaluru
Bash
Python
SQL
JavaScript
Python
Django
Flask
Databases
Amazon Aurora
DynamoDB
ElasticSearch
MySQL
PostgreSQL
Redis
Frontend
GraphQL
React.js
DevOps
Amazon ECS
Amazon EKS
Amazon S3
Ansible
API Gateway
AWS
AWS Lambda
CI/CD
Git
gRPC
Puppet
Rest API
Kubernetes
Analytics
Power BI
Tableau
QA
Swagger
Apply
$26k – $58k per year (Estimated) • In office • Internship • 3+ years exp • Bachelor's Degree • Minsk
Java
Databases
Couchbase
DevOps
Amazon EKS
ArgoCD
AWS
CI/CD
GitHub
GitHub Actions
Kubernetes
Terraform
Apply
$89k – $212k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Rehovot
Bash
Python
DevOps
Amazon CloudWatch
Amazon EKS
AWS
CI/CD
Datadog
GitLab
GitLab CI
GitOps
Grafana
IAM
Kubernetes
PagerDuty
Pulumi
Terraform
Terragrunt
VictoriaMetrics
Apply
$40k – $86k per year (Estimated) • Remote/Hybrid • Full-Time • Élancourt
Bash
PowerShell
Python
SQL
Databases
MySQL
PostgreSQL
Redis
DevOps
Amazon CloudWatch
Ansible
AWS
Azure
CentOS Stream
CI/CD
Debian
Docker
GCP
GitLab
GitLab CI
Grafana
Hyper-V
IAM
Jenkins
Kubernetes
Nagios
Prometheus
Proxmox VE
Terraform
Ubuntu
VMWare
Windows Server
Zabbix
Apply
$20k – $45k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Bengaluru
Verilog
VHDL
Apply
$21k – $59k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru
Python
JavaScript
Python
FastAPI
Databases
Apache Iceberg
PostGIS
PostgreSQL
AI/ML
Airflow
Function Calling
LLM
Prompt Engineering
RAG
Spark
AI Agents
Frontend
React.js
DevOps
AWS
CI/CD
Kubernetes
Terraform
SpaceTech
GDAL
Apply
$21k – $26k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree
Python
Python
Hypothesis
AI/ML
CUDA
CUDA Toolkit
Fine-tuning
Knowledge Distillation
KServe
MLFlow
ONNX
PyTorch
Quantization
TensorRT
Triton
Time Series Forecasting
DevOps
Amazon EC2
AWS
CI/CD
Docker
Kubernetes
Vector
Amazon S3
SpaceTech
GDAL
Apply
$26k – $68k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
JavaScript
Python
Dask
FastAPI
Databases
Amazon Redshift
Apache Kafka
Chroma
FAISS
Google BigQuery
MySQL
Pinecone
PostGIS
PostgreSQL
Snowflake
Weaviate
AI/ML
Airflow
dbt
Embeddings
Function Calling
LangChain
LlamaIndex
LLM
Prompt Engineering
RAG
Ray
Spark
Anthropic
Hugging Face
OpenAI
Frontend
GraphQL
React.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Terraform
Vector
Analytics
ETL/ELT
SpaceTech
GDAL
Apply
$8.5k – $36k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Bengaluru
Java
Python
SQL
DevOps
AWS
Azure
Bitbucket
CI/CD
GCP
GitHub Actions
Jenkins
Rest API
Shift-Left
GitHub
Cybersecurity
Shift-Left Security
Analytics
ETL/ELT
Management
Jira
QA
BrowserStack
JMeter
Playwright
Postman
Rest-Assured
Selenium
Apply
$16k – $34k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Mumbai • Bengaluru
JavaScript
PowerShell
SQL
C#
C#
.NET
Databases
Azure SQL Database
MS SQL
DevOps
Azure
Rest API
Cybersecurity
Microsoft Entra ID
QA
Postman
Swagger
Apply
$41k – $89k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bengaluru
C#
TypeScript
JavaScript
C#
.NET
Databases
Apache Kafka
AI/ML
Copilot
LLM
OpenAI
Frontend
Angular
GraphQL
DevOps
Azure
Azure AKS
Azure DevOps
CI/CD
Docker
GitHub
GitHub Actions
Grafana
Kubernetes
Prometheus
Rest API
Apply
$38k – $83k per year (Estimated) • In office • Full-Time • 12+ years exp • Bachelor's Degree • Bengaluru
Databases
Oracle
DevOps
AWS
Platform Engineering
Apply
Data Architect 2 hours ago
$38k – $91k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru • Pune
Node JS
Python
SQL
JavaScript
Databases
Databricks
MongoDB
Redis
Apply
$28k – $71k per year (Estimated) • In office • Full-Time • 5+ years exp • Bengaluru
DevOps
CI/CD
Platform Engineering
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.