693,133open jobs
40,634companies
98,921added this week
Browse all
Salary
$124k – $240k per year (Estimated)
Location
Remote/Hybrid (United States)
Seniority
Senior · 10+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Mirantis is an American infrastructure software company founded in 1999 that helps enterprises run container and cloud platforms on their own hardware rather than on public cloud. It began as an OpenStack specialist, acquired Docker Enterprise in 2019 and now sells Kubernetes distributions, container runtimes and managed services aimed at organisations with regulatory, cost or latency reasons to keep workloads in their own data centres. Headquartered in Campbell, California with a substantial engineering presence in eastern Europe, it competes with Red Hat and the cloud providers' own managed offerings.

Mirantis, an IREN company, is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment-on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy. https://www.mirantis.com/

We are looking for an experienced Senior Data Platform Engineer to own the event-streaming, transactional database, and custom connectivity backbone driving the k0rdent-ai platform - our multi-tenant control plane for enterprise GPU infrastructure. Every cluster provisioned, every GPU-hour consumed, and every tenant action produces high-cardinality metadata that must be transactionally recorded and securely isolated.

You will architect the core pipeline connecting our streaming data plane (Apache Kafka via Strimzi) to our relational data tier (Open Source PostgreSQL via CloudNativePG), which runs across both elastic Kubernetes clusters and high-performance bare metal hardware. Beyond standard administration, a substantial focus of this role is custom ecosystem engineering: deep Go (Golang) systems programming to build custom open-source clients and connectors, design specialized integration libraries, and develop proprietary middleware that bridges our database multi-tenancy with enterprise identity planes like Active Directory (AD)/LDAP.

Because this system is the transactional record of every tenant action across the platform, this role also owns its disaster-recovery posture - designing how the platform survives a full region or cluster loss, not just a single-node failure.

Main Responsibilities

  • Hybrid Database Architecture: Design, deploy, and operate high-availability Open Source PostgreSQL topologies (via CloudNativePG on Kubernetes and native Patroni-style topologies on bare metal) and distributed Apache Kafka clusters (via Strimzi), across both containerized Kubernetes environments and high-performance bare metal hardware.

  • Custom Client & Connector Development: Write custom, enterprise-grade open-source Kafka clients, standalone connector services, and utility libraries from scratch using preferably Go (Golang) to extend data capabilities where off-the-shelf tooling falls short.

  • Enterprise Identity & Data Integration: Architect and build system integrations connecting PostgreSQL authentication and row-level security (RLS) policy evaluation to enterprise Active Directory (AD), LDAP, and OIDC identity providers - via role-mapping and session-context layers (e.g., mapped Postgres roles, JWT claims consumed by RLS policies) rather than a direct connection.

  • Transactional Architecture & CDC: Scale high-throughput Change Data Capture (CDC) pipelines via Debezium and Kafka Connect. Implement resilient architectural patterns to maintain absolute data integrity between databases and topics without dual-write risk.

  • Cross-Region Resilience & Disaster Recovery: Design and operate cross-region failover and disaster-recovery orchestration for both PostgreSQL and Kafka - including replication topology (sync vs. async trade-offs), split-brain prevention via quorum/witness mechanisms, and explicit RPO/RTO targets for a full region or cluster loss, not just single-node HA.

  • PostgreSQL & Infra Tuning: Optimize PostgreSQL instances for heavy ingestion and zero-downtime operations. Tune Write-Ahead Logs (WAL), logical replication streams, connection pooling (PgBouncer), and configure underlying Kubernetes infrastructure primitives (CSI storage volumes and CNI network paths) to eliminate replication lag and unnecessary cross-node latency.

  • GitOps & Self-Service Platforming: Maintain a strictly declarative infrastructure-as-code (IaC) culture using Terraform and ArgoCD, creating self-service workflows so internal product teams can securely provision databases, topics, schemas, and ACLs through code.

  • Security & Isolation: Implement strict multi-tenant isolation, combining database-level row-level security (RLS) with CNI network policies, mTLS, and Kafka topic-level RBAC.

We don't expect any one candidate to check every box below - if your experience is strong across most of these areas, we encourage you to apply.

Must Have

Platform Seniority: 10+ years in systems, software, data, or platform engineering, with a track record of owning large-scale, business-critical data infrastructure.

  • Programming: Go (Golang) or similar engineering skills - comfortable writing clean, concurrent systems code, building custom client libraries, and interacting directly with database/network APIs.
  • Open Source PostgreSQL Expertise: 5+ years managing native PostgreSQL. Deep knowledge of database internals (WAL streams, logical replication, publications/subscriptions, and connection architectures) across both Kubernetes stateful environments and bare metal physical nodes.

  • Production PostgreSQL-on-Kubernetes: Hands-on experience running PostgreSQL on Kubernetes via a mature Postgres operator - CloudNativePG (CNPG) preferred, but equivalent experience with Zalando's postgres-operator or a comparable operator is acceptable. What matters is depth with declarative cluster lifecycle, failover, and backup/PITR management, not the specific product name.

  • Production Kafka: Strong experience managing Apache Kafka at scale on Kubernetes via a declarative operator - Strimzi preferred, but equivalent experience with Confluent for Kubernetes, Koperator, or a comparable operator is acceptable. What matters is depth with operator-managed broker lifecycle and scaling, not the specific product name.

  • Custom Data Movement & CDC: Proven experience tuning Kafka Connect and Debezium pipelines. Experience writing or contributing to open-source database connectors, Kafka client libraries, or integration frameworks.

  • Enterprise Identity Planes: Hands-on experience building integrations connecting distributed systems to an enterprise IAM/SSO provider - Active Directory, LDAP, Okta, Azure AD, or a generic OIDC provider all count; the specific product matters less than real experience integrating application-level authorization with an enterprise identity system.

  • Kubernetes & Cloud Infrastructure: Strong grasp of K8s primitives (Pod lifecycles, Operators, CSI storage layers, and CNI overlay routing), backed by production experience in AWS or GCP.

Nice to Have

  • Experience designing and operating cross-region disaster recovery and failover orchestration for stateful systems (Postgres and/or Kafka), including split-brain prevention and RPO/RTO trade-off decisions.

  • Active contributions to open-source Kafka connectors, PostgreSQL operators (e.g., CloudNativePG, Zalando), or Go-based database libraries.

  • Experience with schema governance for event streams (e.g., Karapace or Confluent Schema Registry) and compatibility-mode management for evolving event contracts.

  • Familiarity with high-cardinality telemetry ingestion using Apache Flink or Kafka Streams.

  • Familiarity with US enterprise compliance benchmarks (SOC 2, ISO 27001).

  • Hands-on knowledge of open-source Redis (caching, pub/sub, or as a lightweight data store) - a huge plus.

  • Hands-on knowledge of open-source MongoDB (document modeling, replication, sharding) - a huge plus.

Education and Experience

Bachelor's degree in Computer Science & Engineering (or a related field), or 10+ years of equivalent related experience.

What does Mirantis offer you?

  • Work with an established Silicon Valley leader in the cloud infrastructure industry;
  • Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
  • Be a part of cutting-edge, open-source innovation;
  • Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
  • Professional development and training;
  • Attend conferences and working groups;
  • Company outings, happy hours, hackathons, and tech talks;
  • Receive a competitive compensation package with a strong benefits plan.

We are a Leader for Container Management in G2 (#2 after AWS)!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
693,133 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
United States
$137k – $252k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Overland Park
Python
JavaScript
TypeScript
SQL
C#
Frontend
Angular
React.js
Angular Material
Mobile
React Native
Material Design
DevOps
Terraform
GCP
Azure DevOps
Azure
Git
AWS
Docker
Kubernetes
Management
Agile
Scrum
Apply
$73k – $136k per year • Remote/Hybrid • Secret • Full-Time • 3+ years exp • Bachelor's Degree • Arlington
Python
SQL
PowerShell
DevOps
Rest API
Terraform
Azure DevOps
Azure
CI/CD
Git
Docker
Kubernetes
Configuration Management
Cybersecurity
Least Privilege
Management
Agile
QA
Selenium
Playwright
Apply
$210k – $240k per year • In office • TS/SCI • 7+ years exp • Reston
Python
Bash
DevOps
Terraform
Ansible
OpenShift
CloudFormation
GitLab CI
CI/CD
Jenkins
AWS
Docker
Amazon EC2
GitLab
Amazon S3
IAM
Linux
Cybersecurity
Prisma Cloud
SonarQube
OWASP
Management
ServiceNow
Service Desk
Apply
$200k – $225k per year • In office • TS/SCI • 7+ years exp • Reston
Python
Bash
DevOps
Terraform
Ansible
OpenShift
CloudFormation
GitLab CI
CI/CD
Jenkins
AWS
Docker
Amazon EC2
GitLab
Amazon S3
IAM
Linux
Cybersecurity
Keycloak
Prisma Cloud
SonarQube
OWASP
Management
ServiceNow
Service Desk
Apply
$210k – $255k per year • In office • TS/SCI • 10+ years exp • Reston
Python
Bash
DevOps
Terraform
Ansible
OpenShift
CloudFormation
GitLab CI
CI/CD
Jenkins
AWS
Docker
Amazon EC2
GitLab
Amazon S3
IAM
Linux
Cybersecurity
Prisma Cloud
SonarQube
OWASP
Management
ServiceNow
Service Desk
Apply
$142k – $256k per year (Estimated) • Remote/Hybrid • Full-Time • United States
Go
AI/ML
Machine Learning
DevOps
Rest API
gRPC
Terraform
OpenTofu
GitOps
ArgoCD
AWS
Kubernetes
Platform Engineering
Apply
$58k – $107k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Barcelona
Go
Rust
SQL
Rust
Axum
SQLx
Tonic
Databases
PostgreSQL
AI/ML
InfiniBand
NVLink
Machine Learning
DevOps
Rest API
gRPC
Splunk
Cilium
OpenTelemetry
AWS
Kubernetes
Platform Engineering
KVM
KubeVirt
Linux
DNS
DHCP
BGP
Cybersecurity
Keycloak
Calico
PKI
Apply
$142k – $256k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • United States
Go
Rust
SQL
Rust
Axum
SQLx
Tonic
Databases
PostgreSQL
AI/ML
InfiniBand
NVLink
Machine Learning
DevOps
Rest API
gRPC
Cilium
OpenTelemetry
AWS
Kubernetes
Platform Engineering
KVM
KubeVirt
Linux
DNS
DHCP
BGP
Cybersecurity
Keycloak
Calico
PKI
Apply
$151k – $263k per year (Estimated) • Remote • Full-Time • 5+ years exp • United States
Go
TypeScript
AI/ML
Claude Code
Quantization
LLM
OpenAI Codex
Machine Learning
DevOps
Helm
CI/CD
AWS
Kubernetes
Platform Engineering
API Gateway
Apply
$138k – $238k per year (Estimated) • Remote • Full-Time • 5+ years exp • United States
Databases
Milvus
Qdrant
ElasticSearch
OpenSearch
AI/ML
vLLM
TensorRT
TensorRT-LLM
Triton
InfiniBand
Machine Learning
DevOps
Jaeger
Loki
OpenTelemetry
Prometheus
SLURM
AWS
Kubernetes
Grafana
Platform Engineering
Apply
General Cardiologist 3 hours ago
$82k – $220k per year (Estimated) • Remote/Hybrid • Full-Time • United States
Apply
Finance Director 3 hours ago
$178k – $329k per year (Estimated) • Remote/Hybrid • Full-Time • 7+ years exp • Bachelor's Degree • United States
AI/ML
Computer Use
Apply
$56k – $113k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • United States
Apply
$104k – $209k per year (Estimated) • Remote/Hybrid • Full-Time • 9+ years exp • Bachelor's Degree • United States
Management
Microsoft Office
Apply
$52k – $113k per year (Estimated) • Remote/Hybrid • Full-Time • 1+ year exp • United States
AI/ML
Computer Use
Apply
See all jobs
This is one of many
693,133 more open roles from verified company boards, updated every day.