368,530open jobs
9,432companies
50,439added this week
Browse all
Location
Remote (Bulgaria)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Mirantis is a B2B open-source cloud computing, container management, and AI infrastructure company headquartered in Campbell, California. Originally known as a core contributor to OpenStack, Mirantis now provides open-cloud software and managed services centered on Kubernetes, multi-cloud platforms, and enterprise AI workloads.

Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment-on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.

Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Liberty Mutual, PayPal, Reliance Jio, Societe Generale, Splunk, and Volkswagen. Learn more at www.mirantis.com.

We are seeking a highly skilled Senior HPC Networking Engineer to design, deploy, manage, and troubleshoot high-performance networking environments. The ideal candidate will have deep expertise in InfiniBand technologies, strong general networking knowledge, and hands-on experience with Fortinet solutions. You will play a critical role in ensuring the performance, reliability, and scalability of HPC infrastructure.

Key Responsibilities:

  • Design, deploy, and maintain high-performance network infrastructures for HPC environments, with a strong focus on InfiniBand fabrics.
  • Troubleshoot complex network issues across InfiniBand and Ethernet environments, ensuring minimal downtime and optimal performance.
  • Manage and optimize InfiniBand components, including switches, HCAs, subnet managers, and fabric configurations.
  • Perform performance tuning, monitoring, and capacity planning for HPC networking systems.
  • Implement and maintain network security using Fortinet solutions (FortiGate, FortiManager, FortiAnalyzer).
  • Diagnose and resolve issues related to routing, switching, latency, and throughput across hybrid network environments.
  • Collaborate with compute, storage, and platform teams to support HPC workloads and cluster operations.
  • Develop and maintain documentation for network architecture, configurations, and operational procedures.
  • Participate in on-call rotations and provide escalation support for critical incidents.
  • Lead or contribute to network upgrades, migrations, and new deployments.

Required:

  • 5+ years of experience in network engineering, with a focus on HPC or data center environments.
  • Strong hands-on experience with InfiniBand technologies (e.g., Mellanox/NVIDIA).
  • Solid understanding of networking fundamentals: TCP/IP, routing protocols (BGP, OSPF), VLANs, QoS, and network design.
  • Proven experience deploying and troubleshooting Fortinet solutions (FortiGate, FortiManager, VPNs, firewall policies).
  • Experience with network performance analysis and troubleshooting tools.
  • Familiarity with Linux systems and scripting for automation (e.g., Bash, Python).
  • Strong analytical and problem-solving skills.

Preferred:

  • Experience with large-scale HPC clusters or AI/ML infrastructure.
  • Knowledge of RDMA, MPI, and low-latency networking concepts.
  • Certifications such as FCSS/FCNSP (Fortinet), CCNP/CCIE, or equivalent.
  • Experience with automation and Infrastructure as Code tools (e.g., Ansible, Terraform).

Soft Skills:

  • Strong communication and collaboration skills.
  • Ability to work independently and handle complex technical challenges.
  • Detail-oriented with a proactive approach to problem-solving.

What We Offer:

  • Operate some of the most advanced AI infrastructure environments in production today.
  • Work with the latest NVIDIA GPU technologies, Kubernetes platforms, and high-performance networking environments.
  • Help define operational standards and reliability practices for next-generation AI infrastructure services.
  • Influence the adoption of AI-powered operational capabilities through k0rdent AI.
  • Work alongside highly skilled engineers solving complex infrastructure and platform challenges at scale.
  • Join a growing organisation investing heavily in AI infrastructure, platform services, and operational innovation.

We are a Leader for Container Management in G2 (#2 after AWS)!

We are a Leader for Container Management in G2 (#2 after AWS)!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Sofia
$105k – $175k per year • In office • Full-Time • Raleigh
Python
AI/ML
AI Agents
Amazon SageMaker
AutoGen
AWS Bedrock
CrewAI
Edge AI
Function Calling
LangChain
LangGraph
Model Context Protocol
NLP
Prompt Engineering
PyTorch
RAG
TensorFlow
DevOps
Amazon EC2
Amazon ECS
Amazon EKS
Amazon S3
AWS
AWS Lambda
Azure
GCP
Git
Kubernetes
Apply
$44k – $58k per year • Remote/Hybrid • Full-Time • 2+ years exp • Associate's Degree • Vancouver
Bash
Python
SQL
Databases
MS SQL
MySQL
Oracle
PostgreSQL
DevOps
AWS
Azure
GCP
VMWare
Windows Server
Apply
$70k – $150k per year • In office • Full-Time • 5+ years exp • Calgary
Java
Node JS
Python
SQL
TypeScript
JavaScript
Java
Spring Boot
Databases
Apache Kafka
Chroma
OpenSearch
pgvector
Pinecone
PostgreSQL
AI/ML
AWS Bedrock
Copilot
Embeddings
LangChain
LangGraph
LLM
Prompt Engineering
RAG
AI Agents
Anthropic
Function Calling
Human-in-the-Loop
OpenAI
Structured Outputs
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
Docker
Kubernetes
Vector
Apply
$59k – $99k per year • In office • Full-Time • Bachelor's Degree • Boca Raton
C++
JavaScript
Python
SQL
C#
C#
.NET
DevOps
AWS
Azure
Apply
$59k – $99k per year • In office • Full-Time • Bachelor's Degree • Alpharetta
C++
JavaScript
Python
SQL
C#
C#
.NET
DevOps
AWS
Azure
Apply
$72k – $114k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Warsaw
Go
Rust
SQL
Rust
Axum
SQLx
Tonic
Databases
PostgreSQL
AI/ML
InfiniBand
NVLink
DevOps
AWS
Cilium
gRPC
Kubernetes
KubeVirt
KVM
OpenTelemetry
Platform Engineering
Rest API
Splunk
Cybersecurity
Calico
Keycloak
Apply
Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Prague
Go
Rust
SQL
Rust
Axum
SQLx
Tonic
Databases
PostgreSQL
AI/ML
InfiniBand
NVLink
DevOps
AWS
Cilium
gRPC
Kubernetes
KubeVirt
KVM
OpenTelemetry
Platform Engineering
Rest API
Splunk
Cybersecurity
Calico
Keycloak
Apply
Remote/Hybrid • Full-Time • Riga
Go
Python
Rust
Databases
ElasticSearch
OpenSearch
AI/ML
Anomaly Detection
DevOps
AIOps
AWS
Cortex
eBPF
Grafana
Incident Management
Jaeger
Kubernetes
Loki
Mimir
OpenTelemetry
Platform Engineering
Prometheus
SLI/SLO/SLA
Splunk
Thanos
HPC
Apply
Remote/Hybrid • Full-Time • Sofia
Go
Python
Rust
Databases
ElasticSearch
OpenSearch
AI/ML
Anomaly Detection
DevOps
AIOps
AWS
Cortex
eBPF
Grafana
Incident Management
Jaeger
Kubernetes
Loki
Mimir
OpenTelemetry
Platform Engineering
Prometheus
SLI/SLO/SLA
Splunk
Thanos
HPC
Apply
$47k – $131k per year (Estimated) • Remote/Hybrid • Full-Time • Barcelona
Go
Python
Rust
Databases
ElasticSearch
OpenSearch
AI/ML
Anomaly Detection
DevOps
AIOps
AWS
Cortex
eBPF
Grafana
Incident Management
Jaeger
Kubernetes
Loki
Mimir
OpenTelemetry
Platform Engineering
Prometheus
SLI/SLO/SLA
Splunk
Thanos
HPC
Apply
Game Artist 2 hours ago
Remote/Hybrid • Full-Time • 3+ years exp • Sofia
Game Dev
Unity
Design
Adobe Photoshop
Apply
Remote/Hybrid • Full-Time • Sofia
Go
Python
Rust
Databases
ElasticSearch
OpenSearch
AI/ML
Anomaly Detection
DevOps
AIOps
AWS
Cortex
eBPF
Grafana
Incident Management
Jaeger
Kubernetes
Loki
Mimir
OpenTelemetry
Platform Engineering
Prometheus
SLI/SLO/SLA
Splunk
Thanos
HPC
Apply
In office • Full-Time • 1+ year exp • Bachelor's Degree • Sofia
Python
SQL
Apply
In office • Full-Time • 6+ years exp • Sofia
Python
AI/ML
Kubeflow
MLFlow
Amazon SageMaker
DevOps
Amazon EC2
Amazon EKS
ArgoCD
AWS
AWS CDK
AWS Fargate
AWS Lambda
CI/CD
FinOps
GitHub Actions
Grafana
Kubernetes
OpenTelemetry
Splunk
Terraform
Amazon CloudWatch
Amazon ECS
Amazon S3
GitHub
IAM
Apply
Remote/Hybrid • Sofia
Databases
Snowflake
AI/ML
dbt
Management
ServiceNow
Marketing
Salesforce
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.