368,746open jobs
9,444companies
47,506added this week
Browse all
Salary
$24k – $59k per year (Estimated)
Location
Remote/Hybrid (Gurgaon, India)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
AHEAD helps enterprises build modern, secure, and scalable digital platforms by combining cloud, data, AI, and automation. Their consulting and managed services drive real business impact through smarter IT.

The High-Performance Computing Storage Engineer is primarily responsible for the overall health and maintenance of storage technologies in our managed services customer's environments. Our Storage Engineers are a valued member of the Managed Services Infrastructure Practice responsible for Tier 3 incident management, service request management and change management infrastructure support for all Managed Services customers.

Key Responsibilities

    • Provide enterprise-level operational support to Managed Services customers for incident, problem, and change management activities

    • Administer parallel and distributed filesystems such as Lustre, GPFS, BeeGFS, Ceph, Weka, or Vast

    • Optimize storage performance, throughput, metadata operations, and data locality for AI training and inference

    • Build and maintain automation for storage provisioning, monitoring, alerting, quota management, and lifecycle operations

    • Plan and perform maintenance activities

    • Assess customer environments for performance and design issues and propose resolutions

    • Work across technical teams to troubleshoot complex infrastructure issues

    • Create and maintain detailed documentation

    • Serve as a subject matter expert and escalation point for storage technologies

    • Work with vendors to resolve storage issues

    • Communicate with customers and internal team with transparency

    • Support data movement workflows including ingest, replication, caching, tiering, and archiving

    • Troubleshoot storage, Linux, network, and I/O bottlenecks across storage clusters and fabrics

    • Partner with infrastructure, platform, and research teams to support production AI/HPC workloads

    • Evaluate new storage architectures and technologies for scalability, resilience, and cost efficiency

    • Communicate with customers and internal team with transparency

    • Participate in on-call rotation

Required Qualifications

    • 5+ years of experience with HPC, AI infrastructure, or large-scale storage engineering

    • Bachelor’s degree or equivalent Information Systems or related field. Unique education, specialized experience, skills, knowledge, training, or certification may be substituted for education

    • Strong experience with Linux systems administration

    • Hands-on experience configuring, managing, and tuning distributed or parallel filesystems

    • Experience tuning storage for performance-sensitive workloads

    • Knowledge of HPC schedulers such as Slurm and/or container platforms such as Kubernetes

    • Familiarity with high-speed interconnects such as InfiniBand or RDMA

    • Ability to troubleshoot complex issues across storage, compute, and networking layers

    • Understanding of data protection mechanisms, including data replication, backup strategies, and disaster recovery in HPC environments

    • Experience with machine learning or data science workflows in HPC environments

    • Managed Services or consulting experience

    • Strong background with customer service

    • High level problem-solving and communication skills

    • Strong oral and written communications skills

    • Managed Services or consulting experience

Preferred Qualifications

    • Experience supporting storage solutions for GPU clusters and AI/ML workflows

    • Familiarity with object storage such as S3, MinIO, or Ceph Object Gateway

    • Experience with Terraform, Ansible, Helm, or GitOps workflows

    • Knowledge of observability platforms such as Prometheus and Grafana

    • Experience with multi-petabyte environments, caching architectures, and storage isolation in multi-tenant systems

    • Experience with machine learning or data science workflows in HPC environments

    • Scripting or programming experience with Python and Bash

    • Related Storage certifications are a bonus

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Gurgaon
$25k – $42k per year • Equity 0–0.2% • Remote • Full-Time • 3+ years exp
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
$100k – $210k per year • Equity 0–0.5% • Remote • Full-Time • 3+ years exp • San Francisco
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
Team Lead DevOps 1 day ago
$23k – $62k per year (Estimated) • Remote • 5+ years exp • Moscow
Bash
Python
Erlang
Erlang
EMQX
Databases
Apache Kafka
ClickHouse
PostgreSQL
RabbitMQ
Redis
Redpanda
Trino
DevOps
Ansible
AWS
AWX
FinOps
HAProxy
Hetzner
Kubernetes
SLI/SLO/SLA
Terraform
Yandex Cloud
Amazon S3
Apply
$108k – $173k per year • Equity • Remote • Full-Time • 5+ years exp • Bachelor's Degree • United States
DevOps
Ansible
Git
Kubernetes
Red Hat
Apply
$10k – $34k per year (Estimated) • In office • Tashkent
Java
Kotlin
Java
Hibernate
Maven
Spring Boot
Kotlin
Mockito
Databases
PostgreSQL
DevOps
CI/CD
Docker
Git
Grafana
Kubernetes
Loki
Prometheus
GitLab
Apply
$130k – $145k per year • Remote • Full-Time • 3+ years exp • Bachelor's Degree
PowerShell
Python
SQL
DevOps
AWS
Azure
FinOps
GCP
Kubernetes
Analytics
Power BI
Tableau
Management
Jira
ServiceNow
Apply
$28k – $66k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Gurgaon
DevOps
VMWare
Apply
NetSuite Developer 5 days ago
$11k – $41k per year (Estimated) • Remote • Full-Time • 3+ years exp
JavaScript
SQL
Apex
Apex
MuleSoft
Databases
Snowflake
Analytics
ETL/ELT
Marketing
Salesforce
Apply
$230k – $290k per year • Remote • Full-Time • 8+ years exp • PhD
Python
AI/ML
AI Agents
Copilot
Anthropic
LLM Guardrails
OpenAI
DevOps
Azure
Azure DevOps
Bicep
CI/CD
FinOps
GitHub Actions
Platform Engineering
Terraform
GitHub
Cybersecurity
Microsoft Entra ID
Apply
$23k – $57k per year (Estimated) • Remote • Full-Time • 8+ years exp • Bachelor's Degree • Gurgaon
PowerShell
Python
AI/ML
AWS Bedrock
Claude
Claude Code
Copilot
CrewAI
Cursor
LangChain
LangGraph
LlamaIndex
LLM
Prompt Engineering
RAG
Vertex AI
Windsurf
AI Agents
Amazon SageMaker
Devin
OpenAI
DevOps
AWS
Azure
Bicep
CI/CD
FinOps
GCP
GitOps
Kubernetes
Platform Engineering
Service Mesh
Terraform
Vector
VMWare
GitHub
Apply
Lead Data Steward 5 hours ago
$25k – $58k per year (Estimated) • In office • Full-Time • Gurgaon
SQL
Databases
Databricks
PostgreSQL
Snowflake
AI/ML
AI Agents
Copilot
Analytics
Power BI
Apply
$16k – $36k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Gurgaon
Apply
Application Developer 6 hours ago
$26k – $69k per year (Estimated) • In office • Full-Time • 5+ years exp • Gurgaon
Java
DevOps
Git
Apply
$32k – $83k per year (Estimated) • In office • Full-Time • 5+ years exp • Gurgaon
Python
Python
pySpark
Databases
Microsoft Fabric
AI/ML
Spark
DevOps
Azure
Apply
$24k – $65k per year (Estimated) • In office • Full-Time • 3+ years exp • Gurgaon
Databases
Snowflake
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.