552,642open jobs
20,924companies
75,446added this week
Browse all
Salary
$185k – $275k per year
Location
In office (Santa Clara)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match

DDN

DDN is seeking a Staff Engineer to join our Infinia Core team. This is a hands-on technical role combining deep distributed-systems engineering with direct engagement with customers running Infinia in production.

You'll own complex technical escalations end-to-end, from root-cause analysis and incident response through to mitigation, customer communication and product improvements. You'll also help shape engineering best practice, mentor other engineers and drive the use of AI and automation to improve reliability and diagnostics.

If you love the technical depth but want to stay behind the curtain, this probably isn't the right fit but if you want to combine serious engineering with real customer impact, read on.

About Infinia

Infinia is DDN's next-generation, software-defined storage platform, built from the ground up for AI and accelerated computing. It combines separate control and data planes, all-flash performance, sub-millisecond latency and multi-tenancy for demanding enterprise and hyperscale AI and GPU workloads.

What You'll Do

  • Communicate technical issues clearly to customers, engineers and senior stakeholders, including executive audiences.

  • Own complex customer escalations from diagnosis through to resolution, mitigation and RCA.

  • Lead live incident response, war rooms and cross-functional investigations with Engineering, QA and Field teams.

  • Debug complex distributed-systems, storage and performance issues across the system, protocol and application layers.

  • Reproduce customer issues and feed findings into product and reliability improvements.

  • Develop runbooks, troubleshooting guidance and performance-tuning practices.

  • Act as a technical authority on Infinia internals, mentoring engineers and influencing architectural best practice.

  • Partner with Field CTOs, Solutions Architects and Sales Engineers on strategic customer issues.

  • Use AI, automation and observability to improve diagnostics, reliability and MTTR.

  • Communicate technical issues clearly to customers, engineers and senior stakeholders, including executive audiences.

  • This position requires participation in an on-call rotation to provide after-hours support as needed.

What You'll Bring

Must-Haves

  • Significant experience in enterprise storage, distributed systems or cloud infrastructure, with technical leadership at Senior or Staff level.

  • Deep understanding of file systems and storage technologies, including S3, POSIX, NFS and storage performance.

  • Strong Linux systems knowledge, including kernel-level troubleshooting and debugging.

  • Strong coding ability in Python or C++.

  • Proven ability to diagnose complex issues using tools such as strace, tcpdump and perf.

  • Genuine interest in working directly with customers and taking ownership of complex problems through to resolution.

Nice-to-Haves

  • Experience with DDN, VAST, Weka or similar scale-out storage/file systems.

  • Familiarity with observability platforms such as Prometheus, Grafana, ELK or OpenTelemetry.

  • Knowledge of replication, consistency models and data integrity mechanisms.

  • Experience supporting AI/ML, LLM training or other high-performance computing environments.

  • Experience using AI tools for log analysis, troubleshooting, automated RCA or reducing MTTR.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
552,642 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Santa Clara
In office • Bachelor's Degree
Python
Java
TypeScript
C++
DevOps
CI/CD
Apply
$160k – $240k per year • In office • 4+ years exp • Bachelor's Degree
Python
Go
Java
TypeScript
Ruby
C++
DevOps
Ansible
Podman
Chef
Packer
CI/CD
Jenkins
Docker
Kubernetes
Configuration Management
Cybersecurity
SBOM
Apply
$160k – $240k per year • In office • 4+ years exp
Python
Java
DevOps
Terraform
Ansible
eBPF
IAM
Cybersecurity
CyberArk
Least Privilege
Teleport
Apply
$160k – $210k per year • In office • 5+ years exp
Python
SQL
Analytics
Superset
Microsoft Excel
Apply
$160k – $240k per year • In office • 4+ years exp • Bachelor's Degree
Python
AI/ML
DeepSpeed
ONNX
PyTorch
DevOps
GCP
Azure
AWS
Kubernetes
Argo Workflows
Apply
$123k – $245k per year (Estimated) • Remote • Full-Time • 8+ years exp
AI/ML
Cerebras
TPU
DevOps
Kubernetes
OpenStack
HPC
Apply
HR Specialist 2 days ago
$14k – $34k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Shanghai
Apply
Demo Engineer, EMEA 3 days ago
$39k – $87k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Madrid
AI/ML
AI Agents
Apply
$65k – $85k per year • Remote/Hybrid • Full-Time • 5+ years exp • Madrid
Apex
Apex
Lightning Web Components
MuleSoft
Salesforce Flow
Visualforce
AI/ML
Copilot
Claude
Agentforce
DevOps
PagerDuty
CI/CD
Git
Management
Slack
Jira
ServiceNow
Apply
$29k – $71k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree
DevOps
HPC
Marketing
Salesforce
Apply
$46k – $71k per year (Estimated) • In office • Internship • PhD • Santa Clara • Irvine • Austin • Morrisville • Chandler
Python
AI/ML
Copilot
LangGraph
AutoGen
LangChain
Claude
ChatGPT
LlamaIndex
Model Context Protocol
Fine-tuning
Reinforcement Learning
Prompt Engineering
Multimodal AI
Diffusion Models
AI Agents
TensorFlow
PyTorch
CrewAI
LLM
RAG
Hugging Face
A2A
Agentic Workflows
Multi-Agent Systems
Tool Use
DevOps
Git
Management
n8n
Apply
$61k – $100k per year • In office • Bachelor's Degree • Santa Clara
Apply
$167k – $291k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • Santa Clara
Python
SQL
Bash
Databases
Azure Cosmos DB
AI/ML
AI Agents
Anomaly Detection
LLM Guardrails
DevOps
Terraform
GCP
CloudFormation
Azure
CI/CD
GitOps
AWS
Kubernetes
Platform Engineering
Service Mesh
Self-Healing
Amazon EKS
Google GKE
Azure AKS
AWS Lambda
Amazon EC2
AIOps
Incident Management
SLI/SLO/SLA
Management
ServiceNow
Apply
$136k – $213k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara
Apply
$168k – $270k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Santa Clara
Python
SQL
AI/ML
Spark
DevOps
Terraform
Azure
AWS
Docker
Kubernetes
Analytics
Tableau
Power BI
ETL/ELT
SAP BusinessObjects
Management
Outlook
Apply
See all jobs
This is one of many
552,642 more open roles from verified company boards, updated every day.