525,450open jobs
18,278companies
72,729added this week
Browse all
Salary
$38k – $83k per year (Estimated)
Location
In office (Pune)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match

DDN

Staff Engineer

DDN is seeking great candidates to join our dynamic team of passionate customer-enabling technologists!This is an incredible opportunity to be part of a company that has been at the forefront of AI and high-performance data storage innovation for over two decades. DDN Storage is a global market leader renowned for powering many of the world's most demanding AI data centers, in industries ranging from life sciences and healthcare to financial services, autonomous cars, Government, academia, research and manufacturing

"DDN's A3I solutions are transforming the landscape of AI infrastructure." - IDC

“The real differentiator is DDN. I never hesitate to recommend DDN. DDN is the de facto name for AI Storage in high performance environments” - ~ Marc Hamilton

VP, Solutions Architecture & Engineering | NVIDIA

DDN is the global leader in AI and multi-cloud data management at scale. Our cutting-edge storage and data management solutions are designed to accelerate AI workloads, enabling organizations to extract maximum value from their data. With a proven track record of performance, reliability, and scalability, DDN Storage empowers businesses to tackle the most challenging AI and data-intensive workloads with confidence.

Our success is driven by our unwavering commitment to innovation, customer-centricity, and a team of passionate professionals who bring their expertise and dedication to every project. This is a chance to make a significant impact at a company that is shaping the future of AI and data management.

Our commitment to innovation, customer success, and market leadership makes this an exciting and rewarding role for a driven professional looking to make a lasting impact in the world of AI and data storage.

Job Description

As a Staff Engineer - L4, you’ll be an escalation point for the most complex and critical issues affecting enterprise and hyperscale environments. This hands-on role is ideal for a deep technical expert who thrives under pressure and has a passion for solving distributed system challenges at scale. This role is part of the Infinia Core engineering team.

You’ll collaborate with Engineering, Product Management, and Field teams to drive root cause resolutions, define architectural best practices, and continuously improve product resiliency. Leveraging AI tools and automation, you’ll reduce time-to-resolution, streamline diagnostics, and elevate the support experience for strategic customers.

Key Responsibilities

  • Technical Expertise & Escalation LeadershipOwn critical customer case escalations end-to-end, including deep root cause analysis and mitigation strategies.

  • Act as one of the technical escalation points for Infinia incidents - especially in productionimpacting scenarios.

  • Lead war rooms, live incident bridges, and cross-functional response efforts with other engineering, QA, and Field teams.

  • Utilize AI-powered debugging, log analysis, and system pattern recognition tools to accelerate resolution.

Product Knowledge & Value Creation

  • Become a subject-matter expert on Infinia internals: metadata handling, storage fabric interfaces, performance tuning, AI integration, etc.

  • Reproduce complex customer issues and propose product improvements or workarounds.

  • Author and maintain detailed runbooks, performance tuning guides, and RCA documentation.

  • Feed real-world support insights back into the development cycle to improve reliability and diagnostics.

Customer Engagement & Business Enablement

  • Partner with Field CTOs, Solutions Architects, and Sales Engineers to ensure customer success.

  • Translate technical issues into executive-ready summaries and business impact statements.

  • Participate in post-mortems and executive briefings for strategic accounts.

  • Drive adoption of observability, automation, and self-healing support mechanisms using AI/ML tools.

  • Delivering training to customer support and field engineering

Required Qualifications

  • 8+ years in enterprise storage, distributed systems, or cloud infrastructure support/engineering.

  • Deep understanding of file systems (S3, POSIX, NFS), storage performance, and Linux kernel internals.

  • Scripting and Coding using Python, Go, C++

  • Proven debugging skills at system/protocol/app levels (e.g., strace, tcpdump, perf).

  • Hands-on experience with troubleshooting on Linux.

  • Exposure to RDMA, NVMe-oF, or high-performance networking stacks.

  • Exceptional communication and executive reporting skills.

  • Experience using AI tools (e.g., log pattern analysis, LLM-based summarisation, automated RCA tooling) to accelerate diagnostics and reduce MTTR.

Preferred Qualifications

  • Experience with DDN, VAST, Weka, or similar scale-out file systems.

  • Strong scripting/coding ability in Python, Bash, or Go.

  • Familiarity with observability platforms: Prometheus, Grafana, ELK, OpenTelemetry.

  • Knowledge of replication, consistency models, and data integrity mechanisms.

  • Exposure to Sovereign AI, LLM model training environments, or autonomous system data architectures.

  • This position requires participation in an on-call rotation to provide after-hours support as needed.

Success Metrics - First 30 Days

Technical Ramp-Up

  • Complete Infinia training, labs, and architecture deep dives. o Stand up a fully functioning Infinia test system. o Shadow at least 5 complex escalations and participate in 2 customer calls.

Operational Integration

  • Lead one live incident response and deliver a full RCA within 48 hours. o Propose 3+ enhancements to internal tools, AI/automation usage, or documentation. o Establish key partnerships with Engineering and Field teams.

Strategic Insight

  • Deliver a written 30-day reflection with gaps and high-impact recommendations. o Begin identifying patterns where AI or automation can reduce MTTR or improve proactive detection.

Success Metrics - Beyond 30 Days

  • MTTR on high-severity cases is consistently below internal SLAs.

  • Volume and quality of resolved L4 escalations.

  • Strategic tooling or automation contributions adopted across the support org.

  • Executive-ready RCAs that inform product improvement.

  • High-impact engagements with strategic accounts (prevention, performance tuning, etc.).

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
525,450 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Pune
$47k – $103k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Auckland
Python
C++
Apply
$154k – $257k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Vancouver
Python
C++
Apply
Remote/Hybrid • Internship • Bachelor's Degree
Python
Java
Verilog
C++
SystemVerilog
Perl
Chips/EDA
UVM
Apply
$61k – $157k per year (Estimated) • Remote/Hybrid • Full-Time • 7+ years exp • Wellington • Auckland
Python
Java
TypeScript
Java
Spring Framework
Spring Boot
Databases
PostgreSQL
Apache Kafka
DevOps
Rest API
Splunk
OpenShift
Dynatrace
Kong
Azure
CI/CD
Jenkins
AWS
Kubernetes
Spinnaker
Platform Engineering
Amazon EKS
Azure AKS
Management
Agile
Apply
$87k – $184k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Toronto
Python
Apply
$185k – $275k per year • In office • Full-Time • Santa Clara
Python
C++
AI/ML
LLM
DevOps
OpenTelemetry
Prometheus
Grafana
Amazon S3
Cybersecurity
Tcpdump
Apply
$123k – $245k per year (Estimated) • Remote • Full-Time • 8+ years exp
AI/ML
Cerebras
TPU
DevOps
Kubernetes
OpenStack
HPC
Apply
HR Specialist 2 days ago
$14k – $34k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Shanghai
Apply
Demo Engineer, EMEA 3 days ago
$39k – $87k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Madrid
AI/ML
AI Agents
Apply
$65k – $85k per year • Remote/Hybrid • Full-Time • 5+ years exp • Madrid
Apex
Apex
Lightning Web Components
MuleSoft
Salesforce Flow
Visualforce
AI/ML
Copilot
Claude
Agentforce
DevOps
PagerDuty
CI/CD
Git
Management
Slack
Jira
ServiceNow
Apply
Data Engineer 7 hours ago
$26k – $53k per year (Estimated) • In office • Full-Time • 5+ years exp • Navi Mumbai • Pune
Databases
Snowflake
Analytics
ETL/ELT
Apply
Test Automation Lead 7 hours ago
$21k – $56k per year (Estimated) • In office • Full-Time • 12+ years exp • Pune
DevOps
CI/CD
Apply
Security Architect 7 hours ago
$30k – $71k per year (Estimated) • In office • Full-Time • 3+ years exp • Pune
DevOps
AWS
Apply
DevOps Engineer 7 hours ago
$13k – $37k per year (Estimated) • In office • Full-Time • 2+ years exp • Pune
DevOps
GCP
Docker Swarm
Azure
CI/CD
AWS
Kubernetes
Apply
$25k – $58k per year (Estimated) • In office • Full-Time • 5+ years exp • Hyderabad • Chennai • Pune • Bhubaneswar • Ahmedabad
Management
Agile
Apply
See all jobs
This is one of many
525,450 more open roles from verified company boards, updated every day.