950,760open jobs
57,518companies
154,547added this week
Browse all
Salary
$148k – $220k per year
Location
In office (Pittsburgh)
Seniority
Senior · 5+ years exp

Confirmed on the employer's own hiring board on Sep 29, 2026. First seen by Alion on Sep 3, 2026. NetApp scores A on the Alion truth index.

Overview
Company
Impact
Profile match
NetApp provides enterprise data storage and cloud data services. Its ONTAP software manages file, block and object storage across on-premise arrays and public clouds. The company also sells observability and cloud cost optimisation tooling.

Job Summary

As a Cloud Infrastructure / Site Reliability Engineer, you will operate at the intersection of development and operations. You will engage and enhance all aspects of the cloud services lifecycle from design through deployment, operation, and refinement. You will be responsible for maintaining these services by measuring and monitoring their availability, latency, and overall system health and building automation for efficient cloud operations management.

You will play a crucial role in sustainably scaling systems through automation and driving changes that improve reliability and velocity. As part of your responsibilities, you will administer cloud-based environments that support our SaaS/IaaS offerings implemented on a microservices, container-based architecture (Kubernetes). In addition, you will oversee a portfolio of customer-centric cloud services (SaaS/IaaS), ensuring their overall availability, performance, and security. You will work closely with NetApp and cloud service provider teams (to include Azure) from Research Triangle Park (RTP), D.C., Pittsburg and more.

Due to the critical nature of the services we support, this position involves participation in a rotation-based on-call schedule as part of our global team. This role offers the opportunity to work in a dynamic, global environment, ensuring the smooth operation of vital cloud services. To be successful in this role, you should be a motivated self-starter and self-learner, possess strong problem-solving skills, and be someone who embraces challenges.

Responsibilities

  • Automation and Efficiency: Identify tasks and areas where automation can be applied to achieve time efficiencies and risk reduction. Develop software for deployment automation, packaging, and monitoring visibility.
  • Team Collaboration and Influence: Work in tandem with other Cloud Infrastructure Engineers and developers to ensure maximum performance, reliability, and automation of our deployments and infrastructure. Consult and influence developers on new feature development and software architecture to ensure scalability.
  • Debugging, Troubleshooting, and Advanced Support: Undertake debugging and troubleshooting of service bottlenecks throughout the entire software stack. Additionally, provide advanced tier 2 and 3 support for NetApp's Cloud Data Service solutions.
  • Analysis, and Infrastructure Maintenance: Continuously monitor, analyze, and measure system health, availability, and latency using tools like Prometheus, Stackdriver, ElasticSearch, Grafana, and SolarWinds. Develop strategies to enhance system and application performance, availability, and reliability. In addition, maintain and monitor the deployment and orchestration of servers, docker containers, databases, and general backend infrastructure.
  • Incident Response and Troubleshooting: Address and perform Root Cause Analysis (RCA) of complex live production incidents and cross-platform issues involving OS, Networking, and Database in cloud-based SaaS/IaaS environments. Implement SRE best practices for effective resolution.
  • Document system knowledge as you acquire it, create runbooks, and ensure critical system information is readily accessible. Security Management: Stay updated with security protocols and proactively identify, diagnose, and resolve complex security issues. Issue Tracking and Resolution: Use Atlassian’s tool chain along with first party cloud service management tools to track and resolve issues based on their priority.
  • Directly influence the decisions and outcomes related to solution implementation: measure and monitor availability, latency, and overall system health.

Job Requirements

  • This role requires US Citizenship, due to potential need for Security Clearance.
  • 5+ years experience in scripting and infrastructure automation using tools such as PowerShell, Python, Go or Ruby.
  • Deep working knowledge of Containers, Kubernetes, Serverless computing implementation, and distributed systems design patterns. Knowledge of DevOps/SRE development methodologies.
  • Proficiency in Linux/Unix and CoreOS. Experience with cloud platforms such as AWS, Azure, or Google Cloud.
  • Ability to lead a scrum team, influence stakeholders to effectively maintain a product backlog, manage sprints.
  • This position will have ON-CALL rotations as well as an ask to work odd hourss.

Education

  • A Bachelor of Science Degree in Computer Science, a master’s degree; or equivalent experience is required

Compensation:

The target salary range for this position is 147,900 - 220,000 USD. The salary offered will be determined by the candidate's location, qualifications, experience, and education and may be outside of this range. Final compensation packages are competitive and in line with industry standards, reflecting a variety of factors, and include a comprehensive benefits package. This may cover Health Insurance, Life Insurance, Retirement or Pension Plans, Paid Time Off, various Leave options, Performance-Based Incentives, employee stock purchase plan, and/or restricted stocks (RSU’s), with all offerings subject to regional variations and governed by local laws, regulations, and company policies. Benefits may vary by country and region, and further details will be provided as part of the recruitment process.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
950,760 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Pittsburgh
$120k – $235k per year • In office • TS/SCI • Bachelor's Degree • Lorton
Apply
$105k – $135k per year • Equity • Remote (United States) • 3+ years exp • Bachelor's Degree
Python
JavaScript
TypeScript
SQL
C#
Node JS
Python
Django
Databases
MySQL
PostgreSQL
Db2
Oracle
Frontend
React.js
DevOps
Terraform
GCP
Azure DevOps
Azure
Git
AWS
Docker
Kubernetes
Bitbucket
GitHub
Apply
Software Engineer II 10 days ago
$89k – $111k per year • In office • Full-Time • 6+ years exp • Bachelor's Degree • Victor
Verilog
C#
VHDL
MATLAB
DevOps
Linux
Apply
Software Developer 4 8 hours ago
$115k – $235k per year • Equity • In office • High School Diploma • Austin
JavaScript
Java
Kotlin
TypeScript
SQL
Java
Maven
Spring Boot
Kotlin
Mockito
Databases
Oracle
Apache Kafka
Frontend
Angular
React.js
DevOps
Rest API
Splunk
New Relic
CI/CD
Jenkins
Git
Docker
Kubernetes
Bitbucket
API Gateway
Cybersecurity
LDAP
QA
Swagger
Apply
$175k – $250k per year • In office • Full-Time • 4+ years exp • San Francisco
Python
JavaScript
TypeScript
Ruby
Node JS
Ruby
Ruby on Rails
Databases
PostgreSQL
AI/ML
Copilot
Cursor
Claude Code
LLM
Frontend
React.js
DevOps
AWS
Apply
$84k – $209k per year • In office • Full-Time • Bengaluru
Python
Java
SQL
Scala
Databases
Apache Kafka
AI/ML
Hadoop
Spark
Flink
Machine Learning
DevOps
GCP
Azure
AWS
Analytics
ETL/ELT
Apply
≈ $20k – $56k per year (Estimated) • In office • Full-Time • Master's Degree • Bengaluru
Python
Python
pySpark
AI/ML
Spark
AI Agents
NLP
LLM
Anomaly Detection
Time Series Forecasting
Agentic Workflows
DevOps
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Analytics
ETL/ELT
Apply
≈ $47k – $113k per year (Estimated) • Equity • In office • Full-Time • Bremen
Python
MATLAB
AI/ML
Machine Learning
Management
Agile
Apply
In office • Full-Time • PhD • Toulouse
Python
AI/ML
Computer Vision
Apply
≈ $12k – $32k per year (Estimated) • Remote (EAEU) • 2+ years exp • Moscow
Python
SQL
DevOps
Rest API
GitLab CI
CI/CD
Jenkins
Git
Bitbucket
SOAP
Management
Confluence
Jira
QA
Postman
Apply
$228k – $339k per year • Equity • In office • 15+ years exp • San Jose
Python
C++
Databases
Cassandra
DevOps
Amazon S3
Linux
Apply
$113k – $168k per year • Equity • In office • 2+ years exp • Bachelor's Degree • Wichita
DevOps
GCP
Apply
$180k – $210k per year • Equity • In office • 4+ years exp • Bachelor's Degree • San Jose
Python
Go
TypeScript
C#
C++
Databases
PostgreSQL
RabbitMQ
Apache Kafka
DevOps
Rest API
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Apply
$215k – $245k per year • Equity • In office • 8+ years exp • Bachelor's Degree • San Jose
Python
Go
TypeScript
C#
C++
Databases
PostgreSQL
RabbitMQ
Apache Kafka
DevOps
Rest API
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Apply
≈ $208k – $413k per year (Estimated) • Equity • In office • 8+ years exp • San Jose
Python
C++
DevOps
GCP
Azure
CI/CD
AWS
KVM
QEMU
Amazon S3
Linux
Apply
$93k – $189k per year • Remote (United States) • Full-Time • 10+ years exp • Bachelor's Degree • Grand Rapids • Pittsburgh • Charlotte
Analytics
Microsoft Excel
Apply
≈ $51k – $99k per year (Estimated) • In office • 1+ year exp • High School Diploma • Pittsburgh
Apply
Remote (United States) • Pittsburgh
Apply
$26k per year • In office • Pittsburgh
Apply
≈ $114k – $210k per year (Estimated) • In office • Pittsburgh
Apply
See all jobs
This is one of many
950,760 more open roles from verified company boards, updated every day.