Confirmed on the employer's own hiring board on Sep 25, 2026. First seen by Alion on Sep 25, 2026.
### **Key Responsibilities**
* Manage, monitor, and support cloud infrastructure across platforms (AWS / GCP / OCI / Azure)
* Perform VM provisioning, scaling, patching, backup, and lifecycle management
* Configure and maintain monitoring tools such as New Relic, Zabbix, ManageEngine, Azure Monitor, Grafana, or similar
* Set up alerts, dashboards and proactive health checks to ensure high availability
* Perform incident analysis, root cause analysis (RCA), and implement preventive actions
* Handle L1 incidents, troubleshooting, and timely escalation as per SLA
* Work closely with infrastructure, network, and application teams for issue resolution
* Maintain and update documentation, SOPs, and operational runbooks
### **Required Skills**
* Good Exposure in Cloud Infrastructure / Monitoring / Operations
* Exposure to multiple cloud platforms (AWS, Azure, OCI, GCP)
* Strong Knowledge with at least any one cloud platform (Azure /AWS / GCP / OCI preferred)
* Strong knowledge of monitoring and alerting tools
* Basic understanding of networking concepts (VPC/VNet, Subnets, Routing, Load Balancer)
* Familiarity with Linux and Windows system administration
* Scripting knowledge (PowerShell / Bash / Python)
* Strong troubleshooting and analytical skills
### **Good to Have**
* Knowledge of Infrastructure as Code (Terraform, ARM, CloudFormation)
* Familiarity with AIOps / AI-based monitoring and observability platforms
* Basic understanding of AI/ML concepts in IT operations
* Relevant cloud certifications (AWS / Azure / GCP fundamentals or associate level)

