First seen by Alion on Oct 9, 2026.
About the Role :
We are looking for an experienced Senior GCP Cloud Operations Engineer to manage, provision, deploy, monitor, and troubleshoot cloud infrastructure. The ideal candidate will have strong hands-on expertise in Google Cloud Platform (GCP), Terraform, Kubernetes (GKE), infrastructure automation, CI/CD, and production operations.
Key Responsibilities :
- Design, provision, deploy, and maintain cloud infrastructure on Google Cloud Platform.
- Implement and manage Infrastructure as Code (IaC) using Terraform, including reusable modules, remote state, state locking, and environment separation.
- Deploy, operate, and troubleshoot containerised applications using Kubernetes and Google Kubernetes Engine (GKE).
- Configure and manage GCP services, including Compute Engine, Cloud Run, Cloud Storage, Cloud SQL, IAM, VPC, Cloud DNS, and Load Balancing.
- Develop and maintain CI/CD pipelines using Cloud Build, Jenkins, GitHub Actions, GitLab CI, or equivalent tools.
- Manage cloud networking components, including subnets, firewall rules, Cloud NAT, Cloud Router, VPN, private connectivity, and Shared VPC.
- Monitor infrastructure health, investigate incidents, troubleshoot production issues, and improve system reliability and performance.
- Implement cloud security best practices, including least-privilege IAM, service accounts, secrets management, and audit logging.
- Develop automation scripts using Bash, Python, or Go.
- Identify opportunities for cloud cost optimisation and operational efficiency.
Required Skills :
- 6+ years of experience in infrastructure engineering, DevOps, platform engineering, or SRE.
- ~4 years of hands-on experience working with GCP services.
- Strong proficiency in Terraform (HCL, modules, state management).
- Hands-on experience with GKE, Kubernetes deployments, and troubleshooting.
- Strong understanding of GCP networking (VPC, load balancing, VPN).
- Experience with CI/CD pipelines, Linux administration, and scripting (Bash/Python/Go).
- Ability to explain infrastructure trade-offs and failure scenarios.
Good to Have :
- Experience with GKE Autopilot, Anthos, Cloud Deploy, Artifact Registry, FinOps, and Policy as Code (OPA/Sentinel).
Skills
Cloud, Terraform, Kubernetes, CI/CD Pipeline, Linux, Python, Bash Scripting, Google Cloud Platform, Cloud Operations

