Senior DevOps Engineer
Next Generation Monitoring (NGM) Platform - Vertiv IT Systems | Pune, Maharashtra
Experience: 8-12 Years
- Band: Senior Engineer (IC)
- Department: Software Engineering - IT Systems
- Reports To: Engineering Manager
POSITION SUMMARY
We are seeking an experienced Senior DevOps Engineer to own the engineering delivery infrastructure - CI/CD pipelines, container orchestration, multi-deployment packaging, release management, security automation, and platform observability. The engineer will support on-premises, cloud-native (AWS/GCP/Azure), hybrid edge-cloud, and air-gapped deployment targets across a multi-year, multi-release product roadmap.
KEY RESPONSIBILITIES
CI/CD Pipeline Engineering
- Design, build, and optimize GitLab CI/CD pipelines for all microservices with independent deployment cycles
- Implement multi-stage pipelines: build → test → security scan → package → deploy across all environments
- Integrate automated unit, integration, and system-level tests; manage all pipeline configuration as code
- Configure branch-based promotion strategies (dev → staging → production)
Container Orchestration & Infrastructure
- Manage containerized deployments via Docker and Kubernetes on Linux and Windows Server
- Provision and manage Kubernetes clusters (dev, staging, production) with horizontal pod autoscaling
- Configure service mesh (Istio/Linkerd) for inter-service mTLS; manage container image registry lifecycle
- Implement Infrastructure as Code using Terraform and/or Ansible for cloud and on-premises resources
Multi-Deployment Operations
- Build on-premises Docker/Kubernetes deployment automation for Linux and Windows Server targets
- Develop cloud-native automation for AWS, GCP, and Azure (EKS/GKE/AKS) including multi-region config
- Engineer hybrid edge-cloud pipelines; package air-gapped artefact bundles for offline/classified environments
- Implement zero-downtime rolling updates, blue-green deployments, and automated rollback strategies
Release Management
- Orchestrate software releases across the multi-release product roadmap; coordinate cross-functional readiness
- Manage versioning, release tagging, changelog generation, and artefact storage across all microservices
- Produce deployment runbooks, release notes, and rollback procedures for each release milestone
Security & Compliance Automation
- Integrate SAST, DAST, container image scanning (Trivy/Snyk/Clair), and dependency vulnerability checks
- Automate SBOM generation, open-source license compliance, and software supply chain tracking
- Configure secrets management (HashiCorp Vault), TLS certificate automation, and PKI for microservices
- Implement pipeline controls for SOC 2, EU Cyber Resilience Act, and FedRAMP readiness
Monitoring, Observability & Reliability
- Deploy centralized logging (ELK/Loki), metrics (Prometheus/Grafana), and distributed tracing (Jaeger)
- Configure Kubernetes cluster monitoring, pod health dashboards, and alerting/on-call integration
- Drive proactive reliability engineering to meet high-availability platform SLA targets
Automation & Cross-Functional Collaboration
- Create Bash/Python automation for infrastructure operations, environment provisioning, and self-healing
- Interface with software architects, developers, QA, and product management on delivery and planning
- Guide development teams on DevOps best practices, containerization patterns, and pipeline design
REQUIRED TECHNICAL SKILLS
Infrastructure & Platform: Linux (Ubuntu, RHEL, CentOS) production admin (5+ yrs)
- Docker & Kubernetes (Helm, RBAC, autoscaling)
- Service mesh: Istio/Linkerd
- Windows Server 2019+ for on-premises targets
CI/CD & Version Control: GitLab CI/CD administration & pipeline YAML (5+ yrs)
- Multi-service microservices pipelines
- Artefact management: GitLab, ECR, GCR, ACR
- GitOps with Helm chart management
Cloud & Multi-Deployment: AWS, GCP, or Azure (multi-cloud preferred)
- EKS/GKE/AKS auto-scaling & multi-region
- Hybrid edge-cloud
- Air-gapped offline artefact packaging
- Terraform & Ansible (5+ yrs)
Development & Automation: Advanced Bash & Python scripting
- YAML/JSON for Kubernetes manifests & Helm
- REST API automation
- Go (Golang) familiarity for microservices platform context
Security & Compliance: Container scanning: Trivy, Snyk, or Clair
- SAST/DAST pipeline integration
- SBOM: Syft/CycloneDX
- HashiCorp Vault
- TLS/PKI automation
- SOC 2, EU CRA, FedRAMP pipeline controls
Monitoring & Observability: ELK/EFK, Loki, or Splunk
- Prometheus & Grafana (production-grade)
- Distributed tracing: Jaeger/Zipkin
- Kubernetes cluster monitoring
- SLA reliability metrics
GOOD-TO-HAVE SKILLS
- Cloud certifications: AWS DevOps Engineer, Google Professional DevOps, or Azure DevOps Expert
- CKA/CKAD, Kubernetes operator development, and service mesh deep expertise
- GitOps tooling: ArgoCD or Flux for continuous delivery
- SRE practices: error budget management, chaos engineering (LitmusChaos, Chaos Monkey)
- Time-series DB operations: InfluxDB or TimescaleDB backup, scaling, schema migration
- Message broker management: Kafka or NATS cluster operations and monitoring
- Edge node provisioning, remote management, and telemetry collection from distributed agents
- Internal Developer Platform (IDP) tooling and developer self-service workflows
REQUIRED QUALIFICATIONS
- Bachelor's Degree in Computer Science, Information Technology, Electrical Engineering, or related field
- Minimum 8 years of experience in DevOps, Site Reliability Engineering, or Infrastructure Engineering
- 5+ years of Linux system administration in production environments (Ubuntu, RHEL, CentOS)
- 5+ years of Docker and Kubernetes container orchestration in production
- 5+ years of CI/CD pipeline management (GitLab CI preferred)
- 5+ years of cloud platform deployments: AWS, GCP, or Azure
- Proven track record delivering end-to-end CI/CD pipelines for microservices architectures
- Strong background in Infrastructure as Code (Terraform, Ansible, or equivalent)
- Experience integrating security and compliance tooling into DevOps workflows
SOFT SKILLS & EDUCATION
- Strong problem-solving and root-cause analysis in complex distributed systems
- Clear written and verbal communication for cross-functional engineering collaboration
- Proactive ownership mindset with a strong bias toward automation over manual operations
- Ability to manage competing priorities in a fast-paced agile product development environment
- Team collaboration and informal DevOps culture leadership; comfort with ambiguity
Education: B.E. / B.Tech / M.E. / M.Tech in Computer Science, IT, or related field. Cloud/DevOps certifications (CKA/CKAD, AWS/GCP/Azure) are a plus.

