We are seeking a seasoned Site Reliability Engineering to take end-to-end ownership of our retail eCommerce platform's operational excellence. You will drive SRE culture across engineering teams, champion observability, and work with a cross-functional onsite/offshore team to deliver exceptional uptime, performance, and release velocity. This role blends deep technical expertise - you'll be equally comfortable reviewing Kubernetes manifests, tuning Akamai rules, and presenting SLO dashboards to senior leadership
Must have skills and Technologies
Core Technical Skills
SRE Practices & Leadership
- Should have Development experience in an E-Commerce system built on Java/Microservices
- SLI / SLO / Error Budget management
- Akamai CDN (edge config, WAF, TLS, traffic Mgmt)
- Incident Management & Blameless Post-Mortems
- AWS Core Services (EC2, EKS, RDS, S3, CloudWatch)
- Performance tuning & capacity planning
- Kubernetes / Amazon EKS (production operations)
- AWS IAM, VPC, Security Groups, KMS
- Jenkins CI/CD (pipeline design & optimization)
- Container security & image scanning
- Splunk APM & Splunk Cloud (dashboards, alerts)
- Monitoring stack design (metrics, logs, traces)
- Redis & Memcached (distributed caching strategy)
- Distributed systems troubleshooting at scale
- CloudFormation / Auto Scaling Groups
- Cross-team technical leadership & mentoring
- Java and Microservices architecture & service mesh
Preferred / Nice to have skills
Python, Bash, JavaScript, Java, Terraform, AWS CDK
Open Telemetry, Chaos Engineering, GitOps / ArgoCD, PCI-DSS, SOC 2
NGINX / Apache, Service Mesh (Istio), Dynatrace / Datadog, JIRA / Confluence
- Experience supporting or developing eCommerce web applications (JavaScript & Java)
- Familiarity with accessibility and compliance standards in retail platforms
- Prior experience in high-GMV retail or marketplace environments
- Bachelor's or Master's degree in Computer Science, Engineering, or equivalent experience
- 8+ years of SRE, DevOps, or Infrastructure Engineering experience in production environments
- Proven track record managing large-scale AWS cloud-native microservices architecture
- AWS Certified DevOps Engineer - Professional or Solutions Architect (preferred)
- CKA / CKAD (Certified Kubernetes Administrator / Developer) is a strong plus

