This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a DevOps Engineer - Contract based in India.
As a DevOps Engineer, you will help design, build, and continuously evolve a modern cloud-based platform focused on observability, reliability, automation, and data quality. You will work extensively with Microsoft Azure, Kubernetes, CI/CD, Infrastructure as Code, and cloud-native technologies. The role combines DevOps engineering with Site Reliability Engineering (SRE), data infrastructure, monitoring, and FinOps practices. You will build intelligent capabilities that proactively detect issues, identify anomalies, and support automated resolution across critical operational systems. You will also help establish robust data-quality monitoring across complex supply chain and enterprise data pipelines. Working in a collaborative and inclusive environment, you will have a direct impact on platform resilience, operational efficiency, cloud costs, and data-driven decision-making.
Accountabilities:
- Architect, build, and continuously evolve core platform services that provide observability, operational monitoring, data quality, and reliability capabilities.
- Design and maintain highly available, stable, scalable, and performant cloud infrastructure, applying Site Reliability Engineering (SRE) principles throughout platform design and operations.
- Develop proactive monitoring and intelligent detection capabilities to identify technical issues, anomalies, performance bottlenecks, and operational risks before they significantly impact systems.
- Implement automated and semi-automated remediation workflows to accelerate incident response and improve operational resilience.
- Build and integrate capabilities for continuously measuring, monitoring, and reporting data quality across dimensions including timeliness, consistency, completeness, accuracy, validity, and uniqueness.
- Monitor data quality across diverse supply chain and enterprise data pipelines, including warehousing, transformation, SAP, manufacturing, and other critical data sources.
- Develop and manage centralized dashboards using technologies such as Grafana, Power BI, or Databricks to provide visibility into system health, operational KPIs, financial metrics, and data quality.
- Drive infrastructure and operational automation through Infrastructure as Code (IaC), CI/CD pipelines, Python, Shell scripting, and other automation approaches.
- Integrate, manage, and optimize technologies including Fivetran, Azure Kubernetes Service (AKS), Kafka, Azure SQL Server, Databricks, and Flink within an Azure cloud environment.
- Implement and maintain effective CI/CD workflows supporting reliable, repeatable, and efficient software and infrastructure delivery.
- Monitor cloud resource utilization and spending, embedding FinOps practices and reporting to identify cost-optimization opportunities and improve infrastructure investment efficiency.
- Analyze incidents, performance trends, operational metrics, and data-quality issues to identify opportunities for continuous improvement.
- Support data-driven business decision-making by ensuring monitoring and data-quality insights are accurate, timely, accessible, and reliable.
- Help ensure platform processes and data management practices support regulatory requirements for data integrity, governance, traceability, and auditability.
- Collaborate with engineering, data, operations, and business stakeholders to continuously improve platform capabilities and operational processes.
- 5+ years of professional experience in DevOps, Cloud Engineering, Platform Engineering, Site Reliability Engineering, or a closely related role.
- Strong hands-on experience with Microsoft Azure and cloud infrastructure, including designing, deploying, and operating cloud-based environments.
- Practical experience with Azure Kubernetes Service (AKS), Kubernetes, and containerized application environments.
- Strong experience implementing and maintaining CI/CD pipelines and Infrastructure as Code (IaC) solutions.
- Strong scripting and automation capabilities using Python and Shell scripting.
- Hands-on experience with observability and monitoring technologies such as Grafana and Prometheus.
- Strong understanding and practical application of Site Reliability Engineering (SRE) principles, including reliability, availability, performance, monitoring, and incident management.
- Experience working with Kafka, Azure SQL Server, Databricks, Flink, and Fivetran.
- Experience with infrastructure automation, system monitoring, performance optimization, troubleshooting, and incident response.
- Strong understanding of cloud operations, distributed systems, and production environments.
- Analytical and problem-solving mindset, with the ability to proactively identify technical risks and develop effective solutions.
- Strong communication and collaboration skills, with the ability to work effectively across technical and cross-functional teams.
- Ability to work independently, take ownership of complex technical initiatives, and contribute to a culture of continuous improvement.
- Commitment to working in an inclusive, respectful, and collaborative environment where diverse perspectives are valued.
- 12-month contract opportunity with the potential for meaningful work on modern cloud and data platforms.
- 100% remote working opportunity from anywhere in India.
- Opportunity to work with Microsoft Azure, Kubernetes, Databricks, Kafka, Flink, Fivetran, and other modern cloud and data technologies.
- Exposure to advanced DevOps, SRE, observability, automation, data quality, and FinOps practices.
- Opportunity to influence platform architecture, reliability, operational efficiency, and cloud-cost optimization.
- Work on complex enterprise and supply chain data environments with significant business impact.
- Collaborative and inclusive working environment that values diversity, equity, and inclusion.
- Opportunity to develop expertise across cloud engineering, platform operations, data infrastructure, and reliability engineering.
- Competitive contract compensation, subject to experience and applicable contract terms.

