We're looking for a Senior DevOps Engineer to lead the evolution of SpotDraft's infrastructure and platform reliability as we scale. In this role, you'll own critical infrastructure systems, drive cloud architecture decisions, and build scalable, secure, and highly available platforms that power our products and internal engineering workflows. You'll work closely with Engineering and AI teams to improve deployment efficiency, strengthen observability and reliability, and support next-generation AI-powered workloads across the organization. The ideal candidate brings strong hands-on experience with Kubernetes (K8s), Google Cloud Platform (GCP), ArgoCD, GitOps practices, and modern cloud-native infrastructure, preferably within fast-paced B2B product or SaaS environments.
The candidate will have responsibilities across the following functions:
Infrastructure Ownership and Platform Engineering:
- Own and manage SpotDraft's cloud infrastructure and internal platforms.
- Architect and scale secure, reliable infrastructure primarily on Google Cloud Platform (GCP).
- Manage Kubernetes (K8s) clusters and production workloads.
- Drive infrastructure automation using Terraform and GitOps practices.
- Improve platform scalability, reliability, and operational excellence.
AI Infrastructure and Deployment:
- Support and scale AI/ML infrastructure and deployments across the organization.
- Optimize infrastructure for AI-powered workloads and services.
- Ensure reliable deployment, monitoring, and scalability of AI applications.
CI/CD and Deployment Automation:
- Build and optimize CI/CD pipelines and deployment workflows.
- Manage GitOps deployments using ArgoCD.
- Automate application deployments using Helm, Kustomize, and Terraform.
Monitoring, Reliability, and Security:
- Improve monitoring, logging, and observability using Prometheus, DataDog, Grafana, and GCP tools.
- Troubleshoot production issues and drive long-term infrastructure improvements.
- Optimize infrastructure performance, security, and cloud costs.
Cross-Functional Collaboration:
- Partner with Engineering teams to improve developer experience and platform workflows.
- Drive DevOps best practices and operational standards across teams.
- Collaborate with leadership on infrastructure strategy and scaling initiatives.
Requirements:
- 2-4 years of experience in DevOps.
- Strong experience with Google Cloud Platform (GCP), Kubernetes (K8s), Docker, and cloud-native infrastructure.
- Hands-on experience with ArgoCD, GitOps workflows, Terraform, and CI/CD automation.
- Strong understanding of Linux systems, networking, cloud security, and infrastructure scalability.
- Experience with monitoring and observability tools such as Prometheus, DataDog, or Grafana.
- Proficiency in scripting languages such as Bash or Python.
- Familiarity with AI/ML infrastructure and deployment workflows is preferred.
- Prior experience in B2B product-first or SaaS companies is strongly preferred.
- Strong ownership mindset with excellent problem-solving skills.

