Confirmed on the employer's own hiring board on Oct 10, 2026. First seen by Alion on Jun 25, 2026. Pinterest scores B on the Alion truth index.
Join TvScientific, the first CTV advertising platform designed for performance marketers. As a Site Reliability Engineer, you will help operate, scale, and continuously improve our cloud-native platform built on AWS, Kubernetes/EKS, and ArgoCD-driven GitOps workflows. Your responsibilities will include ensuring the reliability and performance of production infrastructure, managing GitOps-based deployment workflows, supporting infrastructure provisioning, and participating in incident response. The ideal candidate will have 4+ years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure, and strong expertise in Kubernetes and AWS.
Missions
- Contribuer à l'amélioration de la fiabilité, de l'évolutivité, de l'automatisation, de l'observabilité et de la maturité opérationnelle de l'infrastructure et de l'écosystème de livraison.
- Assurer la fiabilité, la disponibilité et la performance de l'infrastructure de production et des services de la plateforme, y compris l'exploitation et l'évolutivité des plateformes Kubernetes.
- Participer à la réponse aux incidents, à l'analyse des causes profondes et aux initiatives d'amélioration post-incident, tout en réduisant le travail opérationnel grâce à l'automatisation.
Profil recherché
- The ideal candidate is a hands-on engineer with solid production experience and a strong foundation in building and supporting resilient platforms using infrastructure as code, automation, and modern Kubernetes operational practices- Solid scripting and automation skills using Bash and/or Python
- Experience with Terraform/Terragrunt for infrastructure provisioning and environment management
- Demonstrated ability to use AI to improve speed and quality in your day-to-day workflow for relevant outputs
- High integrity and ownership: you protect sensitive data, avoid over-reliance on AI, and remain accountable for final decisions and deliverables
- Strong hands-on experience with Helm
- Experience with monitoring, alerting, and observability in production environments
- Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems
- Demonstrated ownership mindset with experience handling incidents and resolving production issues
- 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure
- Good expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration
- Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams
- Experience implementing and operating ArgoCD within a GitOps delivery model
- Experience with Kubernetes multi-tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns
- Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions
- Bachelor’s degree in computer science, engineering, a related field or equivalent experience
- Strong hands-on experience operating AWS in production environments
- Strong track record of critical evaluation and verification of AI-assisted work (e.g., testing, source-checking, data validation, peer review)

