386,061open jobs
10,141companies
47,930added this week
Browse all
Salary
$79k – $168k per year (Estimated)
Location
Remote (Canada)
Seniority
Senior · 6+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Jobgether is a Belgian recruitment platform built entirely around remote and flexible work, aggregating openings from thousands of employers that allow work from outside an office. Its matching engine ranks roles against a candidate's skills, seniority and stated preferences on location and flexibility, rather than leaving people to filter a keyword search, and it verifies how genuinely remote each posting is. The company also runs an AI screening layer that shortlists applicants for employers, and publishes research and guidance on distributed work practices alongside the job marketplace itself.

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Platform Engineer, Cloud Infrastructure based in Canada.

This is a senior, hands-on platform engineering opportunity focused on building and operating resilient cloud-native infrastructure at scale.

You’ll own production Kubernetes platforms and the underlying systems that engineering teams rely on to deliver reliable products.

The role combines deep infrastructure engineering with software development in Go, Python, or Java.

You’ll shape platform architecture across networking, service mesh, observability, infrastructure as code, and production operations.

You’ll also lead complex, multi-quarter initiatives, including cloud migrations, reliability improvements, and platform modernization.

Success means creating scalable capabilities, improving developer velocity, and ensuring systems remain secure, observable, and dependable.

Working fully remotely, you’ll collaborate with highly technical distributed teams while mentoring engineers and influencing technical direction.

Accountabilities

    • Design, build, and operate production Kubernetes clusters, including networking, workload isolation, resource management, scheduling, and multi-region architectures.
    • Work deeply with Kubernetes internals, including CNI networking, NetworkPolicy enforcement, cluster behavior under load, resource quotas, and custom controllers or operators.
    • Design and operate service mesh capabilities, including mTLS, workload identity, service-account authentication and authorization, and traffic management.
    • Optimize containerized workloads for performance, cost efficiency, scalability, and effective resource utilization.
    • Develop and maintain production-grade services, Kubernetes controllers, middleware, and platform components using Go, Python, or Java.
    • Build HTTP, REST, and gRPC interfaces used by internal engineering teams, while maintaining strong unit and integration test coverage.
    • Lead platform-level incident response, troubleshoot distributed systems using logs, metrics, traces, and profiling, and produce postmortems that drive lasting improvements.
    • Define and implement SLOs, actionable alerting, dashboards, metrics, and distributed tracing to strengthen platform reliability.
    • Own infrastructure as code at scale, creating reusable modules and evolving them as platform requirements change.
    • Build and improve CI/CD and GitOps workflows that enable engineering teams to release software safely, frequently, and reliably.
    • Plan and execute cloud migration initiatives, including moving production workloads between providers or environments while minimizing disruption and maintaining reliability.
    • Partner with product, security, and infrastructure teams to gather requirements, evaluate trade-offs, conduct design reviews, and establish technical direction.
    • Mentor engineers and contribute to higher standards of engineering, reliability, testing, and operational excellence.
    • Requirements

      • 6+ years of professional experience in software, platform, infrastructure, or site reliability engineering, including substantial experience operating production distributed systems.
      • Demonstrated experience building and operating production Kubernetes platforms, rather than simply deploying applications onto existing clusters.
      • Strong production programming experience in Go, Python, or Java, with the ability to work effectively in complex existing codebases.
      • Experience taking ambiguous technical problems from initial design through production implementation and ongoing operation.
      • Proven experience planning and executing cloud migrations involving production workloads across providers or environments.
      • Degree in Computer Science, Engineering, or a related discipline, or equivalent practical experience.
      • Strong understanding of Kubernetes networking, CNI, NetworkPolicy, resource management, scheduling, and cluster behavior.
      • Hands-on experience with service mesh technologies such as Istio, Envoy, Linkerd, or equivalent, including mTLS and workload identity.
      • Strong Linux fundamentals, including cgroups and resource management.
      • Significant infrastructure-as-code experience using Terraform or an equivalent technology.
      • Production experience with at least one major cloud platform such as AWS, GCP, or Azure; multi-cloud experience is highly valuable.
      • Experience with observability technologies such as Prometheus, Grafana, OpenTelemetry, and PromQL or comparable query languages.
      • Production experience with relational databases, particularly PostgreSQL or managed PostgreSQL-compatible services, including an understanding of replication and failover.
      • Strong Docker and container tooling experience across the software delivery lifecycle.
      • Advanced debugging, troubleshooting, and performance-profiling capabilities.
      • Experience with Kubernetes controllers, operators, API-server extensions, or Go testing frameworks such as Ginkgo and Gomega is an asset.
      • Additional Python expertise and the ability to read Java are valuable.
      • Experience with identity and access technologies such as SSO, Keycloak, OIDC, SAML, Vault, or cloud-based secrets management is preferred.
      • Experience designing high-availability and disaster-recovery architectures across regions or cloud providers is an advantage.
      • Familiarity with monorepos, Bazel, compliance frameworks such as SOC 2 or GDPR, and reliability considerations for LLM-backed systems is beneficial.
      • Experience with Alibaba Cloud and large-scale data technologies such as Apache Spark or Apache Flink is highly preferred.
      • Strong analytical and problem-solving skills, excellent written and verbal communication, and the ability to produce clear design documents and postmortems.
      • Comfortable working independently in a distributed, fast-moving, highly technical environment while collaborating effectively across teams.
      • Benefits

        • Fully remote position open to candidates in Canada.
        • Senior-level opportunity with significant technical ownership and influence over cloud infrastructure.
        • Opportunity to work with Kubernetes, cloud platforms, service mesh, observability, infrastructure as code, and modern software engineering practices.
        • Exposure to complex, multi-quarter initiatives including cloud migrations and large-scale reliability improvements.
        • Collaborative distributed environment with opportunities to mentor engineers and shape engineering standards.
        • Flexible remote collaboration across an international technical team.
        • Opportunity to contribute to innovative cloud-native infrastructure and emerging technologies, including AI-enabled systems.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
386,061 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$111k – $215k per year (Estimated) • Remote • Full-Time • 6+ years exp • Bachelor's Degree
Go
Java
Python
Databases
PostgreSQL
AI/ML
Flink
LLM
Spark
DevOps
AWS
Azure
Bazel
CI/CD
Docker
Envoy
GCP
GitOps
Grafana
gRPC
Istio
Kubernetes
Linkerd
OpenTelemetry
Platform Engineering
Prometheus
Service Mesh
SRE
Terraform
Cybersecurity
GDPR
Keycloak
SOC 2
Apply
$22k – $118k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Node JS
Python
SQL
JavaScript
Python
FastAPI
AI/ML
LangChain
Llama
LLM
OpenAI
Prompt Engineering
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
Grafana
Kubernetes
Apply
$95k – $205k per year (Estimated) • Remote • Full-Time • 7+ years exp
Bash
C#
Go
JavaScript
Python
TypeScript
Python
FastAPI
Databases
PostgreSQL
RabbitMQ
Redis
AI/ML
LLM
Model Context Protocol
RAG
DevOps
ArgoCD
Azure
Azure AKS
Bitbucket
CI/CD
Cloudflare
Docker
Envoy
GitOps
Grafana
Helm
Jenkins
Kubernetes
OpenTelemetry
Platform Engineering
Prometheus
Rest API
Terraform
QA
Sentry
Apply
$130k – $150k per year • Remote • Full-Time • 7+ years exp
C#
SQL
C#
.NET
Databases
Azure Cosmos DB
AI/ML
OpenAI
DevOps
Azure
Azure DevOps
CI/CD
GitHub
Platform Engineering
Terraform
Cybersecurity
HIPAA
QA
Postman
Apply
$93k – $163k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree
DevOps
AWS
Azure
GCP
VMWare
Windows Server
Cybersecurity
Crowdstrike
Sophos
Management
Microsoft Teams
Marketing
Salesforce
Apply
$95k – $205k per year (Estimated) • Remote • Full-Time • 7+ years exp
Bash
C#
Go
JavaScript
Python
TypeScript
Python
FastAPI
Databases
PostgreSQL
RabbitMQ
Redis
AI/ML
LLM
Model Context Protocol
RAG
DevOps
ArgoCD
Azure
Azure AKS
Bitbucket
CI/CD
Cloudflare
Docker
Envoy
GitOps
Grafana
Helm
Jenkins
Kubernetes
OpenTelemetry
Platform Engineering
Prometheus
Rest API
Terraform
QA
Sentry
Apply
$94k – $168k per year (Estimated) • Remote • Full-Time • 7+ years exp
Apply
$130k – $222k per year (Estimated) • Remote • Full-Time • 7+ years exp
Apply
$36k – $44k per year • Remote • Full-Time • 2+ years exp • High School Diploma
Bash
PHP
PHP
WordPress
Apply
$115k – $129k per year • Remote • Full-Time • 10+ years exp
PowerShell
Python
SQL
C#
Java
C#
.NET
Entity Framework Core
Java
Flyway
Liquibase
Databases
Amazon Aurora
MS SQL
Oracle
PostgreSQL
AI/ML
AI Agents
DevOps
Amazon CloudWatch
Amazon EC2
AWS
CloudFormation
IAM
SLI/SLO/SLA
Terraform
Cybersecurity
SOC 2
Apply
See all jobs
This is one of many
386,061 more open roles from verified company boards, updated every day.