368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$107k – $242k per year (Estimated)
Location
Remote/Hybrid (Canada)
Seniority
Staff · 8+ years exp
Overview
Company
Impact
Profile match
ELEKS is a global software engineering and technology consulting firm headquartered in Tallinn, Estonia, and founded in 1991. The company provides a range of services including custom software development, product design, data science, cybersecurity, and quality assurance across various industries such as fintech, healthcare, and logistics. Operating through a network of delivery centers and offices across Europe, North America, and the Middle East, the firm employs over 2,000 specialists to support enterprise-level digital transformation projects.

ELEKS is looking for an Infrastructure/GPU Cluster/Platform Operations Lead in Canada.

Alberta-based candidates are strongly preferred (Calgary or Edmonton). Canada-based candidates will also be considered.

ABOUT CLIENT

Our customer is building a next-generation AI platform that enables organizations to securely develop, govern, and operationalize artificial intelligence while ensuring that sensitive data and organizational knowledge remain fully under their control. The platform combines advanced AI capabilities with enterprise-grade governance, security, and data sovereignty to support mission-critical decision-making.

The solution serves government organizations and enterprise customers operating in highly regulated and security-sensitive environments, where reliability, accountability, and trust are essential. The platform supports intelligent decision-making across strategic planning, workforce intelligence, and organizational operations, helping customers leverage AI without compromising security, compliance, or control over their data.

REQUIREMENTS

  • 8+ years of Infrastructure Engineering or Platform Operations experience
  • Experience managing GPU clusters for AI workloads
  • Strong Kubernetes administration skills
  • Experience with NVIDIA GPU technologies and CUDA ecosystem
  • Experience with cloud infrastructure (Azure, AWS or GCP)
  • Knowledge of storage, networking, and high-performance computing environments
  • Experience implementing Infrastructure as Code (Terraform or similar)
  • Strong operational leadership skills
  • Experience supporting AI platform infrastructure
  • Upper-Intermediate or higher level of English

RESPONSIBILITIES

  • Lead GPU infrastructure design and operations
  • Manage Kubernetes-based AI platform environments
  • Optimize infrastructure for AI training and inference workloads
  • Define operational standards and reliability practices
  • Collaborate with AI engineering teams
  • Implement monitoring, security, and disaster recovery strategies
  • Lead infrastructure capacity planning
  • Support technical roadmap and infrastructure evolution
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$105k – $252k per year • Remote • Full-Time • 18+ years exp • Bachelor's Degree
Python
Java
Java
Gradle
DevOps
Ansible
AWS
CI/CD
CloudFormation
Configuration Management
Docker
GitHub Actions
GitLab CI
Helm
Jenkins
Kubernetes
Platform Engineering
Terraform
GitHub
GitLab
Cybersecurity
Sonatype Nexus IQ
Management
Confluence
Jira
Apply
$54k – $175k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$35k – $113k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$35k – $116k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$61k – $200k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$85k – $172k per year (Estimated) • Remote/Hybrid • 5+ years exp
DevOps
Azure
Apply
$31k – $107k per year (Estimated) • In office • 4+ years exp
Java
Java
Spring Boot
Spring Cloud
DevOps
GitLab
Apply
$44k – $110k per year (Estimated) • Remote
DevOps
Azure
Azure DevOps
CI/CD
Jenkins
Rest API
GitHub
Management
ServiceNow
Apply
$50k – $87k per year (Estimated) • Remote • Full-Time • 5+ years exp • Poland
TypeScript
DevOps
CI/CD
QA
Playwright
Apply
$103k – $230k per year (Estimated) • Remote • Contractor • 5+ years exp
C#
SQL
C#
.NET
Entity Framework Core
Databases
MS SQL
AI/ML
Windsurf
DevOps
Azure
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.