599,985open jobs
30,608companies
86,410added this week
Browse all
Salary
$84k – $184k per year (Estimated)
Location
In office (Toronto)
Seniority
Staff · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match

At Shakudo, we're building the world's first operating system for data and AI. We use the term "operating system" in the truest sense: just like iOS, Windows, or Linux, Shakudo's end-to-end OS provides ever-evolving, fully automated, best-in-class open-source components tailored to each business's unique needs.

We are seeking an Infrastructure Engineer to join our Business Automation team to own and operate the internal systems, infrastructure, and AI Gateway product that power Shakudo at scale. This is a hands-on role for someone who thrives on keeping production systems reliable, secure, and fast. You will be responsible for everything from physical servers and DGX machines to CI/CD pipelines and customer-facing AI Gateway infrastructure. You will also contribute directly to product hardening, security, and DevOps practices across the platform.

At Shakudo, our culture is proactive, collaborative, and supportive - we succeed together by building strong partnerships and solving complex challenges. We expect high ownership: you will be hands-on, driving outcomes directly rather than delegating or waiting for direction. Individual contribution matters here - your work will have a visible, measurable impact on the company's operations and product.

Key Responsibilities

  • Maintain and operate internal services for the rest of the Shakudo employees, including proprietary applications for sales and ETL pipelines
  • Maintain and operate DGX machines that host LLMs for the team's use
  • Maintain and operate Shakudo's product for Shakudo's internal use, and contribute to product hardening, security, and DevOps practices
  • Maintain and operate physical servers for Kubernetes clusters and ensure uptime
  • Create CI/CD pipelines for internal deployments
  • Maintain and operate the AI Gateway product for customers, ensure uptime, and contribute to product roadmap

Qualifications

  • 8+ years of experience across software, data, platform, or AI engineering roles
  • 5+ years of strong experience with Kubernetes cluster operation and DevOps, and bare-metal server operations
  • Experience operating production infrastructure at scale, including physical servers, GPU clusters, and CI/CD systems
  • Strong background in security hardening, observability, and reliability engineering
  • Proficiency in Rust is preferred
  • Experience with AI/ML infrastructure, including LLM hosting and inference serving is preferred

Why Shakudo Stands Out

    Work with cutting-edge technologies in machine learning and high-performance computing. Contribute to a platform that transforms how organizations leverage data and AI. Join a dynamic team that values innovation, efficiency, and diversity.

    Shakudo offers a high-impact package: competitive salary, meaningful equity so you share in the upside of transformational technology, and comprehensive health benefits that have you fully covered. We provide a flexible vacation policy-because building transformational technology requires supporting the people who build it. More importantly, you'll work on technology that matters.

    This role is based onsite in Toronto to support the high security requirements of our clients and enable effective collaboration. We have a welcoming office environment with a very focused and passionate team, doing meaningful, impactful work together.

    Shakudo is an equal opportunity employer and encourages candidates of all backgrounds to apply. We foster diversity and inclusivity and welcome applications from a broad range of backgrounds and experiences.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
599,985 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Toronto
$98k – $245k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Tunis
JavaScript
TypeScript
AI/ML
LLM
Frontend
Redux
Tailwind CSS
Next.js
React.js
Radix UI
shadcn/ui
Redux Toolkit
Mobile
State Management
DevOps
WebSockets
CI/CD
Docker
Kubernetes
Apply
$98k – $244k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Tunis
Python
Python
SQLAlchemy
FastAPI
Pydantic
Databases
PostgreSQL
AI/ML
Model Context Protocol
Prompt Engineering
AI Agents
LLM
DevOps
Kong
Azure
CI/CD
Docker
Kubernetes
API Gateway
Cybersecurity
Microsoft Entra ID
Apply
In office • Full-Time • 3+ years exp • Bachelor's Degree • Tunis
Python
JavaScript
TypeScript
Python
FastAPI
AI/ML
Prompt Engineering
AI Agents
LLM
Frontend
Next.js
React.js
DevOps
CI/CD
Docker
Kubernetes
Platform Engineering
QA
Playwright
Jest
Pytest
Apply
$56k – $164k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Tunis
Python
AI/ML
Model Context Protocol
AI Agents
DevOps
Helm
Azure DevOps
GitHub Actions
Azure
CI/CD
GitOps
ArgoCD
AWS
Docker
Kubernetes
Amazon EKS
Azure AKS
Apply
$80k – $214k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Tunis
Python
AI/ML
Model Context Protocol
AI Agents
DevOps
Helm
Azure DevOps
GitHub Actions
Azure
CI/CD
GitOps
ArgoCD
AWS
Docker
Kubernetes
Amazon EKS
Azure AKS
Apply
Solutions Architect 17 days ago
$30k – $69k per year (Estimated) • In office • Full-Time • 4+ years exp • Bengaluru
Python
JavaScript
TypeScript
AI/ML
MLFlow
dbt
Frontend
React.js
DevOps
GCP
Helm
Azure
CI/CD
AWS
Kubernetes
Apply
$165k – $310k per year (Estimated) • In office • Full-Time • 8+ years exp • Menlo Park
Databases
Snowflake
Databricks
DevOps
Kubernetes
Cybersecurity
HIPAA
Apply
Account Executive 2 months ago
$112k – $236k per year (Estimated) • In office • Full-Time • 5+ years exp • Menlo Park
Apply
$90k – $190k per year (Estimated) • In office • Full-Time • 8+ years exp • Toronto
Python
Java
TypeScript
Scala
Databases
Snowflake
Databricks
AI/ML
Spark
LLM
RAG
Agentic Workflows
DevOps
GCP
PagerDuty
Azure
AWS
Kubernetes
Apply
Solutions Architect 3 months ago
$102k – $207k per year (Estimated) • In office • Full-Time • 4+ years exp • Toronto
Python
JavaScript
TypeScript
AI/ML
MLFlow
dbt
Frontend
React.js
DevOps
GCP
Helm
Azure
CI/CD
AWS
Kubernetes
Apply
Deployment Engineer 6 hours ago
$40k – $116k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Toronto
Python
TypeScript
AI/ML
AI Agents
OpenAI
Scale AI
Apply
$24k per year • In office • Full-Time • High School Diploma • Toronto
Apply
$60k – $70k per year • Remote • Full-Time • 2+ years exp • Bachelor's Degree • Toronto
Apply
$52k – $100k per year (Estimated) • Remote/Hybrid • Contractor • 3+ years exp • Toronto
Analytics
A/B Testing
Microsoft Excel
Management
Slack
Jira
Apply
$32k – $68k per year (Estimated) • Remote/Hybrid • Part-Time • Toronto
AI/ML
OpenAI
Apply
See all jobs
This is one of many
599,985 more open roles from verified company boards, updated every day.