368,910open jobs
9,449companies
47,822added this week
Browse all
Salary
$61k – $120k per year (Estimated)
Location
Remote (Poland)
Seniority
Staff · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
ALTER GPU CENTER is a cloud computing infrastructure company. It provides GPU-based computing resources and related services for high-performance workloads.

About the role

We are looking for a Lead Linux System Administrator to take technical ownership of the Linux environment supporting large-scale GPU infrastructure used for AI training and inference workloads.

This role combines hands-on system administration with team leadership. You will be responsible for the stability, performance, security, and day-to-day management of Linux-based GPU servers, while also supporting and mentoring a team of administrators working in a complex production environment.

Responsibilities

  • Lead, mentor, and support a team of Linux System Administrators responsible for GPU infrastructure operations

  • Manage the full Linux server lifecycle, including provisioning, patching, configuration management, hardening, and performance tuning

  • Maintain and optimize the NVIDIA GPU software stack, including drivers, CUDA, cuDNN, NCCL, and GPU management tools such as DCGM and nvidia-smi

  • Support and manage MIG and GPU time-slicing configurations where needed

  • Develop and maintain automation for bare-metal provisioning, OS image management, and server configuration using tools such as Ansible, Terraform, and scripting

  • Tune Linux systems for demanding workloads, including kernel parameters, local storage, parallel file systems, networking, and scheduler settings

  • Troubleshoot complex issues across hardware, drivers, the operating system, and cluster-level services

  • Work closely with DevOps/SRE, Site Operations, and AI/ML teams to ensure smooth integration between OS-level infrastructure and higher-level orchestration platforms

  • Support security hardening, vulnerability management, patch compliance, and operational standards across the server fleet

  • Participate in on-call support and contribute to continuous improvements in reliability, performance, and operational efficiency

Requirements

  • 7+ years of hands-on experience in Linux system administration in production environments

  • At least 3 years of experience in a technical lead, lead administrator, or people leadership role

  • Strong expertise in administering Linux systems at scale

  • Hands-on experience with NVIDIA GPUs in Linux environments, including drivers, CUDA ecosystem components, and GPU management tools

  • Strong experience with Ansible or other configuration management tools

  • Good scripting skills in Python and/or Bash

  • Experience with Infrastructure as Code and infrastructure automation

  • Good understanding of high-performance computing, storage systems, and high-speed networking technologies such as InfiniBand or RoCE

  • Experience supporting AI/ML or HPC workloads

  • Ability to troubleshoot complex production issues and work effectively in a high-availability environment

  • English proficiency at least at a communicative level is required, as you will be working in an international team

Nice to have

  • Experience with cluster management and orchestration tools such as Slurm, Kubernetes, or Run:ai

  • Familiarity with bare-metal provisioning tools and large server fleet management

  • Experience in AI infrastructure companies, hyperscalers, or HPC/research environments

  • Knowledge of Linux performance tuning for GPU-accelerated workloads

  • Higher education in Computer Science, Engineering, or a related field

What we offer

  • Benefits package

  • Opportunity to lead Linux infrastructure supporting advanced AI workloads at scale

  • Work with modern GPU hardware and software stacks in a technically demanding environment

  • Collaboration with experienced engineers across infrastructure, platform, and AI teams

  • A dynamic workplace with room for ownership, technical influence, and professional growth

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,910 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Łódź
$131k – $237k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Chantilly • Columbia
Bash
Python
TypeScript
JavaScript
Databases
PostgreSQL
Frontend
Angular
DevOps
Ansible
ArgoCD
AWS
Azure
buildah
CI/CD
Docker
Docker Swarm
FluxCD
GitLab
GitLab CI
Grafana
Helm
Kubernetes
Loki
Prometheus
Terraform
Twelve-Factor App
Podman
Apply
$215k – $240k per year • In office • Full-Time • 5+ years exp • Seattle
Python
Rust
AI/ML
Flyte
OpenAI
DevOps
ArgoCD
AWS
Azure
Buildkite
CI/CD
Docker
GCP
Helm
Kubernetes
Terraform
Apply
$82k – $215k per year (Estimated) • In office • 10+ years exp • Sydney
Node JS
Python
SQL
TypeScript
JavaScript
Frontend
Angular
Vue.js
Mobile
Clean Architecture
DevOps
AWS
CI/CD
Apply
$28k – $81k per year (Estimated) • In office • Full-Time • Madrid
Python
SQL
DevOps
Git
Analytics
ETL/ELT
Power BI
Apply
$37k – $83k per year (Estimated) • In office • Full-Time • Bristol
Python
SQL
Analytics
Power BI
Tableau
Apply
$29k – $74k per year (Estimated) • Remote • Full-Time • Poland
Python
SQL
Python
Django
Databases
PostgreSQL
DevOps
GitHub
GitLab
Apply
Lead DevOps Engineer 2 months ago
$61k – $120k per year (Estimated) • Remote • Full-Time • 8+ years exp • Łódź
Python
AI/ML
KServe
Ray
InfiniBand
DevOps
Ansible
CI/CD
Configuration Management
Crossplane
GitOps
Grafana
Incident Management
Kubernetes
Loki
OpenTelemetry
Platform Engineering
Prometheus
Pulumi
SLURM
Terraform
HPC
Apply
$44k – $84k per year (Estimated) • Remote • Full-Time • 4+ years exp • Łódź
Bash
Python
AI/ML
CUDA Toolkit
cuDNN
InfiniBand
NCCL
DevOps
Configuration Management
Debian
Grafana
Kubernetes
Prometheus
Red Hat
SLURM
Ubuntu
HPC
Apply
DevOps Engineer 2 months ago
$46k – $87k per year (Estimated) • Remote • Full-Time • 4+ years exp • Łódź
Python
AI/ML
KServe
Ray
DevOps
Ansible
CI/CD
GitOps
Grafana
Kubernetes
Loki
OpenTelemetry
Platform Engineering
Prometheus
SLURM
Terraform
HPC
Cybersecurity
Crowdstrike
Snyk
Apply
Front-end Developer 2 months ago
$52k – $87k per year (Estimated) • Remote • Part-Time • Łódź
JavaScript
TypeScript
Frontend
React.js
DevOps
Docker
GitHub
GitLab
Apply
$69k – $111k per year (Estimated) • Remote/Hybrid • Full-Time • Łódź
Python
SQL
Databases
Snowflake
AI/ML
Streamlit
DevOps
Azure
Analytics
Power BI
Apply
UX Designer 6 days ago
In office • Full-Time • Łódź
AI/ML
Copilot
Midjourney
Marketing
Hotjar
Apply
$68k – $122k per year (Estimated) • Remote • Full-Time • Łódź
Cybersecurity
ISO 27001
Apply
$54k – $86k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Łódź
C#
SQL
C#
.NET
Entity Framework Core
AI/ML
Semantic Kernel
OpenAI
DevOps
Azure
Azure DevOps
CI/CD
Grafana
OpenTelemetry
Rest API
Apply
$32k – $38k per year • Remote • Łódź
PHP
TypeScript
JavaScript
PHP
Symfony
AI/ML
Claude
Claude Code
Frontend
Angular
Apply
See all jobs
This is one of many
368,910 more open roles from verified company boards, updated every day.