1,423,711open jobs
83,179companies
212,956added this week
Browse all
Salary
≈ $47k – $109k per year (Estimated)
Location
Remote (Bulgaria, Serbia, Spain, Portugal)
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 9, 2026. First seen by Alion on Oct 8, 2026. Cloudlinux scores B on the Alion truth index.

Overview
Company
Impact
Profile match
CloudLinux is on a mission to continually increase security, stability and availability of Linux servers and devices. Headquartered in Palo Alto, California, CloudLinux Inc. develops a hardened Linux distribution, Linux kernel live security patchi...
Backed by Runa Capital

CloudLinux builds Linux infrastructure and security products. You will join our Automation & Management Services cell, working closely with Dmitrii Petrov to solve problems across teams and services: cloud-cost data, infrastructure inventory, network policies and capacity workflows.

Check out our website for more informationhttps://cloudlinux.com/

Inside the Infrastructure Department, the Platform cell is a small team. We run the observability platform, the company's GitLab and the CI runners behind it, a few smaller engineering services, and the automation the department relies on for provisioning and configuration.

We are looking for a Platform Engineer for the PaaS team to take ownership of an agreed set of platform services: keep them reliable and within agreed service levels, make changes and recovery repeatable, and reduce recurring operational work through automation and self-service.

You get real freedom in how you implement things, and you own the result: you pick the approach, defend it in review, and answer for how the service behaves afterwards.

Most of the job is what site reliability engineering is about: services that run well day after day, and requests from product teams handled properly. Some of it is building. The observability platform and the runner cluster were both built from scratch within the last year, and there will be more of that. Expect the work to split between platform engineering, incident response, and helping engineers in other teams use what we run.

What you'll do

  • Run the observability platform. Keep it healthy, onboard teams, watch cost and capacity, and maintain the alerting that runs on top of it.
  • Run GitLab and the CI runner fleet. Upgrades, capacity, access, backups and restore drills.
  • Keep the rest of our services healthy, with the monitoring and runbooks a production service needs.
  • Deploy new services when they are requested. Research the options, pick a design, and stand the service up from scratch according to good practice: as code, monitored, backed up, documented.
  • Work with developers' requests. Access, onboarding, pipeline problems, new exporters and dashboards. Answer them, and turn the recurring ones into self-service.
  • Run incidents. Diagnose and mitigate impact, restore service safely, then complete the root-cause analysis and post-mortem. Deliver the prevention or detection improvements the incident calls for.
  • Ship everything as code, reviewed in merge requests. Plan and check before every change.
  • Write for engineers outside the team. Runbooks, onboarding guides, maintenance notices and status updates that people can act on.
  • Work with AI agents. Delegate collection and drafting to them, review their output as you would a colleague's merge request, and record what you learn where the team can find it.

Requirements

Must have:

  • Senior-level experience in infrastructure, platform or site reliability engineering, including at least one production service you were responsible for keeping up. We will ask you to walk us through it in detail: what broke, how you found out, and what you changed so it would not happen again.
  • Linux systems administration and debugging on bare metal and virtual machines. Much of our infrastructure is not Kubernetes.
  • Kubernetes in production delivered through GitOps, including cluster upgrades you performed yourself.
  • Infrastructure as code as your delivery form: Ansible and Terraform or OpenTofu, changes reviewed in merge requests.
  • GitLab administration and GitLab CI in production, self-hosted or SaaS. Deep experience with another CI system is acceptable if you can show the same depth.
  • Working knowledge of the Prometheus and Grafana ecosystem: you have run it for a team, written alert rules and dashboards, and can read PromQL. Depth here is welcome, and learnable.
  • Written technical explanation for engineers outside your team: runbooks, notices, answers to requests.
  • Strong communication and interpersonal skills. This role deals with people at least as much as with servers: most work starts as a conversation with a product team, and you need to understand what they actually need, agree scope, priority and timing with them, push back politely when a request should not be done as asked, and keep everyone informed while the work is in progress. We are looking for someone other teams enjoy working with.
  • Advanced use of AI engineering assistants such as Claude and Codex: providing context, breaking down tasks, designing agent loops, and delegating plans for unattended, end-to-end execution within defined scope and permissions, with clear stop conditions. You can explain, debug and test the resulting automation, and verify generated commands, scripts and conclusions before they touch production.
  • English - upper-intermediate or higher - to ensure clear communication of progress within the teams.

Nice to have:

  • Alerting design: SLOs, burn-rate alerts, thresholds sized from data.
  • MicroVM isolation for CI: Kata Containers, Firecracker or gVisor.
  • S3-compatible object storage operations: Ceph RGW or similar.
  • AWS with real cost work.
  • Self-hosted Sentry, or another Kafka, ClickHouse and Redis-backed application you have kept alive under load.
  • Python or Go for exporters and small internal services.

You do not need to have run every system on this list. Solid fundamentals and the judgment to pick up an unfamiliar service, make it observable and hand back a runbook matter more than matching every line.

What this role is not focused on:

  • Not a ticket-queue operator. Recurring requests get turned into self-service, not processed one by one forever.
  • Not a pure cloud or Kubernetes role. Bare metal and virtual machines are a large part of our infrastructure.
  • Not the DBA, the network engineer or the security engineer. Those teams run their own systems; we provide the platform they monitor them with.

Benefits

What's in it for you?

  • A focus on professional development.
  • Interesting and challenging projects.
  • Fully remote work with flexible working hours, which allows you to schedule your day and work from any location worldwide.
  • Paid 24 days of vacation per year, 10 days of national holidays, and unlimited sick leaves.
  • Compensation for private medical insurance.
  • Co-working and gym/sports reimbursement.
  • Budget for education.
  • The opportunity to receive a reward for the most innovative idea that the company can patent.

By applying for this position, you consent to the processing of your personal data as described in our Privacy Policy (https://cloudlinux.com/candidate-privacy-notice), which provides detailed information on how we maintain and handle your data.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,423,711 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Madrid
$103k – $191k per year • Remote (United States) • Full-Time • 5+ years exp • Bachelor's Degree • United States
Python
PowerShell
DevOps
Prometheus
Azure
CI/CD
AWS
Docker
Kubernetes
Grafana
Platform Engineering
AWS Lambda
Amazon CloudWatch
Apply
≈ $19k – $44k per year (Estimated) • Remote (India) • 5+ years exp • Bengaluru
Python
JavaScript
TypeScript
Node JS
Databases
DynamoDB
Amazon Aurora
AI/ML
LangGraph
LangChain
Model Context Protocol
AI Agents
AWS Bedrock
CrewAI
LLM
RAG
AWS Bedrock AgentCore
AWS Strands Agents
Agentforce
Structured Outputs
LLM Evaluation
LLM Guardrails
Machine Learning
DevOps
Rest API
WebRTC
AWS CDK
CloudFormation
CI/CD
AWS
AWS Lambda
Amazon S3
IAM
Amazon Kinesis
Amazon EventBridge
AWS Step Functions
API Gateway
Cybersecurity
Least Privilege
Apply
≈ $39k – $100k per year (Estimated) • Remote (Mexico) • Full-Time • 5+ years exp • Bachelor's Degree • Mexico
Python
Java
PowerShell
C#
Bash
C#
.NET
Databases
ElasticSearch
Apache Kafka
DevOps
Terraform
Azure DevOps
GitHub Actions
Azure
CI/CD
ArgoCD
Jenkins
AWS
Kubernetes
Platform Engineering
Amazon EKS
Azure AKS
Linux
Cybersecurity
ISO 27001
SOC 2
Management
Agile
Apply
≈ $39k – $100k per year (Estimated) • Equity • Remote (Brazil) • 12+ years exp
Python
Bash
Databases
InfluxDB
AI/ML
Time Series Forecasting
DevOps
Splunk
Terraform
Ansible
GCP
OpenShift
GitHub Actions
OpenTelemetry
VMWare
Prometheus
Azure
CI/CD
GitOps
Windows Server
Jenkins
AWS
Docker
Kubernetes
Grafana
Platform Engineering
Self-Healing
Telegraf
Incident Management
GitLab
Windows
Management
ServiceNow
ITIL
ITSM
Apply
IT Administrator 1 hour ago
≈ $43k – $98k per year (Estimated) • In office • Full-Time • Santa Cruz de Tenerife
DevOps
Windows
DNS
DHCP
Apply
In office • Full-Time • Bachelor's Degree • Hyderabad
Python
Apex
Apex
Salesforce Data Cloud
AI/ML
Cursor
Claude
AI Agents
Agentforce
DevOps
Terraform
GCP
Helm
Azure
CI/CD
Jenkins
Git
AWS
Docker
Kubernetes
Spinnaker
Analytics
Tableau
Management
Slack
Apply
≈ $18k – $46k per year (Estimated) • In office • 2+ years exp • Almaty
Databases
PostgreSQL
Redis
RabbitMQ
MinIO
MS SQL
DevOps
Ansible
Zabbix
Helm
Cilium
VMWare
cert-manager
CRI-O
etcd
Prometheus
GitLab CI
CI/CD
Docker
Kubernetes
Nginx
Grafana
kubeadm
eBPF
GitLab
Linux
DNS
DHCP
VPN
VLAN
BGP
Apache HTTP Server
Cybersecurity
Calico
Management
Agile
Scrum
Apply
≈ $16k – $36k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Pune
SQL
Databases
Couchbase
Apache Kafka
DevOps
Splunk
GCP
Prometheus
AWS
Kubernetes
Grafana
AppDynamics
Linux
Unix
QA
Sentry
Apply
Technical Architect 8 hours ago
≈ $81k – $170k per year (Estimated) • In office • Full-Time • 5+ years exp • PhD • Madrid
Python
JavaScript
TypeScript
SQL
Node JS
Apex
Apex
MuleSoft
AI/ML
Copilot
Claude
ChatGPT
AI Agents
LLM
OpenAI
Anthropic
Agentforce
DevOps
Rest API
CI/CD
Git
SOAP
Apply
≈ $68k – $166k per year (Estimated) • In office • Full-Time • 18+ years exp • PhD • Bengaluru • Hyderabad • Pune • Mumbai • Gurgaon
JavaScript
Apex
Apex
Salesforce Data Cloud
Salesforce Industries
AI/ML
Copilot
Cursor
Claude
AI Agents
Agentforce
Frontend
Web Components
Apply
≈ $54k – $107k per year (Estimated) • Remote (Bulgaria, Serbia, Spain, Portugal) • Madrid
Python
SQL
DevOps
Terraform
Ansible
OpenTofu
GitLab CI
CI/CD
AWS
Kubernetes
Grafana
FinOps
Linux
Apply
Remote (North America) • Full-Time • New York
DevOps
New Relic
Datadog
Dynatrace
PagerDuty
Docker
Cloudflare
GitHub
GitLab
Linux
Cybersecurity
Snyk
Aqua Security
Veracode
Sonatype Nexus IQ
Mend
Marketing
HubSpot
Apply
≈ $61k – $103k per year (Estimated) • Remote (Bulgaria, Poland, Spain, Armenia, Georgia, European time zones hours) • Full-Time • Warsaw
Python
Go
JavaScript
Node JS
Node JS
Commander.js
Databases
Redis
ElasticSearch
AI/ML
AI Agents
DevOps
Kibana
Debian
Jenkins
Grafana
GitLab
Linux
Management
Jira
Apply
≈ $140k – $255k per year (Estimated) • Remote (EMEA, North America) • Full-Time • Seattle
DevOps
Linux
Apply
≈ $41k – $116k per year (Estimated) • Remote (Brazil, Argentina, Colombia, Mexico, Peru) • Brasília
Python
Go
AI/ML
AI Agents
DevOps
Debian
Jenkins
Git
Docker
QEMU
Linux
Cybersecurity
CVE
Apply
≈ $46k – $92k per year (Estimated) • Hybrid • Full-Time • Madrid
PowerShell
Bash
DevOps
Ansible
Red Hat
Linux
Apply
≈ $23k – $41k per year (Estimated) • In office • Contractor • Madrid
Apply
≈ $22k – $64k per year (Estimated) • In office • Full-Time • 1+ year exp • Madrid
Apply
≈ $41k – $108k per year (Estimated) • Remote (Poland) • Full-Time • 6+ years exp • Madrid
Management
Agile
Apply
≈ $39k – $91k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Madrid
Apply
See all jobs
This is one of many
1,423,711 more open roles from verified company boards, updated every day.