1,443,753open jobs
85,355companies
221,329added this week
Browse all
Location
Remote (Europe, United States)

Confirmed on the employer's own hiring board on Oct 11, 2026. First seen by Alion on Jan 13, 2026.

Overview
Company
Impact
Profile match
Nebius is an international technology company that specializes in building full-stack artificial intelligence infrastructure and high-performance cloud GPU platforms for AI model development. Headquartered in Amsterdam, Netherlands, the firm operates energy-efficient data centers across Europe and North America to provide scalable compute, storage, and software tools for machine learning workloads.

About Nebius:

Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.

Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.

Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.

About the role:

Token Factory is a part of  Nebius Cloud, one of the world’s largest GPU clouds, running tens of thousands of GPUs. We are building an inference platform that makes every kind of foundation model - text, vision, audio, and emerging multimodal architectures - fast, reliable, and effortless to deploy at massive scale.

Responsibilities:

  • Develop and optimizelow-level kernels and runtime components for AI inference 
  • Improve performance of inference engines GPU platforms 
  • Profile and debug system-level and hardware-level performance issues 
  • Integrate support for new hardware architectures (Hopper, Blackwell,Rubin) 
  • Collaborate with ML and backendteams to optimizeend-to-end execution 

Required Qualifications:

  • Strong proficiencyin C++,OR expertisein GPU programming with a focus on low-level high-performance coding and memory management 
  • Experience in GPU programming or systems-level software development, e.g.operating system internals, kernel modules, or device drivers 
  • Hands-on experience with profiling and debugging tools to identifyperformance issues on both CPUs and GPUs, and the ability to optimize code based on those findings. 
  • Solid understanding of CPU/GPU architecture and memory hierarchy 

Preferred Qualifications:  

  • Experience with GPUcomputingprogramming:CUDA, ROCm, CUTLASS, Cute, ThunderKittens, Triton, Pallas, Mosaic GPU 
  • Familiarity with ML inference runtimes (e.g. TensorRT, TVM) 
  • Knowledge of Linux internals, drivers, or compiler toolchains 
  • Experience with tools like perf, VTune, Nsight, or ROCmprofiler 
  • Familiarity with popular inference engines (e.g. such as vLLM, sglang, TGI) 

We conduct coding interviews as part of the process.

Benefits & Perks:

  • Competitive compensation
  • Career growth and learning opportunities
  • Flexibility and ownership
  • Collaborative and innovative culture
  • Opportunity to work on impactful AI projects
  • International environment and talented teams

What's it like to work at Nebius:

Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI 

Equal Opportunity Statement:

Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.

Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. 

If you need accommodations during the application process, please let us know.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,443,753 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
In your city
Remote (Greece) • Full-Time • Athens
DevOps
VMWare
Azure
Cybersecurity
Active Directory
Apply
≈ $41k – $105k per year (Estimated) • Remote (Greece) • Full-Time
Python
PowerShell
Databases
Databricks
DevOps
Terraform
Azure DevOps
GitHub Actions
Prometheus
Azure
CI/CD
Docker
Kubernetes
Grafana
Bicep
Azure AKS
Cybersecurity
GDPR
Apply
≈ $40k – $101k per year (Estimated) • Remote (Serbia) • Serbia
DevOps
Terraform
GitHub Actions
Azure
CI/CD
AWS
Bicep
Cybersecurity
Least Privilege
Threat Modeling
Apply
$73k – $90k per year • Remote (France) • Full-Time • France
Python
Bash
Databases
PostgreSQL
ElasticSearch
Apache Kafka
DevOps
Terraform
Ansible
GCP
Prometheus
Grafana
Configuration Management
Linux
Apply
≈ $113k – $209k per year (Estimated) • Equity • Remote (United States) • United States
AI/ML
NCCL
DevOps
Platform Engineering
Linux
Apply
Staff Engineer 1 hour ago
≈ $71k – $149k per year (Estimated) • In office • Belgrade
Python
JavaScript
C++
Python
Flask
DevOps
CI/CD
AWS
Management
Agile
Apply
In office • Master's Degree
C++
DevOps
RTOS
Robotics
EtherCAT
Apply
In office
C++
C++
Qt
DevOps
Azure DevOps
Azure
CI/CD
Jenkins
Docker
Kubernetes
Linux
Management
Agile
Scrum
Apply
In office • Master's Degree
C++
DevOps
RTOS
Management
Agile
Apply
≈ $33k – $91k per year (Estimated) • In office • 5+ years exp • Shanghai
C#
C++
IoT
OPC UA
Apply
Automation Engineer 3 days ago
Remote (likely United States) • Full-Time • 3+ years exp
PowerShell
DevOps
Rest API
Terraform
GCP
Azure
CI/CD
Git
Bicep
IAM
Cybersecurity
Microsoft Sentinel
ISO 27001
Microsoft Defender
SOC 2
Least Privilege
Microsoft Defender for Cloud
Microsoft Entra ID
DLP
Management
Google Workspace
Power Automate
SharePoint
Service Desk
Apply
Remote (likely United States) • Full-Time • 3+ years exp
PowerShell
DevOps
Rest API
Terraform
GCP
Azure
CI/CD
Git
Bicep
IAM
Cybersecurity
Microsoft Sentinel
ISO 27001
Microsoft Defender
SOC 2
Least Privilege
Microsoft Defender for Cloud
Microsoft Entra ID
DLP
Management
Google Workspace
Power Automate
SharePoint
Service Desk
Apply
$112k – $140k per year • Remote (United States) • Birmingham
DevOps
HPC
Linux
Apply
≈ $89k – $204k per year (Estimated) • Remote (location not specified) • Full-Time
Go
C++
DevOps
gRPC
Linux
BGP
Apply
$147k – $224k per year • Remote (United States) • 5+ years exp • New York
AI/ML
AI Agents
RAG
Tavily
DevOps
Terraform
CI/CD
GitOps
Kubernetes
Apply
See all jobs
This is one of many
1,443,753 more open roles from verified company boards, updated every day.