1,432,387open jobs
84,237companies
218,256added this week
Browse all
Salary
≈ $83k – $218k per year (Estimated)
Location
In office
Seniority
Principal · 10+ years exp
Visa
Licensed UK visa sponsor

Confirmed on the employer's own hiring board on Oct 10, 2026. First seen by Alion on Aug 26, 2026. Nscale scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Nscale is a London-based AI infrastructure company that builds and operates GPU data centres and runs a full-stack AI cloud offering managed inference, Kubernetes and Slurm clusters, bare-metal instances and dedicated GPU capacity. Founded in 2024 by Josh Payne and Nathan Townsend, it develops sites in Norway, the UK, South Korea and North America, works with Microsoft and NVIDIA, and acquired Anyscale in July 2026 to extend its cloud platform. Its hiring spans data centre design and construction, electrical and infrastructure operations, HPC and storage engineering, networking, solutions architecture, legal, finance and marketing.

About Nscale

Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale enables AI-focused companies to achieve superior results by reducing the complexity of AI development. Our GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.

At Nscale, our Engineering team plays a critical role in designing, deploying, and operating the infrastructure and software platforms that power our customers and enable AI workloads at scale.

We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you'll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you'll be contributing to building the technology that powers the future.

About the Role

We are hiring a Principal Network Engineer to act as a senior technical authority for Nscale's AI-optimised network infrastructure.

Our Network Engineering team is responsible for the design, validation, and ongoing operation of the networking services underpinning both our internal management platform and customer-facing cloud infrastructure. This includes high-performance Ethernet fabrics, InfiniBand, RoCE, WAN connectivity, and large-scale data centre networking.

In this role, you'll set technical direction across the low-latency, high-bandwidth networks supporting large-scale AI training and inference workloads. You'll own critical technical domains end-to-end and help raise the bar for architecture, automation, operational rigour, and engineering standards across Nscale.

This is a deeply technical Principal-level role combining hands-on engineering with broad architectural influence. You'll define reference architectures, drive consistency across sites, lead complex technical decisions and escalations, and mentor engineers while partnering closely with Deployment, Data Centre Operations, Platform Engineering, Systems, Storage, and technology vendors.

What you'll be doing

AI & High-Performance Network Architecture

  • Define, design, validate, and evolve large-scale InfiniBand, RoCE, and Ethernet fabric architectures at rack, row, and data centre scale.

  • Design networks that integrate closely with bare-metal provisioning and cluster management systems.

  • Own technical direction for high-performance Ethernet fabrics, including BGP, EVPN, VXLAN, LACP, and QoS.

  • Establish reference architectures and engineering standards that can be implemented consistently across Nscale's data centre estate.

  • Identify systemic risks and architectural gaps and drive durable solutions that improve scalability, reliability, and operational simplicity.

Network Automation & Infrastructure as Code

  • Lead Nscale's network automation strategy using a GitOps operating model.

  • Build and guide Python and Ansible tooling for provisioning, configuration validation, compliance, and operational workflows.

  • Drive version-controlled configuration and CI/CD-based network change across multi-vendor environments.

  • Apply Infrastructure-as-Code and Network-as-Code principles to reduce manual intervention and improve operational consistency.

  • Continuously identify opportunities to automate repetitive operational tasks and reduce reactive toil.

Network Security & Edge Infrastructure

  • Design and engineer perimeter and network security infrastructure across WAN and data centre edge environments.

  • Own architecture across firewalls, NAT, VPN, security policies, and multi-tenant segmentation.

  • Design highly available and scalable security architectures appropriate for mission-critical AI infrastructure.

Reliability, Observability & Operations

  • Lead complex technical escalations and root-cause analysis for network performance, reliability, and stability issues.

  • Establish measurable SLOs and operational standards for network services.

  • Set technical direction for network observability, telemetry, monitoring, and alerting.

  • Ensure clear visibility into fabric health, traffic patterns, performance, and capacity.

  • Develop runbooks, automation, and engineering improvements that systematically reduce operational toil.

  • Act as a senior 3rd/4th line escalation point for complex networking issues.

Network Data & Configuration Management

  • Ensure the accuracy and reliability of source-of-truth network inventory and configuration data.

  • Establish structured engineering and change-management practices for network configuration.

  • Ensure network changes are controlled, auditable, repeatable, and scalable across multiple sites.

Technical Leadership & Collaboration

  • Partner with Deployment, Data Centre Operations, Platform Engineering, Systems, Storage, and vendors on new site delivery and platform evolution.

  • Lead architecture and design reviews for significant network initiatives.

  • Mentor engineers and raise technical capability across the wider networking organisation.

  • Lead complex technical decisions and incidents spanning networking, systems, storage, and AI/HPC workloads.

  • Influence engineering strategy and standards across teams without relying on formal authority.

About You

Required Experience

  • 10+ years of network engineering experience, with significant depth in HPC, AI, hyperscale, or large-scale data centre environments.

  • Extensive hands-on experience with RDMA-aware networking for AI/HPC workloads, including InfiniBand and/or RoCE.

  • Experience with subnet managers and fabric orchestration technologies such as OpenSM or NVIDIA UFM.

  • Expert-level understanding of modern data centre routing and control planes, including BGP, EVPN-VXLAN, and Clos/spine-leaf architectures.

  • Production experience with network platforms such as Cumulus, Nokia, or Arista EOS.

  • Strong network automation expertise using Python and Ansible.

  • Experience with Git-based workflows and modern Infrastructure-as-Code and CI/CD tooling such as Terraform, GitLab CI, or GitHub Actions.

  • Deep experience designing and engineering firewall infrastructure using platforms such as Juniper SRX and/or Palo Alto.

  • Experience designing telemetry and observability solutions for high-throughput, performance-sensitive environments.

  • Proven ability to lead complex technical decisions and incidents across multiple engineering disciplines.

Architecture & Technical Leadership

  • Demonstrated experience defining network architecture, engineering standards, and technical strategy beyond a single project or data centre.

  • Ability to balance performance, reliability, operability, scalability, and delivery velocity when making technical decisions.

  • Experience influencing engineering teams and technical direction without formal authority.

  • Strong communication skills with the ability to explain complex technical trade-offs to engineers, operators, and senior stakeholders.

  • Proven ability to mentor and develop other engineers through design reviews, technical guidance, incident leadership, and knowledge sharing.

Personal Attributes

  • Deeply technical and comfortable remaining hands-on at Principal level.

  • Highly analytical with a structured approach to solving complex infrastructure problems.

  • Strong sense of ownership and accountability.

  • Comfortable operating in a fast-paced environment where priorities and requirements evolve quickly.

  • Pragmatic and able to balance architectural excellence with business and delivery requirements.

  • Passionate about building next-generation infrastructure for AI and ML at scale.

What we can offer you

At Nscale, you'll find a collaborative, supportive, and innovative environment where your contributions spark real impact. We're building something extraordinary, and we want you at the core.

Highly competitive package (base + equity) with reviews every 12 months.

Join one of the fastest-growing AI infrastructure companies - your opportunity to design and build the high-performance networks powering some of the world's most demanding AI workloads.

Expect a dynamic progression plan tailored to your ambitions. Shape network architecture, establish engineering standards, solve complex infrastructure challenges, and influence the technical direction of Nscale's global AI platform.

Human-First Flexibility: We treat you as humans first. Our flexible workplace trusts Nscalers to deliver, giving you the autonomy to shape your day around life's moments.

Equal Opportunities Statement

We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds.

If there's anything we can do to accommodate your specific situation, please let us know.

The responsibilities outlined in this job description are not exhaustive and are intended to provide a general overview of the position. The employee may be required to perform additional duties, tasks, and responsibilities as assigned by management, consistent with the skills and qualifications required for the role.

For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice:  Here.

Nscale does not accept unsolicited candidate submissions from recruitment agencies.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,432,387 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Networking
Similar stack
Same company
In your city
≈ $60k – $157k per year (Estimated) • In office • Manchester
Management
Gmail
Apply
≈ $68k – $178k per year (Estimated) • In office • Full-Time • Teddington
Apply
≈ $66k – $174k per year (Estimated) • In office
Management
Confluence
Apply
$93k – $97k per year • Hybrid • Full-Time • 5+ years exp • Belfast
Python
Apply
≈ $75k – $180k per year (Estimated) • Hybrid • Full-Time • Bristol
Apply
Ingeniero/a DevOps 4 days ago
≈ $42k – $97k per year (Estimated) • In office • Contractor • Madrid
Python
DevOps
Terraform
Ansible
OpenShift
Helm
Loki
Prometheus
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Grafana
Linux
Apply
≈ $80k – $221k per year (Estimated) • In office • Contractor • Tunis
Python
JavaScript
TypeScript
MATLAB
Python
FastAPI
PyQt
Frontend
Angular
DevOps
GCP
GitLab CI
CI/CD
Docker
Google Cloud Run
Apply
≈ $83k – $157k per year (Estimated) • Remote (United Kingdom) • Full-Time • 5+ years exp • Bachelor's Degree • United Kingdom
Python
SQL
Databases
Snowflake
AI/ML
dbt
Machine Learning
Analytics
Alteryx
Apply
≈ $90k – $172k per year (Estimated) • In office • Full-Time • 6+ years exp • Master's Degree • Fort Worth
Python
MATLAB
AI/ML
AI Agents
Supervision
Robotics
Digital Twin
Design
SolidWorks
Apply
≈ $47k – $131k per year (Estimated) • Remote (United States) • Full-Time • PhD • United States
Python
AI/ML
Supervision
Apply
≈ $84k – $220k per year (Estimated) • In office • 7+ years exp
Python
DevOps
gRPC
Terraform
Ansible
GCP
Azure
CI/CD
Git
AWS
Kubernetes
Platform Engineering
Apply
≈ $95k – $228k per year (Estimated) • In office • 10+ years exp • Ireland
Python
AI/ML
InfiniBand
DevOps
Terraform
Ansible
GitHub Actions
GitLab CI
CI/CD
GitOps
Git
Platform Engineering
HPC
VPN
BGP
Apply
≈ $140k – $304k per year (Estimated) • In office • 12+ years exp • New York
Python
AI/ML
Edge AI
DevOps
Ansible
BGP
OSPF
Apply
≈ $125k – $268k per year (Estimated) • In office • New York
Python
DevOps
Prometheus
Grafana
VLAN
BGP
OSPF
MPLS
Apply
≈ $147k – $319k per year (Estimated) • In office • 10+ years exp • New York
AI/ML
InfiniBand
Edge AI
DevOps
HPC
BGP
OSPF
Apply
See all jobs
This is one of many
1,432,387 more open roles from verified company boards, updated every day.