1,253,116open jobs
72,533companies
215,675added this week
Browse all
Salary
≈ $147k – $319k per year (Estimated)
Location
In office (New York)
Seniority
Principal · 10+ years exp
Visa
H-1B filings in 12 months: 10 · for this role: 4

Confirmed on the employer's own hiring board on Oct 5, 2026. First seen by Alion on Aug 10, 2026. Nscale scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Nscale is a London-based AI infrastructure company that builds and operates GPU data centres and runs a full-stack AI cloud offering managed inference, Kubernetes and Slurm clusters, bare-metal instances and dedicated GPU capacity. Founded in 2024 by Josh Payne and Nathan Townsend, it develops sites in Norway, the UK, South Korea and North America, works with Microsoft and NVIDIA, and acquired Anyscale in July 2026 to extend its cloud platform. Its hiring spans data centre design and construction, electrical and infrastructure operations, HPC and storage engineering, networking, solutions architecture, legal, finance and marketing.

About Nscale

Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers.  Nscale enables AI-focused companies to achieve superior results by reducing the complexity of AI development. Our GPU cloud bolsters technical capabilities and directly supports strategic business outcomes, including cost management, rapid innovation, and environmental responsibility.

We thrive on a culture of relentless innovation, ownership, and accountability, where every team member takes pride in their work and drives it with excellence and urgency. As an Nscaler, you’ll build trust through openness and transparency, where everyone is inspired to do their best work. If you join our team, you’ll be contributing to building the technology that powers the future.

About The Role

The Network Operations and Engineering teams at Nscale operate some of the most demanding networking environments in the industry, supporting tightly coupled GPU clusters where network performance directly impacts customer outcomes.

We’re looking for a Principal Network Engineer - AI Infrastructure to provide technical leadership across Nscale’s high-speed networking domain.

This role is focused on owning the reliability, scalability, and long-term evolution of our Infiniband and RDMA-based network fabrics. You will operate as a technical authority, influencing architecture, standards, and operational practices across teams while tackling the most complex network challenges in the platform.

What You'll Be Doing

  • Owning the technical direction and operational strategy for Nscale’s AI interconnect networks
  • Designing, reviewing, and evolving large-scale Infiniband and RoCE fabric architectures to support future growth and workload demands
  • Acting as the senior escalation point for the most complex network incidents, guiding deep technical investigations and systemic fixes
  • Driving cross-team initiatives to improve fabric reliability, performance predictability, and operational maturity
  • Defining standards for hardware configuration, congestion control, routing, firmware lifecycle management, and change safety
  • Partnering with SRE, Compute Platform, and Network Architecture teams to influence end-to-end system design
  • Mentoring senior and mid-level network engineers, raising the bar for operational rigor and technical excellence
  • Driving measurable improvements in uptime, latency consistency, capacity efficiency, and incident reduction

About You (Skills / Qualifications)

  • 10+ years of experience in network engineering, with deep focus on HPC, AI, or hyperscale data centre networking
  • Expert-level operational and architectural experience with Infiniband and/or large-scale RoCE fabrics
  • Deep understanding of RDMA internals, congestion management, and fabric-level failure modes
  • Strong expertise in modern data centre routing and control planes (BGP, OSPF, ECMP)
  • Proven ability to debug and resolve cross-layer issues spanning hardware, firmware, kernel, and application communication libraries
  • Demonstrated ability to lead complex technical initiatives across teams without direct authority
  • A systems-level mindset, balancing performance, reliability, scalability, and operational cost

Nice to Have

  • Extensive experience with NVIDIA/Mellanox networking platforms in production AI or HPC environments
  • Deep familiarity with distributed training frameworks and GPU communication patterns
  • Experience designing network observability systems for high-cardinality, high-throughput environments
  • Prior experience influencing platform or infrastructure strategy at scale

What We Can Offer You

At Nscale, you'll find a collaborative, supportive, and innovative environment where your contributions spark real impact. We're building something extraordinary, and we want you at the core.

  • Highly competitive package (base + equity) with reviews every 12 months. 
  • Join the fastest-growing tech startup, your chance to push boundaries, collaborate with brilliant minds, and make your mark on cutting-edge AI. 
  • Expect a dynamic progression plan tailored to your ambitions. Grow by trying new things, leading, challenging the status quo, and owning your impact, always with our full support. 

Equal Opportunities Statement

We strongly encourage applications from people of colour, the LGBTQ+ community, people with disabilities, neurodivergent people, parents, carers, and people from lower socio-economic backgrounds.

If there’s anything we can do to accommodate your specific situation, please let us know.

The responsibilities outlined in this job description are not exhaustive and are intended to provide a general overview of the position. The employee may be required to perform additional duties, tasks, and responsibilities as assigned by management, consistent with the skills and qualifications required for the role.

For information on how Nscale handles candidate personal data, please see our Employee & Candidate Privacy Notice:  Here.

Nscale does not accept unsolicited candidate submissions from recruitment agencies.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,253,116 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Networking
Similar stack
Same company
New York
Network Architect 2 hours ago
$152k – $208k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Atlanta
DevOps
Azure
AWS
TCP/IP
BGP
OSPF
Cybersecurity
Zero Trust
Apply
NOC Lead Engineer 21 days ago
≈ $89k – $193k per year (Estimated) • In office • Full-Time • 4+ years exp • Fort Lauderdale
DevOps
Nagios
TCP/IP
BGP
OSPF
MPLS
Apply
$110k – $140k per year • Equity • In office • TS/SCI • Full-Time • 4+ years exp • High School Diploma • San Antonio
Assembly
Assembly
Keystone Engine
DevOps
TCP/IP
DNS
DHCP
Cybersecurity
Active Directory
Management
ServiceNow
Outlook
SharePoint
ITIL
Microsoft Office
Apply
$100k – $145k per year • In office • Top Secret • 8+ years exp • Bachelor's Degree • Norfolk
DevOps
Red Hat
VMWare
Windows Server
Windows
Apply
Network Architect 2 months ago
$120k – $170k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Phoenix
DevOps
GCP
Azure
AWS
Cybersecurity
Zero Trust
Apply
≈ $25k – $50k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Moscow
Python
Go
Bash
Databases
PostgreSQL
DevOps
gRPC
Terraform
Ansible
Zabbix
Helm
Loki
Podman
etcd
Packer
Prometheus
VictoriaMetrics
HAProxy
Docker
Kubernetes
Grafana
KVM
eBPF
SLI/SLO/SLA
Linux
Astra Linux
TCP/IP
DNS
DHCP
VLAN
BGP
OSPF
Cybersecurity
Wireshark
Keycloak
Tcpdump
HashiCorp Vault
Zero Trust
Apply
≈ $16k – $37k per year (Estimated) • Remote (EAEU) • Moscow
Python
AI/ML
Machine Learning
DevOps
SLI/SLO/SLA
GitHub
Linux
Unix
DNS
BGP
OSPF
Cybersecurity
Wireshark
Tcpdump
Active Directory
LDAP
Apply
$130k – $145k per year • In office • Top Secret • 7+ years exp • Washington
DevOps
VLAN
BGP
OSPF
MPLS
Apply
In office • 8+ years exp
Python
DevOps
Terraform
Ansible
Azure
AWS
TCP/IP
DNS
DHCP
BGP
OSPF
Cybersecurity
Wireshark
Zero Trust
PKI
Apply
≈ $41k – $106k per year (Estimated) • Remote (likely Poland) • Full-Time • Warsaw
Python
Python
FastAPI
Django
Databases
PostgreSQL
ClickHouse
DevOps
Ansible
Prometheus
CI/CD
Docker
Grafana
Configuration Management
Linux
Unix
TCP/IP
BGP
OSPF
Apply
≈ $140k – $304k per year (Estimated) • In office • 12+ years exp • New York
Python
AI/ML
Edge AI
DevOps
Ansible
BGP
OSPF
Apply
≈ $125k – $268k per year (Estimated) • In office • New York
Python
DevOps
Prometheus
Grafana
VLAN
BGP
OSPF
MPLS
Apply
$150k – $210k per year • In office • 8+ years exp • Bachelor's Degree
Python
AI/ML
InfiniBand
DevOps
TCP/IP
BGP
OSPF
MPLS
Apply
$150k – $210k per year • In office • 8+ years exp
Python
AI/ML
InfiniBand
DevOps
Terraform
Ansible
GitHub Actions
GitLab CI
CI/CD
GitOps
Git
Platform Engineering
HPC
VPN
BGP
Apply
≈ $95k – $228k per year (Estimated) • In office • 10+ years exp • Ireland
Python
AI/ML
InfiniBand
DevOps
Terraform
Ansible
GitHub Actions
GitLab CI
CI/CD
GitOps
Git
Platform Engineering
HPC
VPN
BGP
Apply
≈ $87k – $234k per year (Estimated) • In office • Full-Time • New York
Marketing
LinkedIn
Apply
≈ $72k – $158k per year (Estimated) • In office • 4+ years exp • New York
Management
Microsoft Office
Apply
$36k per year • In office • 1+ year exp • New York
Apply
$117k – $178k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • New York • Boston • Washington • San Francisco
AI/ML
Claude Code
AI Agents
Agentforce
Analytics
Tableau
Marketing
Salesforce
Apply
$158k – $211k per year • Remote (United States) • Full-Time • PhD • New York • Boston
Apex
Apex
MuleSoft
AI/ML
AI Agents
Agentforce
Analytics
Tableau
Management
Slack
Apply
See all jobs
This is one of many
1,253,116 more open roles from verified company boards, updated every day.