1,438,441open jobs
85,144companies
219,560added this week
Browse all
Salary
$180k – $210k per year
Location
In office (Palo Alto)
Seniority
Senior · 10+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Oct 11, 2026. First seen by Alion on Oct 7, 2026. DeepInfra scores C on the Alion truth index.

Overview
Company
Impact
Profile match
DeepInfra is a company founded in 2022 that runs a low-cost inference cloud for open machine learning models. Developers call hosted language, embedding, speech and image models through a simple interface and pay per token or per second without managing graphics hardware. Its focus on cheap serving of open-weight models made it popular with startups building on Llama, Mistral and Qwen.

About DeepInfra

DeepInfra is building the foundation for companies to use modern AI in production - simply, reliably, and at scale. Our team has deep experience building large systems that serve hundreds of millions of users, and we're bringing that same level of rigor to a rapidly evolving AI inference space. Our mission is to make advanced AI available to people and teams everywhere.

We're an early, tight-knit team where you can influence product direction, try bold ideas, and drive meaningful work forward quickly. If you want to join a fast-growing company at a defining moment, we'd love to talk.

DeepInfra is backed by leading investors including A.Capital, Felicis, 500 Global, Georges Harik, Samsung Next, Supermicro, Upper90, Peak6, SVAngel, and NVIDIA.

Why This Role Matters

Demand for AI inference is growing fast, and DeepInfra runs on GPU infrastructure we own and operate. Every new site, rack, and network link we bring online translates directly into capacity our customers can build on - so how quickly, reliably, and cost-effectively we expand our physical footprint is central to how we grow.

This role spans the full lifecycle of that expansion. You'll find and evaluate data center space, negotiate with colocation and hardware vendors, design and build the networks that tie our clusters together, and run the projects that take a site from signed contract to serving production traffic. You'll work closely with Engineering and our co-founders, and you'll be as comfortable reviewing a colo contract as you are on the data center floor with a console cable.

What You'll Do

  • Design, build, and operate high-performance data center networks - including spine/leaf fabrics, BGP, transit and peering, and network security - across multiple colocation sites.
  • Own site bring-ups end to end: low- and high-level designs, capacity planning, rack and power layouts, cabling plans, hardware installation, turn-up, and handoff to production.
  • Operate and improve the underlying infrastructure that our compute platform runs on, with strong Linux fundamentals, monitoring, and automation.
  • Run infrastructure projects across multiple vendors and sites, managing timelines, dependencies, shipment and logistics coordination, and risk.
  • Travel to data centers to install, connect, and troubleshoot hardware alongside remote-hands and vendor teams.
  • Serve as the primary technical point of contact for data center vendors and for customers whose infrastructure we manage.
  • Build repeatable playbooks, standards, and automation that make every new site faster to stand up than the last.

What You Bring

  • Experience: 10+ years building and operating infrastructure for 24/7 production environments, including significant hands-on experience with data center infrastructure and networking.
  • Data Center Networking: Deep experience designing and operating high-performance data center networks, including spine/leaf or Clos architectures, BGP, and network security. Experience with EVPN/VXLAN, transit, and peering is strongly preferred.
  • Site Deployment: Experience bringing colocation or data center infrastructure online, including network design, rack layouts, cabling, hardware installation, turn-up, and coordination with on-site or remote-hands teams.
  • Infrastructure Operations: Strong Linux fundamentals and the ability to troubleshoot across networking, compute, storage, and the physical layer.
  • Automation: Strong scripting and automation skills using Python, Bash, Ansible, or similar tools.
  • Project Ownership: Proven ability to drive complex infrastructure projects across multiple vendors, dependencies, and locations from planning through production.
  • Vendor Management: Experience working directly with data center, network, or hardware vendors. Experience with capacity planning, procurement, or commercial negotiations is a strong plus.

Preferred Qualifications

  • Experience deploying GPU infrastructure, including high-speed interconnects such as InfiniBand or RoCE and high-density power and cooling environments.
  • On-prem private cloud experience with KVM, Proxmox, OpenStack, Ceph, or similar technologies.
  • Kubernetes and containers in production, along with observability tooling such as Prometheus and Grafana.
  • Experience managing internet resources at scale, including ASNs, IP prefixes, IRR records, and routing policy with ISPs and IXPs.
  • Experience owning hardware procurement, colocation contracts, capacity planning, or infrastructure vendor negotiations.
  • Source, evaluate, and secure data center and colocation space across power, cooling, density, connectivity, and cost.
  • Lead commercial negotiations with data center, network, and hardware vendors across pricing, contract terms, SLAs, and expansion options.

How We Work

Three traits define the people who thrive here, and this role leans on all three.

Initiative. We take ownership and step in where we can add value. Whether it's starting something new, improving what exists, or helping move ideas forward, we aim to be proactive and thoughtful in how we contribute.

Drive. We're energized by hard problems. Building AI infrastructure is complex, and we lean into that. We care about doing things well, moving fast, and continuously improving - because solving meaningful challenges is what motivates us.

Grit. Things don't always work on the first try - and that's expected. You'll balance negotiations, network design, and hands-on site work at once, often across multiple vendors, time zones, and shifting delivery timelines. We stay persistent, adapt quickly, and learn as we go.

Compensation

The base pay range for this role is $180,000 - $210,000 per year.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,438,441 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Palo Alto
$88k – $123k per year • In office • 3+ years exp • Bachelor's Degree • Vista
SQL
Apply
$112k – $154k per year • In office • 4+ years exp • Bachelor's Degree • Vista
Python
PowerShell
DevOps
Windows Server
Cybersecurity
Active Directory
IoT
MQTT
Apply
$180k – $224k per year • In office • United States
Python
DevOps
CI/CD
eBPF
Linux
Apply
≈ $113k – $209k per year (Estimated) • Equity • Remote (United States) • United States
AI/ML
NCCL
DevOps
Platform Engineering
Linux
Apply
$147k – $224k per year • Remote (United States) • 5+ years exp • New York
AI/ML
AI Agents
RAG
Tavily
DevOps
Terraform
CI/CD
GitOps
Kubernetes
Apply
In office • Internship • Bachelor's Degree • Nanjing
Python
DevOps
Linux
TCP/IP
DHCP
Wi-Fi
Apply
≈ $108k – $210k per year (Estimated) • In office • 5+ years exp • Sydney
Python
C#
C++
AI/ML
Machine Learning
DevOps
Linux
Apply
≈ $16k – $43k per year (Estimated) • In office • Full-Time • Novosibirsk
Python
Rust
Python
Django
Ruff
Databases
PostgreSQL
Redis
DevOps
Rest API
Ansible
Git
Cybersecurity
Keycloak
LDAP
QA
Pytest
Apply
≈ $37k – $102k per year (Estimated) • In office • Contractor • Associate's Degree • Singapore
Python
JavaScript
TypeScript
SQL
Node JS
Python
Flask
Django
AI/ML
Pandas
NumPy
RAG
Frontend
Vue.js
Angular
React.js
DevOps
Terraform
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Apply
In office • Full-Time • Bachelor's Degree • Singapore
Python
Java
DevOps
CI/CD
Analytics
ETL/ELT
AWS Glue
Management
Agile
Scrum
Apply
$90k – $110k per year • In office • Full-Time • Palo Alto
Marketing
LinkedIn
Apply
In office • Internship • Bachelor's Degree • Sofia
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
CUDA Toolkit
SciPy
Multimodal AI
Diffusers
Transformers
TensorFlow
NumPy
PyTorch
CUDA
OpenAI
NCCL
Edge AI
Machine Learning
DevOps
Git
Management
Agile
Apply
≈ $18k – $38k per year (Estimated) • In office • Full-Time • 2+ years exp • Sofia
Python
AI/ML
Function Calling
Speech Recognition
LLM
Text-to-Speech
Structured Outputs
Prompt Caching
Tool Use
DevOps
Rest API
Loki
Prometheus
Grafana
Apply
Corporate Controller 17 days ago
$200k – $290k per year • In office • Full-Time • 8+ years exp • Palo Alto
Cybersecurity
SOC 2
Management
Stripe
Apply
Director, FP&A 17 days ago
$200k – $290k per year • In office • Full-Time • 8+ years exp • Palo Alto
Apply
$258k per year • In office • 3+ years exp • Bachelor's Degree • Palo Alto
Python
Go
Bash
DevOps
Terraform
Puppet
GCP
GitHub Actions
VMWare
Prometheus
Azure
CI/CD
AWS
Kubernetes
Grafana
Configuration Management
Amazon EKS
Google GKE
Azure AKS
Proxmox VE
OpenStack
IAM
Amazon CloudWatch
Windows
DNS
VPN
Cybersecurity
Wazuh
NIST CSF
CIS Benchmarks
PCI DSS
GDPR
HIPAA
Zero Trust
Active Directory
LDAP
PKI
SIEM
Apply
$80k – $110k per year • In office • Full-Time • 4+ years exp • Palo Alto
AI/ML
Claude
ChatGPT
Perplexity
Apply
≈ $164k – $325k per year (Estimated) • In office • Full-Time • Palo Alto
Marketing
Salesforce
Apply
≈ $49k – $80k per year (Estimated) • In office • Internship • Master's Degree • New York • Seattle • Palo Alto • San Francisco
Python
SQL
AI/ML
Machine Learning
Apply
$156k – $320k per year • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • New York • Seattle • Palo Alto • San Francisco
JavaScript
Frontend
React.js
Apply
See all jobs
This is one of many
1,438,441 more open roles from verified company boards, updated every day.