719,538open jobs
42,887companies
102,020added this week
Browse all
Salary
$113k – $249k per year (Estimated)
Location
Remote/Hybrid (New York, United States)
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 23, 2026. First seen by Alion on May 6, 2026. Montauk Capital scores C on the Alion truth index.

Overview
Company
Impact
Profile match
Montauk Capital is building solutions across the full stack of the Electron Economy, ensuring its success today, tomorrow, and for the future of our planet.

Head of Infrastructure

Full Time, NYC / NJ Area Preferred

About Montauk Capital

Montauk Capital builds and backs companies at the forefront of the Electron Economy, the generational shift towards electrified, intelligent technologies reshaping industries and driving unprecedented demand for energy. Our team combines deep investing acumen with decades of operating experience to give founders the strategic clarity and hands-on support that accelerates the building of enduring companies of consequence.

About Stealth Edge AI Co

Co-founded by Montauk Capital, Stealth Edge AI Co is a pre-seed venture specialized in modular, metro-edge AI capabilities. By leveraging existing infrastructure for inference deployment, Edge AI provides low-latency, SLA-guaranteed performance across diverse GPU SKUs and colocation environments. Our technology intelligently routes traffic based on demand proximity and real-world network limitations, bypassing the heavy power and infrastructure requirements of traditional hyperscalers. Currently initiating operations with pilot nodes in NYC, we are executing a city-by-city expansion strategy with plans for a broader multi-metro rollout.

About the Role

We’re building the automation, orchestration, and monitoring layer that unifies disparate metro edge GPU nodes into a single software-managed compute platform. You’ll own the definition, design, implementation, and execution of the hardware and infrastructure buildout, executing strategy across edge data center requirements, GPU selection, supply chain, technical implementation, operational maintenance and deployment as we scale. You’ll take the foundational groundwork and execute across the entire hardware and infrastructure side of our company, transforming our roadmap into production scale compute for AI inferencing.

You’ll ensure the GPU clusters deliver on customer requirements, are highly-available, and will be the hands on expert for the hardware side of our business. Most importantly, you’ll turn our high-level plans into real, technical execution, and will play a key role in making supply chain decisions about infrastructure and how we deploy, scale, and support it.

What You’ll Do

  • Own GPU infrastructure design and implementation details from planning through deployment

  • Own hardware selection, configuration, and deployment across early compute infrastructure

  • Help turn early technical groundwork into a functioning deployed system

  • Own the GPU roadmap we use to entice customers and build partnerships

  • Deploy, operate, and tune GPU clusters for both bare-metal and internal software stack

  • Own resilient networking implementation from each site to the cluster, including a robust OOB network for constant monitoring and management

  • Manage deployments at production scale

  • Interface with site ops on power, cooling, and connectivity

  • Build the automation and monitoring stack for distributed edge nodes

  • Own the supply chain for all infrastructure gear

  • Manage third party hardware vendors on provisioning, maintenance and break-fix support

What You’ll Bring

You’re a strong infrastructure engineer experienced with hardware deployment, data center environments, GPU selection, systems setup and design. You can manage the implementation details end to end and have ownership over the entire process. If AI infrastructure is your jam and you've built systems in production, we want to talk.

  • Strong infrastructure engineering experience and systems-level technical judgment

  • Experience deploying or managing compute infrastructure in real-world environments

  • Experience with data center, hardware, or GPU-based systems implementation

  • Experience owning GPU provisioning, hardware selection, and systems configuration

  • GPU scheduling and orchestration specifics: GPU type awareness, memory management, topology considerations, placement strategies for multi-GPU jobs, and fragmentation minimization

  • Bare-metal provisioning lifecycle: IPMI/Redfish, BMC-based remote management, PXE boot, and automated OS deployment workflows

  • On-board storage

  • Observability stack: distributed configuration and troubleshooting, plus monitoring, alerting, and tracing

  • Deployment planning, Hardware configuration, Operational troubleshooting

  • Linux systems depth: RHEL/Ubuntu, low-level troubleshooting, shell scripting

  • Security and operational best practices for bare metal

  • Deployment tooling at production scale

  • Networking fundamentals for inference workloads and OOB management

  • Startup / 0→1 DNA: You ship fast and communicate clearly.

Why Join Us

  • Category-Defining Opportunity: Solving the AI inference bottleneck without the burden of power and infrastructure constraints

  • Massive Market Opportunity: AI spending projected to exceed hundreds of billions annually, 54GW of AI Inference demand expected by 2030

  • Studio Support: Leverage Montauk Capital's resources, network, and operational expertise during critical early stages

  • Competitive compensation + equity: True ownership over what you build

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
719,538 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
$22k – $54k per year (Estimated) • Remote • 1+ year exp • Bachelor's Degree • Athens
Python
PowerShell
Bash
DevOps
Linux
Windows
Unix
Cybersecurity
SIEM
Management
Agile
Apply
$90k – $182k per year (Estimated) • Remote • 3+ years exp • Bachelor's Degree
AI/ML
Red Teaming
DevOps
Kali Linux
Linux
Windows
Cybersecurity
Burp Suite
Metasploit
Nessus
Cobalt Strike
Covenant
Management
Agile
Apply
$60k – $70k per year • In office • 3+ years exp • PhD • North Aurora
DevOps
SLI/SLO/SLA
Management
Outlook
Apply
$44k – $57k per year • In office • Full-Time • 5+ years exp • PhD • Tokyo
JavaScript
TypeScript
Frontend
Next.js
React.js
DevOps
New Relic
Incident Management
SLI/SLO/SLA
Management
Kanban
QA
Sentry
Apply
Senior IT Auditor 4 days ago
In office • 3+ years exp
SQL
DevOps
Linux
Windows
Unix
Cybersecurity
ISO 27001
Management
Agile
ITIL
Apply
$186k – $347k per year (Estimated) • Remote/Hybrid • Full-Time • New York
AI/ML
Edge AI
DevOps
SLI/SLO/SLA
HPC
Apply
Co-Founder Zero Proof 3 months ago
In office • Cofounder
AI/ML
AI Agents
Apply
$186k – $347k per year (Estimated) • In office • Full-Time • New York
AI/ML
Edge AI
DevOps
SLI/SLO/SLA
Apply
$164k – $303k per year (Estimated) • In office • Full-Time • New York
AI/ML
Edge AI
DevOps
SLI/SLO/SLA
Apply
$194k – $361k per year (Estimated) • Remote/Hybrid • Full-Time • New York
JavaScript
Rust
C++
Node JS
Node JS
Electron
AI/ML
vLLM
CUDA Toolkit
Triton Inference Server
Quantization
SGLang
TensorRT
TensorRT-LLM
Ray
CUDA
Triton
Edge AI
Speculative Decoding
KV Cache
DevOps
Kubernetes
SLI/SLO/SLA
Apply
$106k – $209k per year (Estimated) • In office • 4+ years exp • Master's Degree • New York
Apply
$80k – $120k per year • In office • Full-Time • 1+ year exp • New York
AI/ML
RAG
OpenAI
Management
n8n
Marketing
Reddit
Apply
$150k – $300k per year • Equity • Remote • Full-Time • 3+ years exp • New York
AI/ML
RAG
DevOps
Grafana
Management
Monday.com
n8n
Apply
GTM Associate (US) 4 hours ago
$70k – $120k per year • In office • Full-Time • 1+ year exp • New York
AI/ML
RAG
OpenAI
DevOps
Grafana
Marketing
Reddit
Apply
$70k – $110k per year • In office • Full-Time • 1+ year exp • New York
AI/ML
RAG
OpenAI
DevOps
Docker
Marketing
Reddit
Apply
See all jobs
This is one of many
719,538 more open roles from verified company boards, updated every day.