368,291open jobs
9,430companies
50,343added this week
Browse all
Salary
$120k – $180k per year
Location
In office (San Francisco)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Models. Agents. GPUs. Whatever you need — we've got it. Ghost agent VMs, Maestro model routing, Engine wholesale GPUs, and Academy, on one platform.

Inference.net is hiring a Senior Full-Stack (Frontend-Focused) Engineer

Help us build beautiful, performant web experiences that give users super-powers over our globally distributed LLM inference platform. If you love shipping React apps that feel snappy at planet-scale, we’d love to meet you.

AboutInference.net

We combine idle GPU capacity from around the world into a single cohesive plane of compute capable of serving models like DeepSeek and Llama 4. At any moment, 5,000+ GPUs and hundreds of terabytes of VRAM are connected to our network.

We’re a small, well-funded team working in-person from downtown San Francisco (hybrid flexibility as needed). Investors include a16z CSX and Multicoin. We’re high-agency, collaborative, and obsessed with craft-whether that’s a distributed scheduler or a pixel-perfect UI.

What you’ll do

  • Own the user experience - design, build, and polish the dashboards, consoles, and customer-facing apps that let users observe, configure, and pay for inference at scale.

  • Ship end-to-end features - from Figma wireframe to React component to backend API and database migration.

  • Design a component system in React + Tailwind that supports rapid iteration and a cohesive design language.

  • Optimize performance - SSR, code-splitting, hydration, and WebSocket-driven real-time updates that hold up under millions of requests per day.

  • Collaborate across disciplines - work shoulder-to-shoulder with distributed-systems engineers, product designers, and founders to turn complex infrastructure into delightful product.

  • Level-up the team - lead design reviews, mentor junior engineers, and introduce best practices for testing, accessibility, and observability.

What we’re looking for

Must-Have

  • 5+ years building production React applications

  • Deep knowledge of Tailwind CSS & modern CSS architecture

  • Typescript mastery and strong fundamentals in JS/DOM/APIs

  • Experience designing REST/JSON or gRPC backends (Node, Go, or similar)

  • AuthN/AuthZ design (OIDC, JWT)

  • Product sense: you care about UX details & accessibility

Nice-to-Have

  • Experience with Tanstack / Next.js

  • Data-viz libraries (Recharts, Visx, D3)

  • tRPC experience

  • Familiarity with GPU or ML tooling dashboards

  • Comfort debugging perf issues (Lighthouse, Chrome DevTools)

  • Dev-ops chops: CI/CD, Docker, Terraform

You don’t need to tick every “nice-to-have” box-curiosity and the ability to learn quickly matter more.

Compensation

  • Base salary: $120,000 - $180,000

  • Equity: significant early-stage grant

  • Benefits: full medical/dental/vision, 401(k) with match, generous PTO, commuter + hardware stipends, daily office lunch

How we work

We iterate fast, test in prod (safely!), and celebrate small wins. You’ll demo work twice a week, pair with systems engineers, and ship to users continuously. Most of us are in the office 3-4 days a week; remote candidates considered if time-zone compatible with Pacific hours.

Equal Opportunity

Inference.net is an equal opportunity employer. We value diversity and do not discriminate on the basis of race, color, religion, gender identity, sexual orientation, national origin, veteran status, disability, age, or any other protected status.

Ready to build the front door to planet-scale AI?

Send a short note and a link to something you’ve shipped (code, demo, or Dribbble shots) to [email protected]. We can’t wait to chat!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,291 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$116k – $210k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Jacksonville
Go
SQL
Databases
Apache Kafka
PostgreSQL
DevOps
CI/CD
gRPC
Kubernetes
Apply
UX Research Lead 1 hour ago
$192k – $288k per year • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • New York
AI/ML
Hallucination
Design
Axure RP
Figma
Sketch
InVision
Apply
$51k – $136k per year (Estimated) • Remote/Hybrid • Full-Time • Belo Horizonte
C#
SQL
C#
.NET
Apply
$197k – $314k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco • Washington
Go
Apex
Apex
MuleSoft
AI/ML
AI Agents
Model Context Protocol
Agentforce
Frontend
GraphQL
DevOps
Kong
Management
ServiceNow
Marketing
Salesforce
QA
Postman
Apply
$140k – $302k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Belfast
JavaScript
TypeScript
AI/ML
A2A
AI Agents
Claude
Claude Code
Model Context Protocol
OpenAI Codex
Frontend
React.js
DevOps
AWS
Design
Figma
Apply
$220k – $320k per year • In office • Full-Time • 2+ years exp • San Francisco
C++
Python
C#
C++
PyTorch C++
C#
.NET
AI/ML
CUDA Toolkit
Embeddings
Knowledge Distillation
LLM
LoRA
PyTorch
Quantization
SGLang
TensorRT
TensorRT-LLM
vLLM
PEFT
CUDA
GPT-5
DevOps
Docker
Kubernetes
GitHub
Apply
$250k – $350k per year • In office • Full-Time • 3+ years exp • San Francisco
C#
C#
.NET
AI/ML
DeepSpeed
Knowledge Distillation
LLM
Multimodal AI
PyTorch
Reinforcement Learning
RLHF
Transformers
TRL
DPO
GPT-5
Hugging Face
Megatron-LM
Post-training
SFT
DevOps
GitHub
Apply
$220k – $320k per year • In office • Full-Time • 2+ years exp • San Francisco
C#
C#
.NET
AI/ML
Axolotl
DeepSpeed
Knowledge Distillation
LLM
Multimodal AI
PyTorch
Transformers
GPT-5
Hugging Face
Post-training
SFT
DevOps
GitHub
Analytics
ETL/ELT
Apply
$83k – $188k per year (Estimated) • In office • 2+ years exp • San Francisco
Python
AI/ML
AI Agents
LLM Guardrails
Model Context Protocol
DevOps
Terraform
Cybersecurity
Crowdstrike
GDPR
Least Privilege
Okta
SentinelOne
Management
Google Workspace
Slack
Apply
$171k – $273k per year • In office • Full-Time • 8+ years exp • PhD • San Francisco • Washington
AI/ML
A2A
Agentforce
AI Agents
Model Context Protocol
DevOps
AWS
GCP
Marketing
Salesforce
Apply
$173k – $260k per year • In office • Full-Time • PhD • San Francisco
JavaScript
Node JS
Python
Python
Celery
Django
Flask
Databases
RabbitMQ
Redis
AI/ML
Agentforce
AI Agents
DevOps
Akamai
AWS
CI/CD
Cloudflare
CloudFormation
Helm
Jenkins
Kubernetes
Spinnaker
Terraform
Marketing
Salesforce
Apply
$182k – $305k per year • Remote/Hybrid • Full-Time • 10+ years exp • PhD • San Francisco
AI/ML
Agentforce
AI Agents
Marketing
Salesforce
Apply
$173k – $260k per year • In office • Full-Time • 7+ years exp • PhD • New York • Indianapolis • San Francisco
AI/ML
AI Agents
Agentforce
Marketing
Salesforce
Apply
See all jobs
This is one of many
368,291 more open roles from verified company boards, updated every day.