368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$150k – $350k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
Prime Intellect is an artificial intelligence infrastructure company headquartered in San Francisco, California, and founded in 2023. The company provides a decentralized platform for training, evaluating, and deploying large-scale AI models, featuring tools for reinforcement learning, agent development, and a global compute marketplace. It operates globally by aggregating computing resources from various providers to enable researchers and developers to build open-source models and autonomous agents.

Own Your Intelligence

Prime Intellect is building the open superintelligence stack: the infrastructure frontier AI labs build internally, made available to every ambitious AI team.

Our platform, Lab, unifies compute, environments, evaluations, secure sandboxes, high-performance training, and deployment into one full-stack system for post-training at frontier scale - from SFT and RL to tool use, agent workflows, and continuously improving production models. We are building open frontier AI: open-source models trained end to end for long-horizon tasks like autonomous research, and the full-stack platform our own research team uses to build them. The next generation of AI companies, enterprises, and research teams do not just need more GPUs. They need the ability to turn their own workflows, tools, data, and feedback loops into superintelligence they own.

We train open frontier models and ship the same stack to our customers. Its spans the full stack of training, deploying and continuously improving models - compute, large-scale RL, environments, sandboxes, evals, and deployment.

Prime Intellect has raised $150M in total funding from Founders Fund, Radical Ventures, NVIDIA, and exceptional AI, infrastructure, and enterprise operators - including Andrej Karpathy, Dwarkesh Patel, and leaders and founders from Ramp, Perplexity, Harvey, Mercor, Zapier, Datadog, Semianalysis, Cognition, OpenAI, Thinking Machines, Together AI, SemiAnalysis, LangChain, Browserbase, Cloudflare, Sierra, Databricks, Airbnb, OpenRouter, Standard Intelligence, Fleet, Core Auto, and more. We are looking for people who want to build at the intersection of frontier research, real infrastructure, and go-to-market for a category that does not fully exist yet.

What You’ll Work On

  • Build and optimize the distributed training infrastructure behind our pre-training and large-scale RL training workloads by contributing to our prime-rl framework.

  • Improve end-to-end training efficiency across compute, memory, networking, and scheduling layers.

  • Design and implement low-level performance optimizations, including kernels, communication paths, and runtime improvements.

  • Work on distributed training systems spanning data, tensor, and pipeline parallel workloads.

  • Help shape the architecture of our RL training stack, including async rollout and post-training systems.

  • Contribute to open-source libraries and internal infrastructure used for frontier-scale model training.

  • Collaborate closely with researchers and infrastructure engineers to translate bottlenecks into concrete systems improvements.

  • Stay at the frontier of training systems, inference systems, compiler/runtime tooling, and hardware-aware optimization techniques.

You May Be a Fit If You Have

  • Strong systems engineering experience in AI/ML infrastructure, especially around large-scale model training or inference.

  • Deep familiarity with PyTorch and distributed training frameworks such as PyTorch Distributed, DeepSpeed, FSDP, Megatron, vLLM, Ray, or related tooling.

  • Experience optimizing training performance across kernels, memory movement, communication overhead, or parallelization strategy.

  • Hands-on experience with large-scale training techniques including data parallelism, tensor parallelism, and pipeline parallelism.

  • Strong understanding of GPU architecture, profiling, and performance debugging.

  • Ability to identify bottlenecks across the stack and drive improvements from first principles.

  • Comfort working in a fast-moving environment with ambiguous problems and high ownership.

Especially Exciting

  • Experience writing or optimizing CUDA / Triton kernels.

  • Experience with compiler or runtime optimization for ML systems.

  • Experience working on RL training infrastructure, rollout systems, or asynchronous training pipelines.

  • Experience with multi-node GPU clusters and high-performance networking.

  • Contributions to open-source ML systems or infrastructure projects.

  • Interest in publishing technical work or sharing insights through engineering blogs and technical writing.

Benefits & Perks

  • Cash Compensation Range of $150-350k, plus equity incentives, aligning your success with the growth and impact of Prime Intellect.

  • Flexible work arrangements, with the option to work remotely or in-person at our offices in San Francisco.

  • Visa sponsorship and relocation assistance for international candidates.

  • Quarterly team off-sites, hackathons, conferences and learning opportunities.

  • Opportunity to work with a talented, hard-working and mission-driven team, united by a shared passion for leveraging technology to accelerate science and AI.

If you’re excited about building the systems foundation for frontier-scale training and open superintelligence, we’d love to hear from you.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$34k – $83k per year (Estimated) • Remote/Hybrid • Full-Time • Bengaluru
Node JS
JavaScript
Databases
Apache Kafka
DevOps
ArgoCD
Azure
Azure AKS
Azure DevOps
CI/CD
Datadog
FinOps
GitHub Actions
Grafana
Istio
Kubernetes
Platform Engineering
Prometheus
SLI/SLO/SLA
Terraform
GitHub
IAM
Cybersecurity
GDPR
Microsoft Defender
Microsoft Defender for Cloud
Okta
PCI DSS
Management
ServiceNow
Apply
Data Engineer 3 1 day ago
$21k – $52k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Bengaluru
Java
Python
Scala
SQL
Java
Maven
Python
pySpark
Databases
Apache Kafka
Databricks
Snowflake
AI/ML
Hadoop
Spark
DevOps
Azure
Analytics
Tableau
ETL/ELT
Apply
$91k – $216k per year (Estimated) • In office • Full-Time • 10+ years exp • Amsterdam
Java
JavaScript
Java
Gradle
Frontend
npm
pnpm
DevOps
ArgoCD
AWS
Bazel
CI/CD
Datadog
GitHub Actions
GitOps
Helm
JFrog Artifactory
Kubernetes
SRE
Terraform
GitHub
IAM
Cybersecurity
GDPR
Apply
$101k – $242k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Amsterdam
Python
Databases
Databricks
AI/ML
Spark
Apply
AI Engineer 1 day ago
$25k – $103k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Gurgaon
Python
SQL
Databases
Databricks
Microsoft Fabric
AI/ML
AI Agents
Embeddings
Gemini
Hallucination
LangChain
LangGraph
LLM
Multimodal AI
Prompt Engineering
PyTorch
RAG
Semantic Search
Spark
TensorFlow
Hugging Face
LLM Guardrails
LLMOps
OpenAI
Semantic Search
DevOps
AWS
Azure
CI/CD
Apply
$180k – $350k per year • Equity • In office • Full-Time • 5+ years exp • San Francisco
Python
Rust
Databases
Databricks
AI/ML
LangChain
OpenRouter
Perplexity
Together AI
OpenAI
Post-training
Red Teaming
SFT
AI Agents
Function Calling
DevOps
Cloudflare
Datadog
eBPF
GCP
Kubernetes
Service Mesh
Cybersecurity
Falco
FedRAMP
ISO 27001
SOC 2
Threat Modeling
Zero Trust
Management
Zapier
Apply
$150k – $300k per year • Equity • In office • Full-Time • San Francisco
TypeScript
JavaScript
Databases
Databricks
AI/ML
AI Agents
DSPy
LangChain
LangGraph
LLM
OpenRouter
Perplexity
Ray
Reinforcement Learning
RLHF
SGLang
Synthetic Data
Together AI
vLLM
GRPO
LLM Evaluation
OpenAI
Post-training
SFT
Function Calling
Model Context Protocol
Frontend
Next.js
React.js
DevOps
Cloudflare
Datadog
Docker
Grafana
Kubernetes
Prometheus
Terraform
Management
Zapier
Apply
$184k – $372k per year (Estimated) • In office • Full-Time • San Francisco
Python
Rust
TypeScript
JavaScript
Python
FastAPI
Databases
Databricks
AI/ML
LangChain
OpenRouter
Perplexity
Together AI
OpenAI
Post-training
SFT
TPU
Function Calling
Frontend
Next.js
React.js
Tailwind CSS
DevOps
Ansible
Cloudflare
Datadog
GCP
Grafana
Kubernetes
Prometheus
Rest API
Terraform
WebSockets
Management
Zapier
Apply
$150k – $300k per year • Equity • In office • Full-Time • San Francisco
Go
Python
Rust
TypeScript
JavaScript
Python
FastAPI
Databases
Databricks
AI/ML
LangChain
OpenRouter
Perplexity
Together AI
OpenAI
Post-training
SFT
TPU
Function Calling
Frontend
Next.js
React.js
Tailwind CSS
DevOps
Ansible
Cloudflare
Datadog
GCP
Grafana
Kubernetes
Prometheus
Rest API
Terraform
WebSockets
Management
Zapier
Apply
$150k – $300k per year • In office • Full-Time • San Francisco
Python
TypeScript
JavaScript
Python
FastAPI
SQLAlchemy
Databases
Databricks
AI/ML
Fine-tuning
LangChain
LLM
LoRA
OpenRouter
Perplexity
QLoRA
RLHF
SGLang
TensorRT
TensorRT-LLM
Together AI
vLLM
PEFT
NCCL
NVLink
OpenAI
Post-training
SFT
Function Calling
Frontend
Next.js
React.js
shadcn/ui
Tailwind CSS
tRPC
Radix UI
DevOps
Ansible
Cloudflare
Datadog
GCP
GitOps
Google Cloud Run
Google GKE
Grafana
Helm
KEDA
kubectl
Kubernetes
Loki
OpenTelemetry
Prometheus
Rest API
Terraform
Management
Zapier
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 1 hour ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
$222k – $277k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Francisco
DevOps
CI/CD
Immutable Infrastructure
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.