368,746open jobs
9,444companies
47,506added this week
Browse all
Salary
$193k – $391k per year (Estimated)
Location
In office (San Francisco)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Radical Numerics is an artificial intelligence research lab dedicated to developing general biological intelligence and generative genomics models. Headquartered in San Francisco, California, the company builds multimodal platforms capable of reading, writing, and engineering biological sequences across DNA, RNA, and proteins. Its technology aims to accelerate biopharmaceutical research, enhance early disease diagnostics, and establish robust biodefense capabilities.

About Us

Radical Numerics is an AI research lab building general biological intelligence. Our mission is to master the code of life, and our purpose is to reduce human suffering.

Our team created Evo, and started the field of generative genomics. Our work was featured on the cover of Science, and presented by our CEO on the main stage of TED2025. Evo was used to create the first AI gene therapy tool CRISPR-Cas9, and the first AI whole genome from scratch. Evo 2, featured in Nature, is the largest fully open source AI project across any domain.

Radical Numerics is bringing the rigor of distributed systems, model architecture, and numerics research to the challenges of biology. We’ve redesigned the foundation model training stack to turn the world’s raw scientific data (e.g. biological sequences, experiments, and physical processes), into intelligible, generative models that can expand and accelerate what humanity can understand, design, and cure.

The same generative breakthroughs that enable life-saving cures also lowers the barrier to creating engineered threats and AI-generated bioweapons. We believe these forces are inseparable. Radical Numerics was founded to develop both the power to design and the responsibility to defend.

About the Role

As a Member of Technical Staff, Inference at Radical Numerics, you will build and optimize the systems that bring frontier biological AI models into production. Your work will focus on delivering state-of-the-art inference performance for large-scale genome and multimodal biological models across a wide range of real-world applications, including therapeutics, diagnostics, synthetic biology, and biodefense.

This is a highly technical role at the intersection of AI systems, distributed computing, and model deployment. You will work closely with research, infrastructure, and external partners to ensure our models can be efficiently deployed, scaled, and integrated into production environments. Success in this role requires deep expertise in large language model inference, kernel optimization, GPU systems, and performance engineering.

You should be excited by questions such as: How do we reduce inference latency for 100B MoE models? How do we maximize throughput across heterogeneous hardware environments? How do we optimize custom kernels for emerging hybrid model architectures? How do we deploy foundation models reliably across cloud, on-premise, and highly regulated environments? How do we enable our partners to transform biological research and development through production-grade AI systems?

What You'll Do

Drive end-to-end performance improvements. Identify and eliminate bottlenecks across the inference stack, from model execution and memory management to networking, scheduling, and hardware utilization.

Develop high-performance inference primitives. Build and optimize GPU kernels, numerical operators, and serving infrastructure to maximize throughput, latency, and efficiency on modern accelerator platforms.

Partner with external customers and collaborators. Work directly with pharmaceutical companies, biotech organizations, research institutions, and government partners to deploy models in production environments and solve challenging technical problems.

Build scalable deployment infrastructure. Create systems for serving, monitoring, benchmarking, and operating foundation models reliably across cloud, enterprise, and secure environments.

Collaborate with research and platform teams. Ensure new model architectures can be efficiently deployed at scale and help translate frontier AI research into real-world impact.

What We're Looking For

Expertise in large-scale AI inference systems. Proven experience optimizing, deploying, and operating LLMs or other foundation models in production environments.

Strong performance engineering and kernel development skills. Deep understanding of GPU architectures and experience with CUDA, Triton, or equivalent technologies for building high-performance numerical software.

Systems-level thinking. Ability to diagnose and solve bottlenecks across the full stack, including model architectures, serving systems, networking, memory management, and distributed infrastructure.

Hands-on builder. Strong software engineering fundamentals with proficiency in Python and modern ML frameworks such as PyTorch.

Customer and deployment orientation. Experience working closely with users, customers, or cross-functional stakeholders to deliver production AI systems that solve real-world problems.

Excellent technical communication. Ability to collaborate effectively across research, engineering, infrastructure, and scientific teams.

Nice to Have

  • Experience with inference frameworks such as vLLM, TensorRT-LLM, SGLang, DeepSpeed, or similar systems.

  • Contributions to open-source AI infrastructure, inference frameworks, compilers, or kernel libraries.

  • Experience with distributed systems, cloud infrastructure, and large-scale GPU clusters.

  • Familiarity with biological foundation models, computational biology, genomics, or scientific AI applications.

  • Experience operating AI systems in regulated, secure, or mission-critical environments

Why Radical Numerics

Help build the infrastructure that powers the next generation of biological AI models and deploy them into some of the world's most important scientific and healthcare applications.

Work on some of the largest and most capable open biological AI models, helping transform breakthroughs in AI research into real-world impact across therapeutics, diagnostics, synthetic biology, and biodefense.

Join a team that brings together expertise in distributed systems, model architecture, numerics, AI safety, and biology.

Collaborate with leading researchers across AI labs, biotechs, pharmaceutical companies, hospital systems, government programs, and scientific institutions.

Radical Numerics is committed to equal employment opportunity and does not discriminate in any employment opportunities or practices based on an individual's race, color, creed, gender (including gender identity and gender expression), religion (all aspects of religious beliefs, observance or practice, including religious dress or grooming practices), marital status, registered domestic partner status, age, national origin or ancestry (including language use restrictions and possession of a driver’s license issued under California Vehicle Code section 12801.9), natural hair, physical or mental disability, political affiliation, medical condition (including cancer or a record or history of cancer, and genetic characteristics), sex (including pregnancy, childbirth, breastfeeding or related medical condition), genetic information, sexual orientation, military and veteran status or any other consideration made unlawful by federal, state, or local laws. It also prohibits unlawful discrimination based on the perception that anyone has any of those characteristics, or is associated with a person who has or is perceived as having any of those characteristics.

Radical Numerics participates in E-Verify and will provide the federal government with your Form I-9 information to confirm that you are authorized to work in the U.S.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$23k – $58k per year (Estimated) • In office • Full-Time • 7+ years exp • Gurgaon
Python
Databases
Apache Kafka
ElasticSearch
Redis
DevOps
Error Budget
Grafana
Incident Management
Kubernetes
OpenShift
Platform Engineering
Prometheus
Self-Healing
Cybersecurity
HashiCorp Vault
Cryptography
Vault
Apply
Remote • Full-Time
JavaScript
Python
TypeScript
AI/ML
AI Agents
Function Calling
LLM
Prompt Engineering
Structured Outputs
Management
n8n
Zapier
Apply
$36k – $78k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Pune
PowerShell
Python
SQL
DevOps
AppDynamics
AWS
Azure
GCP
Grafana
Prometheus
Splunk
Apply
$37k – $80k per year (Estimated) • Remote/Hybrid • Full-Time • 7+ years exp • Pune
PowerShell
Python
AI/ML
Anomaly Detection
DevOps
AppDynamics
AWS
Azure
CI/CD
GCP
Grafana
Incident Management
Prometheus
Splunk
Apply
$27k – $110k per year (Estimated) • Remote • Full-Time
JavaScript
Python
TypeScript
AI/ML
AI Agents
Function Calling
LLM
Prompt Engineering
Structured Outputs
Management
n8n
Zapier
Apply
$195k – $394k per year (Estimated) • In office • Full-Time • San Francisco
Python
AI/ML
Diffusion Models
Fine-tuning
JAX
PyTorch
Apply
$182k – $369k per year (Estimated) • In office • Full-Time • San Francisco
Python
AI/ML
LLM
TensorRT
TensorRT-LLM
Triton
vLLM
Apply
$200k – $406k per year (Estimated) • In office • Full-Time • San Francisco • Tokyo
Python
AI/ML
Multimodal AI
PyTorch
Apply
$191k – $388k per year (Estimated) • In office • Full-Time • San Francisco • Tokyo
C++
Python
Rust
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
Multimodal AI
PyTorch
Triton
Megatron-LM
NCCL
DevOps
Kubernetes
SLURM
Apply
$193k – $391k per year (Estimated) • In office • Full-Time • San Francisco • Tokyo
Python
AI/ML
Multimodal AI
PyTorch
Pre-training
Apply
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 3 hours ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.