368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$182k – $369k per year (Estimated)
Location
In office (San Francisco)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Radical Numerics is an artificial intelligence research lab dedicated to developing general biological intelligence and generative genomics models. Headquartered in San Francisco, California, the company builds multimodal platforms capable of reading, writing, and engineering biological sequences across DNA, RNA, and proteins. Its technology aims to accelerate biopharmaceutical research, enhance early disease diagnostics, and establish robust biodefense capabilities.

About Us

Radical Numerics is an AI research lab building general biological intelligence. Our mission is to master the code of life, and our purpose is to reduce human suffering.

Our team created Evo, and started the field of generative genomics. Our work was featured on the cover of Science, and presented by our CEO on the main stage of TED2025. Evo was used to create the first AI gene therapy tool CRISPR-Cas9, and the first AI whole genome from scratch. Evo 2, featured in Nature, is the largest fully open source AI project across any domain.

Radical Numerics is bringing the rigor of distributed systems, model architecture, and numerics research to the challenges of biology. We’ve redesigned the foundation model training stack to turn the world’s raw scientific data (e.g. biological sequences, experiments, and physical processes), into intelligible, generative models that can expand and accelerate what humanity can understand, design, and cure.

The same generative breakthroughs that enable life-saving cures also lowers the barrier to creating engineered threats and AI-generated bioweapons. We believe these forces are inseparable. Radical Numerics was founded to develop both the power to design and the responsibility to defend.

About the role

As a Member of Technical Staff, ML Engineer, this role owns the layer between the models and the people using them: the APIs, batch and compute systems, and services that make inference fast, reliable, and cheap at genome scale. You would build for our scientists and for the partners and developers who build on Omnii, our next-generation genome language model.

You care about throughput, latency, cost, and uptime, and you have real opinions about what a good developer experience feels like. The ML Engineer you build is how Omnii reaches the people using it to advance human health and biosecurity.

This role sits close to our product and data-partnership functions and our modeling team. You translate between what the science needs, what partners can consume, and what the infrastructure can deliver.

What you'll do

  • Build the inference APIs behind Omnii: real-time serving plus large batch scoring for genome- and variant-scale workloads, where a single job can be millions of sequences, latency-tolerant and cost-sensitive.

  • Design the compute and orchestration layer: job queues, autoscaling GPU inference, retries, fault tolerance, and the observability to run all of it against real SLAs.

  • Build the developer-facing product: the API surface, SDKs, and documentation that let an outside team depend on Omnii without hand-holding.

  • Drive down cost and latency per inference, and make performance predictable enough to price and guarantee.

  • Package the platform to run inside partner environments that cannot let data leave, including on-prem and air-gapped installs.

  • Work directly with our scientists and partners to shape the API around real inference workloads.

What we're looking for

  • You have built and operated backend or distributed systems at production scale, and you owned their reliability.

  • You have shipped a model as a service that other people depended on. You think in inference, serving, and throughput.

  • You have built asynchronous or batch job systems at scale. Genome-scale workloads should feel like familiar ground.

  • You write production code in Python and at least one typed language, and you are comfortable with containers and infrastructure as code.

  • You have product judgment. You can decide what to expose, what to hide, and how an API should feel to the person calling it.

  • You operate well with ambiguity and want the ownership of an early technical role.

Nice to have

  • GPU inference internals: vLLM, TensorRT-LLM, or Triton, with hands-on performance tuning.

  • Experience deploying software into locked-down environments such as enterprise VPC, on-prem, or air-gapped.

  • Familiarity with genomics or computational biology, or a real appetite to learn it fast.

  • Experience standing up an external API product from the first endpoint forward.

Radical Numerics is committed to equal employment opportunity and does not discriminate in any employment opportunities or practices based on an individual's race, color, creed, gender (including gender identity and gender expression), religion (all aspects of religious beliefs, observance or practice, including religious dress or grooming practices), marital status, registered domestic partner status, age, national origin or ancestry (including language use restrictions and possession of a driver’s license issued under California Vehicle Code section 12801.9), natural hair, physical or mental disability, political affiliation, medical condition (including cancer or a record or history of cancer, and genetic characteristics), sex (including pregnancy, childbirth, breastfeeding or related medical condition), genetic information, sexual orientation, military and veteran status or any other consideration made unlawful by federal, state, or local laws. It also prohibits unlawful discrimination based on the perception that anyone has any of those characteristics, or is associated with a person who has or is perceived as having any of those characteristics.

Radical Numerics participates in E-Verify and will provide the federal government with your Form I-9 information to confirm that you are authorized to work in the U.S.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$20k – $52k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • India
Python
SQL
C#
TypeScript
JavaScript
C#
.NET
AI/ML
AI Agents
Frontend
Angular
Apply
$24k – $63k per year (Estimated) • In office • Full-Time • 6+ years exp • Master's Degree • India
Crystal
Groovy
JavaScript
Perl
Python
Ruby
SQL
TypeScript
Java
Java
Apache Tomcat
Gradle
Hibernate
Maven
Spring Boot
Spring MVC
Databases
Apache Kafka
Db2
Oracle
PostgreSQL
RabbitMQ
AI/ML
Fine-tuning
Frontend
Angular
JQuery
DevOps
Apache HTTP Server
AWS
Azure
CI/CD
Docker
GCP
Jenkins
Kubernetes
Rest API
Cybersecurity
Checkmarx
SonarQube
Apply
$96k – $163k per year • Equity • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Boston
C#
JavaScript
Python
TypeScript
DevOps
Azure AKS
CI/CD
Kubernetes
Rest API
Azure
Apply
$10k – $26k per year (Estimated) • In office • 2+ years exp • Moscow
Python
SQL
Databases
PostgreSQL
DevOps
Graylog
Apply
$64k – $189k per year (Estimated) • Remote/Hybrid • Full-Time • 1+ year exp • Bachelor's Degree • Singapore
Python
SQL
Databases
Apache Kafka
AI/ML
Amazon SageMaker
Kubeflow
MLFlow
Spark
Vertex AI
DevOps
AWS
Azure
Azure DevOps
CI/CD
Docker
GCP
GitLab
GitLab CI
Jenkins
Kubernetes
Apply
$195k – $394k per year (Estimated) • In office • Full-Time • San Francisco
Python
AI/ML
Diffusion Models
Fine-tuning
JAX
PyTorch
Apply
$193k – $391k per year (Estimated) • In office • Full-Time • San Francisco
Python
AI/ML
CUDA
CUDA Toolkit
DeepSpeed
LLM
Multimodal AI
PyTorch
SGLang
TensorRT
TensorRT-LLM
Triton
vLLM
Mixture of Experts
Apply
$200k – $406k per year (Estimated) • In office • Full-Time • San Francisco • Tokyo
Python
AI/ML
Multimodal AI
PyTorch
Apply
$191k – $388k per year (Estimated) • In office • Full-Time • San Francisco • Tokyo
C++
Python
Rust
C++
PyTorch C++
AI/ML
CUDA
CUDA Toolkit
Multimodal AI
PyTorch
Triton
Megatron-LM
NCCL
DevOps
Kubernetes
SLURM
Apply
$193k – $391k per year (Estimated) • In office • Full-Time • San Francisco • Tokyo
Python
AI/ML
Multimodal AI
PyTorch
Pre-training
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
$160k – $283k per year • Equity • In office • 5+ years exp • San Francisco
AI/ML
AI Agents
Apply
$185k – $385k per year • Remote/Hybrid • Full-Time • 5+ years exp • San Francisco
JavaScript
Python
Databases
MySQL
PostgreSQL
AI/ML
OpenAI
Frontend
React.js
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.