368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$136k – $167k per year
Location
In office (San Francisco)
Employment
Full-Time
Overview
Company
Impact
Profile match
Benchling is a biotechnology software company headquartered in San Francisco, California, and founded in 2012. The company provides a cloud-based platform for research and development that includes tools for electronic lab notebooks, biological registration, inventory management, and workflow automation. Serving over 200,000 scientists across academic institutions, startups, and major biopharmaceutical firms, the platform integrates artificial intelligence to accelerate scientific discovery and data management globally.

We are rebuilding biotech for the AI era.

When a breakthrough is delayed, the world waits. Getting a molecule from discovery to patients, or a crop from lab to field, involves thousands of slow, manual, disconnected steps. AI has the potential to change this, compressing decades of R&D work into years. But that only happens when clean, structured scientific data and AI are built into how science gets done.

Benchling is the AI platform for biotech R&D. Scientists use Benchling to design experiments, capture structured data, and run AI agents and models directly in their workflows. Over 200,000 scientists around the world trust Benchling to power their most important work, from academic labs to Sanofi, Moderna, and more than half of the world's top 50 biopharma.

We’re building an AI scientist for our customers. We can’t do that if we haven’t built the muscle ourselves. AI fluency is the foundation we build on; it's core to how we work, and we're committed to helping every new hire integrate it into their day-to-day. As part of our interview process, you'll complete a brief AI-focused exercise or discussion so we can understand how you think about and use AI to drive impact in your role. Feel free to reference any tools, platforms, or workflows you use today.

Role Overview

We’re a team focused on making frontier AI models better at science. LLMs know an extraordinary amount of biology, but there’s still a large gap in reasoning for the real-world problems scientists face every day. We recently published some of our work here.

You’ll build the datasets, evaluations, and systems that help close that gap. You’ll work with scientists to turn complex scientific work into rigorous tasks that models can learn from and be evaluated against. You’ll partner with leading AI labs to understand where models fail and how to improve them.

This is an early and rapidly evolving area. You’ll work at the intersection of software engineering, biology, and frontier AI: finding tasks that are challenging for LLMs and valuable to scientists, designing evaluations that capture real scientific judgment, and building systems to create these tasks at scale.

RESPONSIBILITIES

  • Build datasets for evaluating and improving frontier models, turning complex scientific data into high-quality tasks and environments for LLMs.

  • Analyze model failure modes, running experiments across frontier models to understand where they struggle and identify opportunities for improvement.

  • Build scalable data infrastructure, creating pipelines that curate, transform, and validate large volumes of scientific data into tasks for model evaluation and improvement.

  • Collaborate with frontier AI labs, helping develop and evaluate new approaches for improving models on challenging scientific tasks.

  • Work closely with scientists, translating expert judgment into problems and evaluation criteria that can reliably distinguish strong model behavior.

QUALIFICATIONS

  • 2+ years at the intersection of biology and AI, with experience evaluating and improving scientific models or LLMs for biological applications.

  • Experience building with LLMs, with an intuition for where current models excel, where they struggle, and how to design systems around their capabilities.

  • Curiosity and excitement about frontier AI, with a desire to understand and push the capabilities of rapidly improving models.

  • Comfort working on ambiguous problems, where the playing field is rapidly shifting and the right technical approaches are still being discovered.

  • Collaborative mindset, able to work closely with engineers, scientists, and external research partners.

  • Desire to work in a fast-paced environment, where priorities can shift and rapid experimentation is encouraged.

HOW WE WORK

This is an in-person team in San Francisco built around collaborating in the office in a fast-paced environment. We’re in the office Monday through Friday.

#LI-KW1

Benchling welcomes everyone.

We believe diversity enriches our team so we hire people with a wide range of identities, backgrounds, and experiences.

We are an equal opportunity employer. That means we don’t discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. We also consider for employment qualified applicants with arrest and conviction records, consistent with applicable federal, state and local law, including but not limited to the San Francisco Fair Chance Ordinance.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$70k – $150k per year • In office • Full-Time • 5+ years exp • Calgary
Java
Node JS
Python
SQL
TypeScript
JavaScript
Java
Spring Boot
Databases
Apache Kafka
Chroma
OpenSearch
pgvector
Pinecone
PostgreSQL
AI/ML
AWS Bedrock
Copilot
Embeddings
LangChain
LangGraph
LLM
Prompt Engineering
RAG
AI Agents
Anthropic
Function Calling
Human-in-the-Loop
OpenAI
Structured Outputs
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
Docker
Kubernetes
Vector
Apply
Senior AI Architect 5 hours ago
$138k – $304k per year (Estimated) • In office • Full-Time • 6+ years exp • Bachelor's Degree • Singapore
Python
SQL
Databases
Databricks
AI/ML
AI Agents
LangGraph
OpenAI
RAG
Spark
LangChain
DevOps
Azure
Apply
$118k – $168k per year • In office • Full-Time • Ottawa • Halifax
Databases
Snowflake
AI/ML
AI Agents
AutoGen
CrewAI
Human-in-the-Loop
LangChain
LLM
DevOps
AWS
Azure
Apply
$73k – $183k per year (Estimated) • Remote/Hybrid • Full-Time • 6+ years exp • Zug
Python
Python
FastAPI
Databases
PostgreSQL
Redis
AI/ML
AI Agents
AWS Bedrock
AWS Bedrock AgentCore
LLM
DevOps
Amazon EKS
AWS
AWS CDK
CI/CD
Datadog
Kubernetes
OpenTelemetry
Platform Engineering
Apply
SR AI ENGINEER, SMAI 5 hours ago
In office • Full-Time • 2+ years exp • Bachelor's Degree • Taoyuan
JavaScript
Python
SQL
TypeScript
Python
FastAPI
AI/ML
AI Agents
Function Calling
LLMOps
Prompt Engineering
Quantization
RAG
Streamlit
Frontend
Angular
React.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
GitHub Actions
Kubernetes
OpenShift
Apply
$72k – $177k per year (Estimated) • Remote • Full-Time • 5+ years exp • Bachelor's Degree • London
Apex
Go
JavaScript
Python
Apex
Gearset
Visualforce
AI/ML
AI Agents
Agentforce
DevOps
CI/CD
Management
Jira
Marketing
Salesforce
Apply
Data Engineer 3 days ago
$83k – $207k per year • Remote • Full-Time • 3+ years exp
Python
SQL
Databases
Snowflake
AI/ML
AI Agents
dbt
LLM
DevOps
AWS
CI/CD
Analytics
Tableau
ETL/ELT
Marketing
Salesforce
Apply
Data Engineer 3 days ago
$153k – $207k per year • Remote/Hybrid • Full-Time • 3+ years exp • San Francisco
Python
SQL
Databases
Snowflake
AI/ML
AI Agents
dbt
LLM
DevOps
AWS
CI/CD
Analytics
Tableau
ETL/ELT
Marketing
Salesforce
Apply
$184k – $393k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Zurich
AI/ML
AI Agents
Apply
$152k – $340k per year (Estimated) • Remote • Full-Time • 2+ years exp • London
AI/ML
AI Agents
Apply
$293k – $385k per year • In office • Full-Time • San Francisco
AI/ML
OpenAI
Apply
$180k – $260k per year • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco
Python
AI/ML
ChatGPT
OpenAI
OpenAI Codex
Cybersecurity
FedRAMP
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.