368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$138k – $248k per year (Estimated)
Location
Remote/Hybrid (Emeryville, United States)
Seniority
Senior · 5+ years exp
Overview
Company
Impact
Profile match
Profluent is a company founded in 2022 that trains large language models on protein sequences to design new functional proteins. It released OpenCRISPR-1, a gene editor designed entirely by a model and made freely available for ethical research and commercial use. The company works on editors, antibodies and enzymes that have no direct natural counterpart.

Profluent is the frontier AI lab for biology. Profluent builds powerful foundation models for all of life's molecules, unlocking solutions that transform medicine, agriculture, and beyond. Founded in 2022 and headquartered in Emeryville, CA, Profluent is backed by leading investors including Altimeter Capital, Bezos Expeditions, Spark Capital, Insight Partners, Air Street Capital, AIX Ventures, and Convergent Ventures and has raised over $150M to date.

We’re looking for a Senior Software Engineer to help design, build, and scale Profluent’s data platform. This platform houses data from protein engineering campaigns, including protein designs, experimental results, partner datasets, analytical outputs, and model-ready training data. It enables rapid machine learning, biological discovery, and secure collaboration across internal and external programs.

This role is ideal for an engineer who enjoys building robust data systems: secure ingestion pipelines, well-structured warehouses, reliable data models, access controls, auditability, and infrastructure that makes complex scientific data usable at scale. You will work closely with ML, bioinformatics, and program teams to ensure Profluent’s data is organized, governed, accessible, and protected.

Responsibilities

  • Design, build, and maintain scalable data infrastructure for protein engineering campaigns, including ingestion, transformation, validation, storage, and retrieval of large scientific datasets
  • Develop secure data pipelines for internal and partner-generated data, with strong attention to access control, data siloing, provenance, auditability, and compliance with data use restrictions
  • Own core components of Profluent’s data warehouse and data platform, using Python, GCP, PostgreSQL, BigQuery, and related cloud-native technologies
  • Build systems that transform raw experimental, computational, and partner data into structured, reliable, analysis-ready and model-ready datasets
  • Establish best practices for data modeling, metadata management, data quality checks, schema evolution, versioning, and documentation
  • Collaborate with ML engineers, computational biologists, data scientists, and program stakeholders to understand data requirements and translate them into scalable technical systems
  • Improve engineering quality through thoughtful system design, code review, testing, CI/CD, observability, and maintainable development workflows
  • Contribute to architectural decisions for how Profluent stores, secures, organizes, and uses data across programs and partnerships

Qualifications

  • 5+ years of software engineering, data engineering, or data platform experience
  • Strong proficiency in Python and modern software development practices, including git, testing, code review, CI/CD, and production deployment
  • Experience designing and operating production data pipelines, data warehouses, and data models at scale
  • Hands-on experience with cloud platforms, preferably GCP, and technologies such as BigQuery, PostgreSQL, object storage, workflow orchestration, and containerized services
  • Strong understanding of data security, access control, data partitioning or siloing, audit logging, and managing sensitive or restricted datasets
  • Experience working with complex, heterogeneous datasets and building systems that make them reliable, discoverable, and usable
  • Ability to work independently, make sound technical decisions, and drive projects from ambiguous requirements to production systems
  • BS, MS, or PhD in Computer Science, Engineering, Data Science, Bioinformatics, or a related technical field, or equivalent practical experience

Preferences

  • Experience with scientific, biological, clinical, genomic, laboratory, or high-throughput experimental data
  • Experience managing external partner, customer, or restricted-access datasets
  • Familiarity with data governance, lineage, metadata systems, schema registries, or data catalogs
  • Experience with research data systems, LIMS, ELNs, Benchling, or adjacent scientific platforms
  • Background working with ML, data science, computational biology, or cross-disciplinary technical teams
  • Interest in learning biology, gene editing, protein design, or machine learning concepts

What We Offer

  • High-growth opportunity with meaningful impact on the future of protein design
  • Competitive compensation package with equity participation
  • 401(k) with a strong employer match
  • Comprehensive benefits including health/dental/vision insurance
  • Generous PTO policy and commitment to work-life balance
  • Professional development opportunities in a cutting-edge field at the intersection of AI and biology

Profluent Bio, Inc is an equal opportunity employer promoting diversity and inclusion in the workspace. We do not discriminate on the basis of race, color, religion, marital status, age, national origin, ancestry, physical or mental disability, medical conditions, veteran status, sexual orientation, gender (including gender identity and gender expression), sex (which includes pregnancy, childbirth, and breastfeeding), genetic information, taking or requesting statutorily protected leave, or any other basis protected by law.

Work Authorization Requirement

Applicants must have ongoing work authorization in the United States that does not require employer sponsorship. Sponsorship will not be provided now or at any time in the future for this position.

Employment Eligibility Verification

Legal authorization to work in the United States is required. In compliance with federal law, all persons hired must verify their identity and work eligibility and complete the required employment verification form upon hire.

Hiring Salary Range

$170,000—$220,000 USD

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Emeryville
$133k – $161k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Westminster
Bash
C++
Java
Python
DevOps
CI/CD
Git
Apply
$105k – $252k per year • Remote • Full-Time • 18+ years exp • Bachelor's Degree
Python
Java
Java
Gradle
DevOps
Ansible
AWS
CI/CD
CloudFormation
Configuration Management
Docker
GitHub Actions
GitLab CI
Helm
Jenkins
Kubernetes
Platform Engineering
Terraform
GitHub
GitLab
Cybersecurity
Sonatype Nexus IQ
Management
Confluence
Jira
Apply
$158k – $189k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Westminster
Python
DevOps
Bitbucket
Git
GitLab
Management
Confluence
SpaceTech
GMAT
Apply
$191k – $267k per year • Equity • Remote • Full-Time • 4+ years exp • Master's Degree
Python
SQL
AI/ML
Anomaly Detection
Apply
$217k – $304k per year • Equity • Remote • Full-Time • 8+ years exp
Go
Databases
Apache Kafka
ClickHouse
Google BigQuery
AI/ML
Flink
Recommender Systems
DevOps
Incident Management
Kubernetes
Apply
$139k – $269k per year (Estimated) • Remote/Hybrid • 5+ years exp • Bachelor's Degree • Emeryville
AI/ML
MLFlow
PyTorch
DevOps
CI/CD
Apply
$132k – $248k per year (Estimated) • Remote/Hybrid • 3+ years exp • Bachelor's Degree • Emeryville
Python
AI/ML
CUDA
CUDA Toolkit
Fine-tuning
PyTorch
Spark
Triton
FSDP
SFT
DevOps
AWS
Azure
Docker
GCP
Kubernetes
Analytics
ETL/ELT
Apply
$156k – $299k per year (Estimated) • Remote/Hybrid • 7+ years exp • Bachelor's Degree • Emeryville
Python
AI/ML
Spark
DevOps
CI/CD
Apply
$134k – $293k per year (Estimated) • Remote/Hybrid • PhD • Emeryville
AI/ML
PyTorch
Reinforcement Learning
Post-training
DevOps
AWS
Azure
GCP
Apply
$131k – $288k per year (Estimated) • Remote/Hybrid • PhD • Emeryville
AI/ML
PyTorch
Pre-training
DevOps
AWS
Azure
GCP
Apply
$139k – $269k per year (Estimated) • Remote/Hybrid • 5+ years exp • Bachelor's Degree • Emeryville
AI/ML
MLFlow
PyTorch
DevOps
CI/CD
Apply
$132k – $248k per year (Estimated) • Remote/Hybrid • 3+ years exp • Bachelor's Degree • Emeryville
Python
AI/ML
CUDA
CUDA Toolkit
Fine-tuning
PyTorch
Spark
Triton
FSDP
SFT
DevOps
AWS
Azure
Docker
GCP
Kubernetes
Analytics
ETL/ELT
Apply
$179k – $200k per year • Remote/Hybrid • 8+ years exp • Bachelor's Degree • Emeryville
DevOps
Platform Engineering
Cybersecurity
HIPAA
SOC 2
Apply
$156k – $299k per year (Estimated) • Remote/Hybrid • 7+ years exp • Bachelor's Degree • Emeryville
Python
AI/ML
Spark
DevOps
CI/CD
Apply
$134k – $293k per year (Estimated) • Remote/Hybrid • PhD • Emeryville
AI/ML
PyTorch
Reinforcement Learning
Post-training
DevOps
AWS
Azure
GCP
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.