402,911open jobs
14,044companies
78,108added this week
Browse all
Salary
$92k – $181k per year (Estimated)
Location
In office (London)
Employment
Full-Time
Overview
Company
Impact
Profile match
Relation Therapeutics is a London biotechnology company founded in 2019 that applies graph machine learning to disease biology. Its platform combines single-cell data with causal models to identify drug targets in bone and fibrotic disease. The company runs its own therapeutic programmes alongside pharmaceutical collaborations.

About Relation

Relation is a sector defining TechBio company developing transformational medicines, with technology at our core. Our ambition is to understand human biology in unprecedented ways, discovering therapies to treat some of life’s most devastating diseases. We leverage single-cell multi-omics from patient tissue, functional assays, and machine learning to drive disease understanding, from cause to cure.

We are scaling rapidly and building a team of exceptional individuals to push the boundaries of drug discovery. You will work in highly interdisciplinary teams where biology, computation, and engineering come together to solve complex problems that have not been solved before. Our state-of-the-art wet and dry labs in the heart of London are designed to accelerate this integration and translate insight into impact.

We are committed to building diverse and inclusive teams. Relation is an equal opportunities employer and does not discriminate on the basis of gender, sexual orientation, marital or civil partnership status, gender reassignment, race, colour, nationality, ethnic or national origin, religion or belief, disability, or age.

By joining Relation, you will help define how medicines are discovered and deliver meaningful impact for patients.

The opportunity

Relation is offering an outstanding opportunity for a Research Data Engineer to design and build the data systems that power the next generation of predictive models of cellular behaviour. We generate complex and high-dimensional datasets across modalities at scale. What we can model, and the pace at which we iterate, is determined by the quality of the data layer. This role focuses on making our datasets efficient for analysis and ML model training through deliberate systems design: storage layouts, access patterns, distributed systems to move data from instrument to models, modality-specific format decisions, query and serving layers tuned for GPU-saturated training, and upholding the data contract between the wet lab and ML teams.

Day to day, you will

  • Design, build, and maintain scalable data pipelines that ingest multi-modal scientific data.

  • Optimise data movement, storage layouts, and access patterns for analytical and ML workloads.

  • Stand up and evolve cloud-native data lake / lakehouse infrastructure.

  • Implement data versioning, lineage, and quality monitoring.

  • Partner with data scientists day-to-day to ensure fast and seamless data workflows.

  • Collaborate closely with ML scientists and research engineers to design data representations, storage layouts, and access patterns that enable efficient model training and experimentation.

  • Build and operate workflow orchestration for both production pipelines and large-scale batch jobs.

  • Ensure infrastructure meets security, audit, and governance requirements.

  • Champion engineering best practices across the data platform.

  • Contribute to architecture decisions across the broader ML platform, including how compute, data, and training systems integrate.

Professionally, you will have

  • A degree in Computer Science, Engineering, or a related quantitative discipline; significant industry experience in data engineering, MLOps, or data platform roles.

  • Excellent Python engineering skills.

  • Deep experience with cloud-native data infrastructure (AWS S3 / GCS, plus the surrounding ecosystem) and Infrastructure-as-Code (Terraform or equivalent).

  • A track record of designing data pipelines and storage layouts for large, heterogeneous datasets.

  • Experience building scalable analytical data processing workflows using modern engines and frameworks (e.g. Spark, Polars, Dask, DuckDB, or equivalent), with an understanding of their performance and architectural trade-offs.

  • Hands-on experience with workflow orchestration (e.g. Airflow, Dagster, Prefect, or equivalent) and containerised environments (e.g. docker, k8s).

  • Working knowledge of modern columnar / scientific data formats (Parquet, Zarr, TileDB, HDF5) and lakehouse technologies.

  • Experience partnering closely with scientific or research users and comfortable with the messiness of real-world experimental data.

  • Bonus experience: biomedical or genomics data (BAM, FASTQ, AnnData, OME-Zarr); regulated or pharma-partnered environments; data governance, FAIR principles, or research data management; feature store implementations.

Personally, you

  • Are comfortable working in a matrixed environment, balancing multiple stakeholders and contributing effectively across teams.

  • Take ownership of your work, proactively seek opportunities to contribute, and enable others to do their best work.

  • Communicate openly and directly, give and receive feedback constructively, and handle challenging conversations with respect.

  • Actively seek out diverse perspectives, build strong working relationships, and contribute to shared goals across teams.

  • Embrace challenges with openness and resilience, set high standards for yourself, and strive to deliver meaningful outcomes.

Working style & culture at Relation

At Relation, we operate in a matrixed, interdisciplinary environment, where impact is driven through collaboration across scientific, technical, and operational domains. We collaborate, and you will partner with colleagues across multiple teams and projects, contributing your expertise while aligning to shared company priorities. We work together and win together! The patient is waiting!

Recruitment agencies

Please note that Relation does not accept unsolicited resumes from agencies. Resumes should not be forwarded to our job aliases or employees. Relation will not be liable for any fees associated with unsolicited CVs.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
402,911 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
London
Staff Engineer 25 min ago
$40k – $86k per year (Estimated) • Remote • Bengaluru
C#
JavaScript
SQL
TypeScript
C#
ASP.NET Core
Databases
MySQL
PostgreSQL
AI/ML
ChatGPT
Claude
Copilot
Frontend
Angular
React.js
DevOps
Amazon CloudWatch
Amazon EC2
Amazon ECS
Amazon EKS
Amazon S3
AWS
AWS Lambda
Azure
CI/CD
CloudFormation
Datadog
Docker
GCP
Git
GitHub
GitLab
GitLab CI
Kubernetes
Octopus Deploy
Service Mesh
Shift-Left
Splunk
TeamCity
Terraform
Cybersecurity
Shift-Left Security
Marketing
LinkedIn
Apply
$112k – $140k per year • In office
Java
TypeScript
DevOps
AWS
AWS Lambda
Azure
CI/CD
GitLab
GitLab CI
Jenkins
Cybersecurity
FedRAMP
NIST 800-53
SBOM
SonarQube
STRIDE
Threat Modeling
Apply
DevOps Engineer 22 min ago
In office • Full-Time • Bachelor's Degree • Santiago
Python
DevOps
ArgoCD
AWS
Bitbucket
CI/CD
Docker
GCP
Grafana
Helm
Jenkins
Kubernetes
Spinnaker
Terraform
Apply
Fullstack Engineer 1 hour ago
$100k – $150k per year • Equity 0.2–2% • Remote • Full-Time • 1+ year exp • San Francisco
C++
JavaScript
Python
DevOps
AWS
Apply
$84k – $100k per year • Remote • 8+ years exp • Bachelor's Degree
Python
SQL
Databases
Oracle
DevOps
Ansible
AWS
Azure
CI/CD
Apply
$93k – $224k per year (Estimated) • In office • Full-Time • London
AI/ML
Kubeflow
DevOps
CI/CD
IAM
Kubernetes
Platform Engineering
Cybersecurity
GDPR
ISO 27001
SOC 2
Apply
$103k – $216k per year (Estimated) • In office • Full-Time • PhD • London
Python
AI/ML
PyTorch
TensorFlow
Apply
$94k – $194k per year (Estimated) • In office • Full-Time • PhD • London
Python
AI/ML
Agentic Workflows
AI Agents
Apply
$79k – $154k per year (Estimated) • In office • Full-Time • PhD • London
Python
AI/ML
Agentic Workflows
AI Agents
Apply
$97k – $200k per year (Estimated) • In office • Full-Time • Bachelor's Degree • London
Python
AI/ML
Agentic Workflows
AI Agents
CUDA
CUDA Toolkit
Multimodal AI
PyTorch
Triton
Apply
In office • London
Apply
In office • London
Apply
Content Strategist 1 day ago
$80k – $120k per year • In office • Full-Time • 3+ years exp • London
Marketing
LinkedIn
Apply
Marketing Lead 1 day ago
$80k – $130k per year • In office • Full-Time • 3+ years exp • London
Apply
Operations 1 day ago
$35k – $50k per year • In office • Full-Time • London
Apply
See all jobs
This is one of many
402,911 more open roles from verified company boards, updated every day.