{"id":1310782,"url":"https://alion.io/job/imu-biosciences-senior-software-engineer","title":"Senior Software Engineer","company":{"id":2382805,"name":"IMU Biosciences","domain":"imubiosciences.com","url":"https://alion.io/company/imubiosciences","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Ashby","truth_index":null},"role":"Backend","role_family":"Backend","seniority":"senior","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["London, United Kingdom"],"countries":["GB"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":92000,"max_usd":183000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":87},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AWS","optional":false},{"name":"CI/CD","optional":false},{"name":"Knowledge Graph","optional":false},{"name":"Machine Learning","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Apache Iceberg","optional":true},{"name":"Delta Lake","optional":true},{"name":"DuckDB","optional":true},{"name":"GitHub Actions","optional":true},{"name":"HIPAA","optional":true},{"name":"ISO 27001","optional":true},{"name":"Pulumi","optional":true},{"name":"Python","optional":true},{"name":"Spark","optional":true}],"status":"live","first_seen_at":"2026-05-27T15:50:43Z","employer_posted_date":"2026-05-27","last_verified_at":"2026-09-27T22:15:59Z","board_verified":true,"closed_at":null,"days_open":123,"trust":{"level":"stale","repost_count":0,"flags":["stale"],"days_open":122},"description":"About IMU Biosciences\nIMU Biosciences has developed proprietary platform technologies that generate and translate vast system-level immune data into actionable insights and tools to drive the development of precision medicines across a variety of diseases. Built on over a decade of research at King’s College London and the Francis Crick Institute, IMU leverages advanced immune profiling with proprietary AI and machine learning analytics to uncover novel clinical immune signatures. IMU continues to establish partnerships with leading pharma and biotech companies to advance disease diagnosis, optimise product selection, and improve patient stratification and monitoring - while also building its own pipeline of innovative products.\nAbout the role\nIMU is applying cutting-edge immune system science, data engineering, and machine learning to understand human health in a deeper and more actionable way. As our work scales, we are building the platform capabilities needed to make complex biological and computational data structured, traceable, reproducible, and reusable across the organisation.\nWe are looking for a Senior Software Engineer, Semantics to help build the semantic and metadata capabilities of the core Data Platform.\nThis is not a standalone ontology or knowledge graph initiative. The role sits directly within the Data Platform team and focuses on building practical platform systems that help scientists, computational immunologists, and engineers work with trusted and well-structured scientific data.\nYou will work closely with the Computational Immunology team to understand how analytical and machine learning workflows produce data, and help ensure those outputs are consistently structured, versioned, lineage-aware, discoverable, and reusable inside the platform.\nA core part of the role is helping turn fragmented scientific and computational outputs into usable data products that can be reliably found, assembled, interpreted, and reused by laboratory, project management, Computational Immunology, and Data Platform teams.\nThe work includes metadata systems, lineage and provenance capture, dataset contracts, FAIR data practices, data catalog capabilities, and semantic integration between pipelines, datasets, and scientific outputs.\nThis is a hands-on engineering role for someone who enjoys working across platform engineering, scientific workflows, metadata systems, and cloud-native infrastructure, while staying grounded in practical delivery and operational ownership.\nThe role is hybrid, with an expectation of working from our London office a couple of days per week.\nTeam and ways of working\nYou will join a small, growing Data Platform function working closely with Computational Immunology, wet lab, and clinical-facing teams.\nThis role sits within the platform layer: helping build the shared foundations that allow teams to ingest, structure, govern, discover, process, assemble, and reuse scientific data reliably.\nYou will not be working in isolation. The role is deeply connected to the wider platform effort and will involve close collaboration with engineers responsible for ingestion, orchestration, infrastructure, security, transformation, and platform operations.\nAs a senior engineer, you will be expected to:\nOwn substantial technical problems from discovery through implementation and operation.\n\nWork directly with stakeholders to gather requirements and translate them into practical platform capabilities.\n\nInfluence platform architecture and engineering standards alongside the wider Data Platform team.\n\nMake pragmatic technical trade-offs in environments with evolving scientific and operational requirements.\n\nContribute to shaping how the platform evolves as usage, scale, and regulatory expectations grow.\n\nYou should be comfortable balancing hands-on implementation work with technical leadership, collaboration, and operational ownership.\nWhat you will do\nBuild and maintain metadata, semantic, lineage, and provenance capabilities within the core Data Platform.\n\nWork closely with the Computational Immunology team to understand analytical workflows and translate them into dataset contracts, metadata standards, and reusable platform capabilities.\n\nDevelop systems for structuring, validating, registering, versioning, and governing scientific datasets and computational outputs.\n\nBuild ingestion pathways that return outputs from analytical and machine learning pipelines to the Data Platform as governed and reusable scientific datasets.\n\nHelp establish practical patterns for discovering, assembling, and delivering trusted datasets to laboratory, Computational Immunology, and operational teams.\n\nImprove data discoverability, usability, lineage, provenance, auditability, reproducibility, and reuse through well-structured, named, versioned, archived, and analysis-ready data products aligned with practical FAIR data principles.\n\nContribute to data catalog and scientific knowledge graph capabilities that connect datasets, workflows, biological entities, analytical outputs, and scientific conclusions.\n\nBuild AWS-native services, APIs, automation, and platform tooling using modern engineering practices.\n\nWork closely with scientists, computational immunologists, software engineers, and platform users to turn real scientific and data problems into reliable platform capabilities.\n\nImprove observability, reliability, documentation, maintainability, and operational maturity across semantic and metadata services.\n\nWhat we are looking for\nCore experience\nWe do not expect every candidate to have used every technology in our stack. We are mainly looking for strong engineering judgement, practical delivery experience, and evidence of building reliable systems in complex data environments.\nStrong software engineering experience in Python.\n\nPractical experience building and operating cloud-native systems on AWS.\n\nExperience working with data platforms, metadata systems, or data-intensive distributed systems.\n\nExperience with APIs, distributed services, infrastructure-as-code, CI/CD, and production engineering practices.\n\nExperience working with data lineage, metadata, schema management, versioning, validation, or data governance concepts.\n\nExperience supporting or integrating analytical, machine learning, or scientific workflows into production systems.\n\nTechnologies we use or value highly\nThe closer your experience is to this stack, the faster you are likely to be productive, but we care more about engineering depth and judgement than keyword matching.\nAWS ecosystem\n\nPulumi or equivalent infrastructure-as-code\n\nGitHub Actions or equivalent CI/CD\n\nNextflow or equivalent scientific workflow orchestration systems\n\nData lakes, metadata systems, lineage systems, schema management, and governed dataset patterns\n\nFAIR data principles, provenance, auditability, and reproducibility\n\nKnowledge graphs, semantic modelling, ontologies, RDF / OWL, or related technologies\n\nData catalogs and metadata management systems\n\nContainers and modern DevOps practices\n\nNice to have\nExperience in scientific, bioinformatics, computational biology, immunology, clinical, healthcare, or regulated data environments.\n\nExperience with large-scale analytical or biomedical datasets.\n\nExperience implementing data catalogs, lineage systems, or knowledge graph capabilities.\n\nExperience with modern data lakehouse architectures, including open table formats such as Delta Lake or Apache Iceberg; query engines such as AWS Athena, or DuckDB; processing frameworks such as Apache Spark; and orchestration/compute platforms such as AWS Batch.\n\nExperience with regulated software, security, quality, or healthcare frameworks such as ISO 27001, IEC 62304, HIPAA, or similar.\n\nExperience building self-service platform capabilities for scientific or data-intensive teams.\n\nWhat success looks like\nYou will help make the Data Platform a dependable foundation for scientific discovery and computational research.\nSuccess means computational outputs are no longer treated as disconnected pipeline artifacts, but as trusted and reusable scientific assets with clear structure, metadata, provenance, lineage, and lifecycle management.\nLaboratory and Computational Immunology teams should be able to reliably find, assemble, interpret, and reuse trusted datasets without needing bespoke manual data wrangling for each project.\nThe platform should increasingly support discoverable, versioned, semantically connected scientific datasets that can be reused across workflows, studies, and future AI systems.\nWhy join us\nIMU is building a company around a big scientific idea: that a deeper, data-driven understanding of the immune system can change how we understand, monitor, and ultimately improve human health.\nWe work with rich, complex biological and computational data and combine scientific expertise, machine learning, and modern platform engineering to generate insight from the immune system. As our work moves closer to the clinic, the systems we build now need to support both rapid scientific discovery and the discipline required for reproducibility, governance, and future regulated use.\nThis is a rare opportunity to help shape the semantic and metadata foundations of a modern scientific data platform while the architecture and operating model are still being defined.\nYou will not be maintaining a legacy metadata system or building isolated ontology models disconnected from real workflows. You will help build practical platform capabilities that directly support computational science, reproducibility, AI-readiness, and scientific reuse at scale.\nThe team is collaborative, scientifically curious, and pragmatic. We value strong engineering, sensible architecture, operational ownership, and people who can work effectively across disciplines while staying focused on practical delivery and real scientific outcomes.\nThe Data Platform is central to IMU’s future, which means this role offers genuine scope to influence technical direction, shape engineering standards, and help define how scientific data is structured, governed, discovered, and reused across the organisation as the platform scales.\nYou will join while foundational decisions are still being made, with the opportunity to build systems and patterns that can grow with the company from research-scale workflows toward future clinical and regulated environments.\nWe support hybrid working, with time in our London office a couple of days per week, and flexibility around how people do their best work.","description_format":"text","description_chars":10555,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Hybrid work"],"hiring_locations":[{"name":"United Kingdom","iso":"GB","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Biotechnology","Robotics","Bioinformatics","Robotic Components & Sensors"],"lifecycle":[{"event":"open","at":"2026-09-26T16:06:31Z"}],"liveness":{"score":7,"band":"cold","label":"Long shot","p_open":1,"p_active":0.263,"p_room":0.28,"age_days":122,"expected_fill_days":24,"reasons":["conf:2","win:tail","crowd:"],"computed_at":"2026-09-27T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/imu-biosciences-senior-software-engineer","json_url":"https://alion.io/job/imu-biosciences-senior-software-engineer.json","meta":{"generated_at":"2026-09-28T01:12:06Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":733,"day_limit":5000,"remaining_today":4267,"minute_limit":60,"resets_at":"2026-09-29T00:00:00Z"}}}