{"id":1331696,"url":"https://alion.io/job/novogaia-ml-research-engineer-evaluation","title":"ML Research Engineer, Evaluation","company":{"id":3786674,"name":"Novogaia","domain":"novogaia.bio","url":"https://alion.io/company/novogaia","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Ashby","truth_index":null},"role":"AI/ML","role_family":"AI/ML","seniority":null,"employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["London, United Kingdom"],"countries":["GB"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":96000,"max_usd":262000,"period":"year","method":"role_country_seniority_unknown","sample_n":101},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Machine Learning","optional":false},{"name":"Python","optional":false}],"status":"live","first_seen_at":"2026-08-07T06:45:27Z","employer_posted_date":"2026-08-07","last_verified_at":"2026-09-29T22:08:49Z","board_verified":true,"closed_at":null,"days_open":53,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":53},"description":"Overview\nNovogaia is an applied AI drug discovery company. We build machine learning systems that decode the chemistry of natural organisms, starting with fungi, to find the next generation of medicines.\nWe are a small team of AI engineers, computational biologists and chemists building foundation models for molecular structure prediction from mass spectrometry data. We are seeking a machine learning research engineer who excels at designing rigorous tests for molecular AI systems, and who can turn \"does this model actually work\" into a concrete, defensible answer. You will own the design of our internal benchmarks and evaluation pipelines; build the adversarial checks that catch shortcut learning and leakage. You will work closely with our modeling team to translate evaluation results into research priorities.\nThe Role\nDevelop a deep understanding of Novogaia's models, data, and evaluation needs\n\nDesign benchmark tasks that reflect real discovery problems: de novo molecular structure generation, molecular and spectral retrieval, mass spectrum simulation, molecular formula prediction, and compound property/class prediction\n\nBuild the datasets and controls that make a benchmark trustworthy: hard negatives, leakage-safe splits, and null baselines that catch a model exploiting shortcuts instead of genuine signal\n\nCommunicate evaluation results as clear findings for the modeling team, and as documentation and data cards that others can trust and reproduce\n\nWork across the team to scope and lead evaluation work, including:\nDefining what \"good\" looks like for a given model or task, and choosing the right test for it\n\nAnalyzing model behavior and interpreting results for researchers and non-technical stakeholders alike\n\nWorking with engineers to turn one-off analyses into repeatable, reproducible evaluation pipelines\n\nTranslate lessons from evaluation work into research priorities, data requirements, and R&D direction\n\nIn your first year, you'll take the lead on how Novogaia measures model quality, from benchmark design through adversarial testing and reporting. The tests you build will decide how much weight anyone can put on our models' outputs.\nWhat We Require\nResearch or applied experience in machine learning, with direct experience building or rigorously evaluating ML benchmarks\n\nExceptional technical communication skills, including the ability to explain evaluation findings clearly to both researchers and non-technical stakeholders\n\nAbility to analyze model behavior and interpret computational results critically\n\nStrong proficiency in Python, and comfort with reproducible, containerized pipelines\n\nFamiliarity with cheminformatics representations (SMILES, InChIKey, molecular fingerprints), or willingness to pick these up quickly\n\nFamiliarity with computational mass spectrometry or eagerness to learn\n\nWhat We Value\nAbility to dive deep into a result until you know whether it's real or an artifact\n\nStrong scientific judgment and a willingness to question the benchmark's own assumptions, not just the model's\n\nMotivation to build infrastructure other people can confidently rely and build upon\n\nComfortable being the person who tells the team a result doesn't hold up\n\nCuriosity, low ego, and a willingness to learn the chemistry side quickly, even if it's outside your original training","description_format":"text","description_chars":3337,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-09-27T10:28:40Z"}],"liveness":{"score":14,"band":"cold","label":"Long shot","p_open":1,"p_active":0.395,"p_room":0.35,"age_days":53,"expected_fill_days":21,"reasons":["conf:7","win:tail"],"computed_at":"2026-09-30T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/novogaia-ml-research-engineer-evaluation","json_url":"https://alion.io/job/novogaia-ml-research-engineer-evaluation.json","meta":{"generated_at":"2026-09-30T06:30:02Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":4547,"day_limit":5000,"remaining_today":453,"minute_limit":60,"resets_at":"2026-10-01T00:00:00Z"}}}