{"id":1544007,"url":"https://alion.io/job/nationalarchives-data-engineer","title":"Data Engineer","company":{"id":17743,"name":"Nationalarchives","domain":"nationalarchives.gov.uk","url":"https://alion.io/company/nationalarchives","size_band":"501-1000","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":"junior","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Kew, United Kingdom"],"countries":["GB"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":45000,"max":null,"currency":"GBP","period":"year","gross":null,"usd_annual":59786},"salary_estimate":null,"experience_years_min":null,"visa_sponsorship":true,"relocation_package":false,"has_equity":false,"technologies":[{"name":"ETL/ELT","optional":false},{"name":"Python","optional":false},{"name":"SQL","optional":false},{"name":"Agile","optional":true},{"name":"Talend","optional":true}],"status":"live","first_seen_at":"2026-09-30T21:49:07Z","employer_posted_date":"2026-09-30","last_verified_at":"2026-10-01T21:50:10Z","board_verified":true,"closed_at":null,"days_open":1,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":1},"description":"As the living, growing home of our national story, The National Archives is a special place to work. We’re an institution nearly 200 years old with a collection spanning 1,000 years of history.\nWe’ve set ourselves the challenge of becoming the 21st Century national archive - a different kind of cultural and heritage institution: Inclusive, Entrepreneurial, Disruptive. We won’t become this overnight. It will take time, focus, effort and daring.\nThat’s where you come in. Because we can’t do this without you.\nJob Overview\nSalary: £45,000 per annum (£43,367 salary + £1633 market supplement)Contract type: Permanent\nBand: E / Higher Executive Officer\nClosing date: Wednesday 14th October 2026 at midnight\nShape the Future of the Nation’s Memory\nAt The National Archives, we are more than custodians of the past - we are pioneers in digital archiving, ensuring the UK’s public records are robust, connected, and accessible for generations to come. As a Data Engineer in our small, dynamic Digital Archiving team, you’ll help transform how archival data is managed, structured, and enriched, unlocking its value for public use and discovery\nWhy Join Us?\nYour work will directly support the preservation and accessibility of the UK’s documentary heritage, including a new Parliamentary Archives project.\nWe nurture talent, value every voice, and support personal and professional development. You’ll join a collaborative, forward-looking team where your ideas are heard and your growth is championed.\nWork with a modern tech stack including Python, SQL, Shell scripting, Talend, XSLT, XML, CSV, JSON, RDF, RDBMS, NoSQL, and graph databases. You’ll help design and build scalable data pipelines, automate data processing, and develop tools for data enrichment and integration\nWe invest in your learning, offering opportunities to develop your skills in emerging technologies and data practices that enhance public access to archives.\nThis Job Would Suit…\nIf you’ve been working in an entry-level data role and are ready to take the next step, this is an ideal opportunity to grow your skills and responsibilities. We welcome those who are passionate about data, eager to learn, and keen to make a real impact. Data skills are highly valued at The National Archives, and you’ll be part of an active internal data community where specialists from across the organisation share knowledge, collaborate, and solve problems together.\nWhat You’ll Do\nDesign and develop scalable, repeatable data pipelines (ETL/ELT) to process, transform, and load archival data.\nAutomate data cleansing, validation, standardisation, and enrichment.\nDevelop and document schemas, ontologies, and models for archival data.\nIntegrate data from diverse sources and formats, supporting better discovery and reuse.\nCollaborate with product teams, archivists, and developers to deliver practical, data-centric solutions.\nContribute to a culture of openness, quality, and continuous improvement.\nWho We’re Looking For\nStrong experience in large-scale data analysis and manipulation using relevant programming languages and tools (e.g. Python, Shell scripting, SQL, XSLT).\nSolid understanding of data formats (XML, CSV, JSON, RDF) and database technologies (RDBMS, NoSQL, graph databases).\nExcellent analytical, problem-solving, and communication skills.\nAbility to work collaboratively in multidisciplinary teams and explain technical issues to non-technical colleagues.\nAwareness of the value and meaning of archival data and the ethical responsibilities of working with public records.\nThis is a full time post. However, requests for part-time working, flexible working and job share will be considered, taking into account at all times the operational needs of the Department.\nA combination of onsite and home working is available and applicants should be able to regularly travel to our Kew site for a minimum of 60% of their work time.\nReady to help shape the future of digital archiving?\nApply now and join a team where your work matters - for today, and for generations to come.\nApplication Process\nInterview: Interviews will be held on-site, ideally week commencing 2nd November 2026. The interview will include a technical test.\nPersonal Statement: We ask all applicants to upload their CV and to submit a personal statement of 600-700 words total.\nSelection for interview will be based on the 3 ‘essential’ criteria listed below so please ensure that your application demonstrates in detail how you meet these requirements.\nData manipulation Strong experience of accurate, reliable, large-scale analysis and manipulation of complex datasets using relevant programming languages and tools (e.g. Shell scripting, SQL or XSLT).\nDatabase technologies Practical knowledge of database technologies (e.g. RDBMS, noSQL, graph databases or linked data stores).\nAnalysis and problem solving Excellent analytical skills, with a structured and proactive approach to solving technical problems. Able to apply a range of techniques to capture, document and communicate requirements, issues, designs and solutions.\nApplicants invited to interview will be assessed against ALL of the essential criteria (see job description below).\nSC clearance/willingness to obtain SC clearance will be required for this role. This requires candidates to have been resident in the UK for at least the past five years. Please do not apply if you have been resident in the UK for less than five years as your application will be rejected.\nArtificial Intelligence can be a useful tool to support your application, however, all examples and statements provided must be truthful, factually accurate and taken directly from your own experience. Where plagiarism has been identified (presenting the ideas and experiences of others, or generated by artificial intelligence, as your own) applications may be withdrawn and internal candidates may be subject to disciplinary action. Please visit the Civil Service Careers website where you can find further information on the use of AI in the application guidance section.\nSponsorship:\nWe are unable to offer sponsorship for this role.\nJob Description\nDepartment Digital Archiving\nReports to Lead Data Engineer\nBand Band E - (Civil Service Equivalent - HEO)\nLocation Kew (Hybrid working available)\nLine manages Junior staff, suppliers and contractors as required\nJob purpose\nArchives are the homes of our collective memory, past and future. The National Archives is the archive of UK government, and one of the world’s leading digital archives. As the scale and complexity of our digital collections continue to grow, we ensure our data is robust, connected, and accessible.\nWe are looking for a Data Engineer who thrives on solving complex problems, working at scale, and creating the pipelines and models that power access to the nation’s records. You will play a key role in transforming how archival data is managed, structured, and enriched, helping to unlock its value for public use and discovery.\nThis role is ideal for someone who enjoys designing and building efficient data workflows, developing tools to automate data processing, and collaborating with product teams and specialists to understand data needs and deliver practical solutions. You will bring strong programming skills and a good understanding of data modelling, transformation, and integration across a variety of formats and technologies.\nYou will work in a supportive, forward-looking team that values transparency, curiosity, and collaboration. You’ll have opportunities to develop your skills in existing and emerging technologies that enhance public access to archives and promote the re-use of data in meaningful ways.\nRole and responsibilities\nThis role falls under the GDD (formerly DDaT) Data Engineer and Data Analyst roles with elements of the Data Scientist role.\nKey Responsibilities\nData processing and pipelines\nDesign and develop scalable, repeatable data pipelines (ETL/ELT) to process, transform and load archival data.\nCreate scripts and tools for data cleansing, validation, standardisation, enrichment, and transformation.\nAutomate routine data tasks to improve efficiency, accuracy, and consistency.\nIdentify and manage connections between datasets, including those external to TNA.\nResearch and implement tools for entity extraction and semantic tagging to improve the usability and findability of archival data.\nData modelling and integration\nDevelop and document schemas, ontologies, and other models that represent and structure archival data.\nConnect and integrate data from different internal and external sources, including structured and semi-structured formats (such as XML, CSV, JSON or RDF).\nIdentify patterns and relationships within data to support better discovery and reuse.\nTechnical advice and collaboration\nDevelop excellent relationships with product teams, advocate for data-centric approaches to software development and help solve data problems.\nMaintain effective communication with external technical partners (e.g. suppliers), representing The National Archives to resolve complex data issues, clarify requirements, and ensure smooth delivery of digital archiving workflows\nWork closely with product teams, archivists, and developers to support data-centric solutions.\nProvide advice and input on data engineering approaches within the team.\nShare knowledge on data standards, validation rules, and transformation practices.\nQuality, openness and growth\nWork in the open and share code and documentation internally (and externally where appropriate).\nParticipate in code reviews, technical discussions, and knowledge sharing.\nContribute to building a data culture focused on reuse, quality, and continuous improvement.\nWorking conditions\nNormal office environment\nDisplay Screen Equipment user\nHybrid working pattern, 60% of time on-site, possibly more when new to the role\nPerson specification\nEssential skills and experience:\n Data manipulation Strong experience of accurate, reliable, large-scale analysis and manipulation of complex datasets using relevant programming languages and tools (e.g. Shell scripting, SQL or XSLT).\n Database technologies Practical knowledge of database technologies (e.g. RDBMS, noSQL, graph databases or linked data stores).\nAnalysis and problem solving Excellent analytical skills, with a structured and proactive approach to solving technical problems. Able to apply a range of techniques to capture, document and communicate requirements, issues, designs and solutions.\nCommunication and relationships Strong relationship building and communication skills, with an excellent user focus. Able to work collaboratively in multidisciplinary teams and explain technical issues clearly to non-technical colleagues.\nPrioritisation and organisation Ability to work with high accuracy, attention to detail and organisation, both independently and as a project team member. Able to prioritise competing tasks and deliver high-quality work to agreed deadlines.\nDesirable knowledge and experience:\nSolid understanding of working with different data formats (e.g. XML, CSV, JSON or RDF)\nAwareness of the value and meaning of archival data and the ethical responsibilities that come with working with public records.\nUnderstanding of probabilistic, messy, or uncertain data handling.\nExperience in Agile development environments.\nFamiliarity with archival principles or metadata standards (e.g. EAD, ISAD(G) or Dublin Core)\nThe Civil Service is committed to attract, retain and invest in talent wherever it is\nfound. To learn more please see the Civil Service People Plan and the Civil Service\nD&I Strategy.\nBenefits\nGenerous benefits package, including pension, sports and social club facilities, onsite gym, discounted rates at our on-site cafe and opportunities for training and development. Annual leave entitlement of 22 days per calendar year (rising to 25 after the first year, and incrementally to 30 days after six years) and 10½ days public and privilege holida...","description_format":"text","description_chars":15641,"description_truncated":true,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Annual leave","Flexible schedule","Hybrid work","Professional development"],"hiring_locations":[{"name":"United Kingdom","iso":"GB","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Government"],"lifecycle":[{"event":"open","at":"2026-09-30T21:49:07Z"}],"liveness":{"score":86,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.86,"p_room":1,"age_days":0,"expected_fill_days":16,"reasons":["conf:7","win:early","comp:junior"],"computed_at":"2026-10-01T05:45:00Z"},"pay":{"stated_usd_annual":59786,"is_top_pay":false},"html_url":"https://alion.io/job/nationalarchives-data-engineer","json_url":"https://alion.io/job/nationalarchives-data-engineer.json","meta":{"generated_at":"2026-10-01T21:56:26Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":4521,"day_limit":5000,"remaining_today":479,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}