{"id":1419572,"url":"https://alion.io/job/prosapient-data-engineer","title":"Data Engineer","company":{"id":685515,"name":"proSapient","domain":"prosapient.com","url":"https://alion.io/company/prosapient","size_band":"201-500","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workable","truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":null,"employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"inferred","locations":[],"countries":[],"hiring_countries":["GB","PT"],"hiring_countries_total":2,"salary":null,"salary_estimate":{"min_usd":74000,"max_usd":174000,"period":"year","method":"role_country_seniority_unknown","sample_n":93},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"BigQuery","optional":false},{"name":"ElasticSearch","optional":false},{"name":"Google BigQuery","optional":false},{"name":"Knowledge Graph","optional":false},{"name":"OpenSearch","optional":false},{"name":"PostgreSQL","optional":false},{"name":"Python","optional":false},{"name":"SQL","optional":false},{"name":"Apache Kafka","optional":true},{"name":"AWS","optional":true},{"name":"dbt","optional":true},{"name":"Docker","optional":true},{"name":"GCP","optional":true},{"name":"Kubernetes","optional":true}],"status":"closed","first_seen_at":"2026-09-28T21:50:59Z","employer_posted_date":"2026-09-28","last_verified_at":"2026-09-29T11:20:56Z","board_verified":false,"closed_at":"2026-09-29T11:20:56Z","days_open":0,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":0},"description":"Every day, somewhere in the world, important decisions are made. Whether it is a private equity company deciding to invest millions into a business or a large corporation implementing a new strategic direction, these decisions impact employees, customers and other stakeholders.\nConsulting and private equity firms come to proSapient when they need to discover knowledge to help them make great decisions and succeed in their goals. It is our mission to support them in their discovery of knowledge.\nWe help our clients find industry experts who can provide their knowledge via interview or survey; we curate this knowledge in a market-leading software platform; and we help clients surface knowledge they already have through expansive knowledge management.\nWe are looking for a Data Engineer to help build and scale the Knowledge Graph at the heart of our AI roadmap.\nYou will join the Data Foundation squad, building the pipelines that populate, maintain and serve the graph as a reliable Content / Data product. This is a hands-on role for someone who enjoys meaningful data modelling challenges across entities, relationships, taxonomies and versioning.\n\nThe key duties of this role will include:\nGraph Construction Pipelines:\n· Build ingestion and transformation pipelines that populate the Knowledge Graph from internal systems, enrichment outputs and third-party sources.\n· Implement entity resolution and linking logic in production.\n· Manage late-arriving data, conflicting sources and graph schema changes over time.\nServing the Graph:\n· Build query and serving layers used by search, matching and data product teams.\n· Model the graph for analytical and application use cases across systems such as BigQuery, PostgreSQL and Elasticsearch / OpenSearch.\n· Support external-facing data products with reliable export and delivery pipelines.\nQuality & Operations:\n· Implement data quality checks, lineage and monitoring across graph pipelines.\n· Own production operations, including alerting, backfills and incremental reprocessing.\n· Contribute to data contracts so downstream teams can rely on the graph.\nRequirements\nThe key skills needed for this role are:\n· 3+ years building production data pipelines.\n· Strong Python and SQL skills, with experience writing clean, production-ready code.\n· Experience with PostgreSQL or other relational databases.\n· Strong data modelling skills, including schema design for evolving requirements.\n· Experience integrating messy, multi-source data, including deduplication and normalisation.\n· Hands-on experience with Elasticsearch / OpenSearch or a comparable serving layer.\n· Comfortable with testing, code review, CI and operating your own pipelines.\n\nNice to have:\n· Experience with graph databases, graph data modelling, taxonomies or ontologies.\n· Experience with entity resolution at scale.\n· Experience with BigQuery, Kafka, dbt or similar data platforms and frameworks.\n· Familiarity with Docker, Kubernetes and cloud environments such as AWS or GCP.\nBenefits\nWhat we can offer you:\n· Tenure gifts, including vouchers, extra holiday and sabbaticals for each year of employment.\n· Health insurance through Vitality.\n· Remote working for up to 20 days each year, giving you flexibility and a change of scenery.\n· Employee Assistance Programme with personalised health and wellbeing advice from specialist teams.\n· Enhanced maternity and paternity pay.\n· 25 days’ annual leave plus bank holidays, including a week’s closure over Christmas.\n· MyMindPal app for online mental fitness support.\n· Corporate events, from quarterly gatherings to annual winter and summer parties.\nWe are committed to building an inclusive workplace. Marginalised groups are often less likely to apply unless they meet every requirement listed, so if you are interested in this role but do not tick every box, we encourage you to apply anyway - it could still be a great match.","description_format":"text","description_chars":3906,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Annual leave","Equity","Health insurance"],"hiring_locations":[{"name":"United Kingdom","iso":"GB","kind":"country"},{"name":"Portugal","iso":"PT","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Machine Learning"],"lifecycle":[{"event":"open","at":"2026-09-28T21:50:59Z"},{"event":"close","at":"2026-09-29T11:20:56Z"}],"liveness":null,"pay":null,"html_url":"https://alion.io/job/prosapient-data-engineer","json_url":"https://alion.io/job/prosapient-data-engineer.json","meta":{"generated_at":"2026-09-30T21:05:13Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"assistant","counted_by":"address","units_charged":1,"used_today":1930,"day_limit":2000,"remaining_today":70,"minute_limit":60,"resets_at":"2026-10-01T00:00:00Z"}}}