{"id":1330244,"url":"https://alion.io/job/dotlinkers-senior-data-engineer","title":"Senior Data Engineer","company":{"id":1813,"name":"dotLinkers","domain":"dotlinkers-itrecruitment.com","url":"https://alion.io/company/dotlinkers","size_band":"11-50","is_staffing_agency":true,"employer_type":"agency","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":"senior","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"explicit","locations":["Kraków, Poland"],"countries":["PL"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":24000,"max":29000,"currency":"PLN","period":"month","gross":null,"usd_annual":90768},"salary_estimate":null,"experience_years_min":3,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Amazon S3","optional":false},{"name":"Apache Kafka","optional":false},{"name":"AWS","optional":false},{"name":"BigQuery","optional":false},{"name":"Dagster","optional":false},{"name":"Databricks","optional":false},{"name":"GDPR","optional":false},{"name":"Google BigQuery","optional":false},{"name":"Python","optional":false},{"name":"Snowflake","optional":false},{"name":"SQL","optional":false},{"name":"Trino","optional":false}],"status":"live","first_seen_at":"2026-09-27T09:00:17Z","employer_posted_date":null,"last_verified_at":"2026-09-27T09:00:17Z","board_verified":false,"closed_at":null,"days_open":2,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":2},"description":"Position: Senior Database Engineer\n\nLocation: Kraków, Poland\n\nWorking mode: Initially remote; once the Kraków office is ready, 3-4 days per week onsite\n\nAbout our client\nOur client is an Israeli sports-tech company combining advanced data engineering, AI/ML and mathematical modelling to build high-performance real-time sports analytics solutions. The company processes large volumes of real-time data across sports including football, basketball, baseball and American football, transforming it into predictions and actionable insights. Its technology supports data-driven organizations operating in environments where speed, scalability and data accuracy are critical. The company is now building its new R&D team in Kraków, giving engineers the opportunity to join at an early stage and help shape the technical foundations of the new organization.\nThe Role\nYou'll own our data lake and lakehouse architecture - the system of record for millions of financial events every day, generated by 12M+ active users, including bets, wallet movements and live odds.\nYou will be responsible for how data lands, is stored, retained, governed and made available across AWS S3, Snowflake and Databricks. Your work will ensure that analytics, finance and regulatory teams have accurate, reconciled data with no drift from source systems.\nThis is an architecture and ownership role rather than a heads-down coding position. You'll make key decisions around data architecture, performance, reliability, governance and cost.\nWhat you'll be doing\nOwn the lakehouse architecture, including bronze/silver/gold layers, Iceberg/Delta tables and schema evolution.\n\nDesign and maintain CDC-based streaming ingestion using Kafka, Debezium or equivalent, including handling late and duplicate events.\n\nDesign data layouts for performance and cost efficiency, including partitioning, compaction, file sizing and query optimization across Trino, Athena and Snowflake.\n\nOwn data retention and archival, including storage tiering, regulatory retention, immutability and GDPR deletion.\n\nGuarantee data correctness through freshness SLAs, drift detection and reconciliation against source wallet and ledger systems.\n\nOwn data governance, including catalog and lineage, row/column-level access control, PII masking, encryption and audit trails.\n\nMonitor ingestion health, data anomalies and cloud storage/compute spend.\n\nWork closely with engineering, analytics, finance and other stakeholders to ensure the data platform supports business and regulatory requirements.\n\nWhat we're looking for\n3+ years of experience in data engineering, with real ownership of a large-scale data lake or lakehouse.\n\nProduction experience with cloud object storage such as AWS S3 and Snowflake, Databricks or BigQuery.\n\nHands-on experience with an open table format: Iceberg, Delta or Hudi.\n\nDeep understanding of partitioning, file layout and query optimization at terabyte-plus scale.\n\nExperience with CDC and streaming ingestion, preferably Kafka + Debezium or equivalent.\n\nStrong SQL and data modelling skills, with a good understanding of relational/OLTP systems and financial data flows.\n\nExperience with data lake governance: catalogs, lineage, access control, PII, retention and security.\n\nPython and an orchestrator such as Airflow or Dagster - used as tools rather than being the primary focus of the role.\n\nNice to have: experience in fintech, iGaming or another regulated, audit-heavy environment.\n\nWhy this role is interesting\nYou'll have real ownership of a critical data platform, rather than simply implementing tasks defined by someone else.\n\nYou'll work with large-scale financial data, where correctness, traceability and reliability are critical. The role gives you the opportunity to influence architecture, technology choices, data governance, performance and cloud costs.\n\nThe scale is significant: 12M+ active users and millions of financial events processed every day.\n\nWhat We Offer\nOpportunity to join a new R&D organization in Kraków and help shape its technical foundations from the beginning.\n\nB2B or perm contract (UoP)\n\nInitially fully remote work, followed by a hybrid model with 3-4 days per week from the Kraków office once the office is ready.\n\nOpportunity to work with high-volume transactional and real-time data systems serving millions of active users.\n\nInternational environment and collaboration with experienced technical teams and company leadership.\n\nCompetitive compensation and the equipment required to perform the role.\n\nOpportunity to work on challenging problems around database scalability, data consistency, performance and real-time processing.","description_format":"text","description_chars":4651,"description_truncated":false,"requirements":{"experience_years_min":3,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"Poland","iso":"PL","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-09-27T09:23:52Z"}],"liveness":{"score":57,"band":"ok","label":"Likely open","p_open":1,"p_active":0.568,"p_room":1,"age_days":1,"expected_fill_days":30,"reasons":["seen:1","agency","urgency","win:early"],"computed_at":"2026-09-29T05:45:00Z"},"pay":{"stated_usd_annual":90768,"is_top_pay":false},"html_url":"https://alion.io/job/dotlinkers-senior-data-engineer","json_url":"https://alion.io/job/dotlinkers-senior-data-engineer.json","meta":{"generated_at":"2026-09-30T02:36:50Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1822,"day_limit":5000,"remaining_today":3178,"minute_limit":60,"resets_at":"2026-10-01T00:00:00Z"}}}