{"id":2117631,"url":"https://alion.io/job/samba-senior-data-engineer","title":"Senior Data Engineer","company":{"id":1800218,"name":"Samba","domain":"samba.com","url":"https://alion.io/company/samba-com","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":null,"truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":"senior","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Warsaw, Poland"],"countries":["PL"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":312000,"max":null,"currency":"PLN","period":"year","gross":null,"usd_annual":79821},"salary_estimate":null,"experience_years_min":8,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Airflow","optional":false},{"name":"AWS","optional":false},{"name":"Databricks","optional":false},{"name":"GCP","optional":false},{"name":"pySpark","optional":false},{"name":"Python","optional":false},{"name":"Spark","optional":false}],"status":"live","first_seen_at":"2026-09-14T10:17:19Z","employer_posted_date":null,"last_verified_at":"2026-09-14T10:17:19Z","board_verified":false,"closed_at":null,"days_open":27,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":27},"description":"About Samba\n\nSamba is a media intelligence company. We know what the world is watching, reading, and thinking about — in real time, at scale, across every screen. Our data exists with the consent of over a billion people, organized into the most complete picture of consumer attention ever built. The biggest brands in the world use that picture to make smarter decisions. We think it’s the most interesting data asset on the planet, because it’s the most culturally relevant.\n\nRole Overview\n\nAs a Senior Data Engineer, you will be responsible for leading the development of scalable, high-performance data pipelines and infrastructure that power Samba analytics and insights. You will play a critical role in designing and implementing architectural improvements, ensuring best practices, and mentoring a team of engineers. You will collaborate closely with Data Science, Analytics, and Product teams to deliver robust, production-ready data solutions that drive business impact.\n\nWhat You'll Do\n\nArchitectural Leadership: Lead the design and development of scalable, high-performance data pipelines and infrastructure that power Samba TV's analytics.\n\nComplex Problem Solving: Resolve complex technical issues in creative and effective ways, understanding the interrelationships of different disciplines.\n\nProduction Dataset Management: Lead the design, build, and maintenance of high-scale production datasets. Ensure the delivery of versioned outputs and reliable customer-facing reports.\n\nSchema Evolution & Lifecycle: Manage the end-to-end lifecycle of data features - adding new attributes, updating business logic, and executing safe rollouts (staging → performance check → production), including complex backfills and reprocessing.\n\nPerformance & Cost Engineering: Architect scalable and efficient solutions for data ingestion, transformation, and storage, ensuring performance, reliability, and security. Proactively drive down compute and storage costs.\n\nIncident Response & Debugging: Serve as a technical lead for production incidents. Investigate root causes, implement permanent fixes, and validate data recovery across the ecosystem.\n\nCross-Functional Influence: Network with key contacts outside your own area of expertise and frequently advise others on complex matters.\n\nMentorship: Guide the development of new policies and ideas while mentoring and guiding other engineers to foster a culture of technical excellence.\n\nObservability: Enhance monitoring and observability of data processes, improving debugging, error detection, and system reliability at scale.\n\nWymagania\nWho You Are\n\nExperience: Typically 8+ years of related experience with a Bachelor’s degree (or 6 years with a Master’s; 3 years with a PhD).\n\nTechnical Mastery: Advanced knowledge of Python and deep understanding of distributed data processing frameworks like Apache Spark or PySpark.\n\nOrchestration Expertise: Must-have expertise in Apache Airflow and Databricks for orchestration and scalable data processing.\n\nInfrastructure: Extensive experience with cloud-based data infrastructure (AWS, GCP) and modern data lake architectures.\n\nData Modeling: Strong knowledge of data modeling, database design, and query optimization for both relational and non-relational databases.\n\nSoftware Excellence: Proven track record of driving best practices for code quality, testing, and software design in a data engineering context.\n\nCommunication: Ability to adapt your communication style and use persuasion when delivering messages that relate to the wider firm business.\n\nOferujemy\nSalary range: 312,000 - 395,000 PLN per year.","description_format":"text","description_chars":3612,"description_truncated":false,"requirements":{"experience_years_min":8,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Media & Audience Measurement"],"lifecycle":[{"event":"open","at":"2026-10-08T23:22:33Z"}],"visa":[],"liveness":{"score":30,"band":"fade","label":"Fading","p_open":0.85,"p_active":0.652,"p_room":0.55,"age_days":25,"expected_fill_days":22,"reasons":["seen:25","win:tail"],"computed_at":"2026-10-10T05:45:15Z"},"pay":{"stated_usd_annual":79821,"is_top_pay":false},"html_url":"https://alion.io/job/samba-senior-data-engineer","json_url":"https://alion.io/job/samba-senior-data-engineer.json","meta":{"generated_at":"2026-10-11T19:07:56Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler_verified","counted_by":"address","units_charged":1,"used_today":5658,"day_limit":null,"remaining_today":null,"minute_limit":300,"resets_at":"2026-10-12T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":1800218},"rest":"https://alion.io/mcp/rest/get_company?id=1800218"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fsamba-senior-data-engineer"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fsamba-senior-data-engineer"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fsamba-senior-data-engineer"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/samba-senior-data-engineer\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fsamba-senior-data-engineer"}]}