{"id":1047726,"url":"https://alion.io/job/software-mind-c3f-data-platform-engineer-3","title":"[C3F] Data Platform Engineer","company":{"id":7464,"name":"Software Mind","domain":"softwaremind.com","url":"https://alion.io/company/softwaremind","size_band":"5000+","is_staffing_agency":true,"is_intermediary":false,"listed_via":null,"ats_vendor":"SmartRecruiters","truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":null,"employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Warsaw, Poland"],"countries":["PL"],"hiring_countries":["PL"],"hiring_countries_total":1,"salary":null,"salary_estimate":{"min_usd":61000,"max_usd":102000,"period":"year","method":"role_country_seniority_unknown","sample_n":275},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Apache Kafka","optional":false},{"name":"Azure","optional":false},{"name":"Azure AKS","optional":false},{"name":"Bicep","optional":false},{"name":"CI/CD","optional":false},{"name":"Databricks","optional":false},{"name":"Kubernetes","optional":false},{"name":"PostgreSQL","optional":false},{"name":"Python","optional":false},{"name":"SQL","optional":false},{"name":"Terraform","optional":false}],"status":"live","first_seen_at":"2026-09-16T13:51:51Z","employer_posted_date":"2026-09-16","last_verified_at":"2026-09-24T21:08:10Z","board_verified":true,"closed_at":null,"days_open":8,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":8},"description":"Software Mind develops solutions that make an impact for companies around the globe. Tech giants & unicorns, transformative projects, emerging technologies and limitless opportunities - these are a few words that describe an average day for us. Building cross-functional engineering teams that take ownership and crave more means we’re always on the lookout for talented people who bring passion and creativity to every project. Our culture embraces openness, acts with respect, shows grit & guts and combines employment with enjoyment.\n Project - the aim you’ll have\nOur customer provides innovative solutions and insights that enable our clients to manage risk and hire the best talent. Their advanced global technology platform supports fully scalable, configurable screening programs that meet the unique needs of over 33,000 clients worldwide. Headquartered in Atlanta, GA, they have an internationally distributed workforce spanning 19 countries with about 5,500 employees. Our partner perform over 93 million screens annually in over 200 countries and territories.\nRole - how you’ll contribute\nBuild and operate change-data-capture pipelines from PostgreSQL into Azure, using Kafka Connect and Debezium as the core of the platform\nConfigure, deploy and scale connectors end to end - connector setup, task management, offsets, schema history, and snapshot strategy\nRun these pipelines as stateful workloads on Kubernetes (AKS), covering configuration, secrets, networking and resource tuning\nMonitor and troubleshoot the platform in production: connector failures, task rebalances, restarts, throughput and backpressure, message-size limits, retries and recovery\nAutomate the platform in Python - configuration-driven onboarding of new data sources, pipeline orchestration, monitoring and alerting, recovery workflows, and automated testing\nIntegrate CDC streams with the wider Azure data stack: Event Hubs, ADLS, Azure PostgreSQL, ADF and Databricks\nManage platform infrastructure as code, so environments are reproducible and changes are reviewable\nApply data protection requirements to sensitive data flowing through the pipelines - masking, hashing, access control and retention\n Expectations - the experience you need\nSolid commercial experience as a data or platform engineer, with hands-on work on streaming or CDC pipelines rather than batch reporting alone\nPractical Kafka knowledge - topics, partitions, offsets, consumer groups and delivery semantics - including at least one Kafka Connect deployment you ran yourself\nStrong SQL and PostgreSQL skills, with working knowledge of WAL, logical replication, replication slots and replication lag\nWorking understanding of CDC concepts: initial snapshots, inserts, updates and deletes, event ordering, at-least-once delivery, and schema evolution\nConfident Python for automation and tooling - orchestration, monitoring, recovery scripts, and automated tests\nHands-on experience with Azure data services, for example Event Hubs, ADLS or Azure PostgreSQL\nComfortable working with Kubernetes as a user: deploying workloads, handling configuration and secrets, reading logs, debugging failing pods\nAbility to debug a running pipeline from metrics and logs - telling throughput problems from backpressure, retries or a genuine connector failure\nAdditional skills - the edge you have\nProduction experience with Debezium specifically - snapshot strategies on large tables, schema history recovery, offset loss, and bringing connectors back after failure\nExperience operating stateful workloads on AKS: StatefulSets, stable worker identity, and resource tuning under load\nInfrastructure-as-code and CI/CD for data platform components (Terraform, Bicep or similar)\nHands-on work with Databricks and ADF at production scale\nExperience implementing data protection controls for sensitive data - masking, hashing, access control and retention policies\n What we offer:\nFlexible and remote cooperation options \nInternational projects with leading global clients \nTravels related to international projects \nNon-corporate atmosphere \nAccess* to language classes\nAccess* to knowledge-sharing initiatives \nAccess* to private healthcare and life insurance \nAccess* to multisport card \n*The availability and terms of individual benefits may vary depending on the chosen form of cooperation.","description_format":"text","description_chars":4331,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[{"language":"English","level":"All levels","optional":false}]},"benefits":["Life insurance"],"hiring_locations":[{"name":"Poland","iso":"PL","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-09-18T22:26:20Z"}],"liveness":{"score":42,"band":"fade","label":"Fading","p_open":1,"p_active":0.448,"p_room":0.945,"age_days":7,"expected_fill_days":18,"reasons":["conf:3","agency","velocity","win:mid"],"computed_at":"2026-09-24T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/software-mind-c3f-data-platform-engineer-3","json_url":"https://alion.io/job/software-mind-c3f-data-platform-engineer-3.json","meta":{"generated_at":"2026-09-25T00:28:30Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":459,"day_limit":5000,"remaining_today":4541,"minute_limit":60,"resets_at":"2026-09-26T00:00:00Z"}}}