{"id":28050,"url":"https://alion.io/job/addepto-senior-data-engineer-databricks-3","title":"Senior Data Engineer (Databricks)","company":{"id":2713,"name":"Addepto","domain":"addepto.com","url":"https://alion.io/company/addepto","size_band":"201-500","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Recruitee","truth_index":{"grade":"C","score":60,"open_postings":15,"ghost_share":0.667,"stale_share":0,"repost_share":0,"time_to_fill_p50_days":null,"computed_at":"2026-09-28T05:45:00Z"}},"role":"Data Science","role_family":"Data Science","seniority":"senior","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Warsaw, Poland"],"countries":["PL"],"hiring_countries":["PL"],"hiring_countries_total":1,"salary":{"min":21000,"max":28560,"currency":"PLN","period":"month","gross":null,"usd_annual":91620},"salary_estimate":null,"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Airflow","optional":false},{"name":"Apache Kafka","optional":false},{"name":"Azure","optional":false},{"name":"CI/CD","optional":false},{"name":"Dagster","optional":false},{"name":"Databricks","optional":false},{"name":"Machine Learning","optional":false},{"name":"Power BI","optional":false},{"name":"pySpark","optional":false},{"name":"Python","optional":false},{"name":"Spark","optional":false},{"name":"SQL","optional":false}],"status":"live","first_seen_at":"2026-07-31T12:33:48Z","employer_posted_date":null,"last_verified_at":"2026-07-31T12:33:48Z","board_verified":false,"closed_at":null,"days_open":59,"trust":{"level":"not_scored","repost_count":null,"flags":[],"days_open":59},"description":"Addepto is a leading AI consulting (\n\nhttps://addepto.com/ai-consulting/\n) and data engineering (\n\nhttps://addepto.com/data-engineering-services/\n) company that builds scalable, ROI-focused AI solutions for some of the world's largest enterprises and pioneering startups, including Rolls Royce, Continental, Porsche, ABB, and WGU. With an exclusive focus on Artificial Intelligence and Big Data, Addepto helps organizations unlock the full potential of their data through systems designed for measurable business impact and long-term growth.\nThe company's work extends beyond client engagements. Drawing from real-world challenges and insights, Addepto has developed its own \n\nproduct - ContextClue\n - and actively contributes open-source solutions to the AI community. This commitment to transforming practical experience into scalable innovation has earned Addepto recognition by \n\nForbes as one of the top 10 AI consulting companies worldwide\n.\nAs part of \n\nKMS Technology\n, a US-based global technology group, Addepto combines deep AI specialization with enterprise-scale delivery capabilities-enabling the partnership to move clients from AI experimentation to production impact, securely and at scale.\nAs a Senior Data Engineer, you will have the exciting opportunity to work with a team of technology experts on challenging projects across various industries, leveraging cutting-edge technologies. Here are some of the projects we are seeking talented individuals to join:\nDesign and development of a universal data platform for global aerospace companies. This Azure and Databricks powered initiative combines diverse enterprise and public data sources. The data platform is at the early stages of the development, covering design of architecture and processes as well as giving freedom for technology selection.\n\nData Platform Transformation for energy management association body. This project addressed critical data management challenges, boosting user adoption, performance, and data integrity. The team is implementing a comprehensive data catalog, leveraging Databricks and Apache Spark/PySpark, for simplified data access and governance. Secure integration solutions and enhanced data quality monitoring, utilizing Delta Live Table tests, established trust in the platform. The intermediate result is a user-friendly, secure, and data-driven platform, serving as a basis for further development of ML components.\n\nDesign of the data transformation and following data ops pipelines for global car manufacturer. This project aims to build a data processing system for both real-time streaming and batch data. We'll handle data for business uses like process monitoring, analysis, and reporting, while also exploring LLMs for chatbots and data analysis. Key tasks include data cleaning, normalization, and optimizing the data model for performance and accuracy.\n\nYour main responsibilities:\nDesign and optimize scalable data processing pipelines for both streaming and batch workloads using Big Data technologies such as Databricks, Apache Airflow, and Dagster.\n\nArchitect and implement end-to-end data platforms, ensuring high availability, performance, and reliability.\n\nLead the development of CI/CD and MLOps processes to automate deployments, monitoring, and model lifecycle management.\n\nDevelop and maintain applications for aggregating, processing, and analyzing data from diverse sources, ensuring efficiency and scalability.\n\nCollaborate with Data Science teams on Machine Learning projects, including text/image analysis, feature engineering, and predictive model deployment.\n\nDesign and manage complex data transformations using Databricks, DBT, and Apache Airflow, ensuring data integrity and consistency.\n\nTranslate business requirements into scalable and efficient technical solutions while ensuring optimal performance and data quality.\n\nEnsure data security, compliance, and governance best practices are followed across all data pipelines.\n\nWhat you'll need to succeed in this role:\nAt least 5 years of commercial experience implementing, developing, or maintaining Big Data systems.\n\nStrong programming skills in Python: writing a clean code, OOP design.\n\nStrong SQL skills, including performance tuning, query optimization, and experience with data warehousing solutions.\n\nExperience in designing and implementing data governance and data management processes.\n\nDeep expertise in Big Data technologies, including Apache Airflow, Dagster, Databricks, and other modern data orchestration and transformation tools.\n\nExperience implementing and deploying solutions in cloud environments (with a preference for Azure).\n\nKnowledge of how to build and deploy Power BI reports and dashboards for data visualization.\n\nExcellent understanding of dimensional data and data modeling techniques.\n\nConsulting experience and the ability to guide clients through architectural decisions, technology selection, and best practices.\n\nAbility to work independently and take ownership of project deliverables.\n\nFamiliarity withSpark, Azure Event Hub or Kafka.\n\nMaster's or Ph.D. in Computer Science, Big Data, Mathematics, Physics, or a related field.\n\nDiscover our perks & benefits:\nWork in a supportive team of passionate enthusiasts of AI & Big Data.\n\nEngage with top-tier global enterprises and cutting-edge startups on international projects.\n\nEnjoy flexible work arrangements, allowing you to work remotely or from modern offices and coworking spaces.\n\nAccelerate your professional growth through career paths, knowledge-sharing initiatives, language classes, and sponsoredtraining or conferences, including a partnership with Databricks, which offers industry-leading training materials and certifications.\n\nChoose your preferred form of cooperation - B2B or a contract of mandate - and enjoy 20 fully paid days off\n\nParticipate inteam-building events and utilize the integration budget.\n\nCelebrate work anniversaries, birthdays, andmilestones.\n\nAccess medical andsports packages, eye care, and well-being support services, including psychotherapy and coaching.\n\nGet full work equipment for optimal productivity, including a laptop and other necessary devices.\n\nWith our backing, you can boost\nyour\npersonal brand by speaking at conferences, writing for our blog, or participating in meetups.\n\nExperience a smooth onboarding with a dedicated buddy, and start your journey in our friendly, supportive, and autonomous culture.\n\nAre you interested in Addepto and would like to join us?\nGet in touch! We are looking forward to receiving your application. Would you like to know more about us?\nVisit our website (career page) and social media (Facebook, LinkedIn, Instagram).","description_format":"text","description_chars":6688,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"master","optional":false},"security_clearance":false,"languages":[]},"benefits":["Flexible schedule"],"hiring_locations":[{"name":"Poland","iso":"PL","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["AI Consulting & Integration","Data Engineering & Migration Services"],"lifecycle":[{"event":"open","at":"2026-09-05T11:00:00Z"}],"liveness":{"score":13,"band":"cold","label":"Long shot","p_open":0.4,"p_active":0.576,"p_room":0.55,"age_days":58,"expected_fill_days":40,"reasons":["seen:58","stale_co","velocity","win:tail"],"computed_at":"2026-09-28T05:45:00Z"},"pay":{"stated_usd_annual":91620,"is_top_pay":false},"html_url":"https://alion.io/job/addepto-senior-data-engineer-databricks-3","json_url":"https://alion.io/job/addepto-senior-data-engineer-databricks-3.json","meta":{"generated_at":"2026-09-29T03:50:28Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":3644,"day_limit":5000,"remaining_today":1356,"minute_limit":60,"resets_at":"2026-09-30T00:00:00Z"}}}