Overview
Company
Profile match
Impact
Conditions
Benefits
Hiring process
Similar jobs

Sigma Software Group

Sigma Software Group is a global software engineering and IT consulting company founded in 2002. Part of the Swedish IT consultancy group Sigma AB (owned by Danir AB) since 2006, the company combines Nordic management practices with Eastern European software engineering talent.

Are you a Senior Data Engineer passionate about building scalable, high-performance data platforms and working with modern Lakehouse technologies? Join Sigma Software’s Data Engineering Center of Excellence and contribute to the modernization of an enterprise-scale analytics ecosystem for the retail domain.

We are looking for a Senior specialist with strong Databricks, PySpark, and cloud data engineering expertise to participate in the migration of a large-scale analytical platform from BigQuery to Databricks. You will collaborate with international teams, contribute to architectural decisions, and help shape reliable and scalable data solutions.

We at Sigma Software create opportunities for continuous learning, technology growth, and meaningful engineering impact while working on complex international projects.

CUSTOMER

Our Customer is a leading retail technology company specializing in AI-driven pricing optimization solutions for enterprise retailers. The company helps businesses improve profitability and competitiveness through advanced analytics, automation, and intelligent pricing strategies. Their platform combines business intelligence with sophisticated algorithms to support data-informed pricing decisions at scale for global retail organizations.

PROJECT

The project focuses on the strategic migration of a large-scale analytical platform from a legacy BigQuery ecosystem to a modern Databricks Lakehouse architecture. The platform processes high-volume retail datasets, machine learning workloads, analytics pipelines, and customer-specific business logic.

As part of the modernization initiative, the engineering team is implementing scalable Spark-based processing, Delta Lake architecture, medallion data layers, and modern governance practices. The role offers an opportunity to work with distributed data processing systems, optimize large-scale workloads, and contribute to the evolution of an enterprise-grade data platform.

Key Technologies: Databricks, Apache Spark, PySpark, Delta Lake, Python, SQL, Airflow, GCP, CI/CD, Unity Catalog

  • Participate in the migration of a large-scale analytical platform from BigQuery to Databricks
  • Design and implement scalable Lakehouse architectures using Databricks and Delta Lake
  • Analyze existing ETL / ELT workloads and define migration approaches
  • Develop and optimize data pipelines processing large volumes of retail and analytical data
  • Implement incremental processing strategies and scalable transformation frameworks
  • Build and maintain Spark-based data processing solutions using PySpark
  • Design and maintain medallion architecture layers including Bronze, Silver, and Gold
  • Implement data governance and security best practices using Unity Catalog
  • Collaborate with Data Science, Analytics, Product, and Customer Engineering teams
  • Participate in architecture discussions and technical solution design
  • Develop reusable data platform components and engineering standards
  • Conduct code reviews and contribute to platform reliability and maintainability
  • Troubleshoot and optimize complex SQL and Spark workloads
  • Support production deployments and platform modernization activities
  • 5+ years of professional experience as a Data Engineer
  • Strong programming skills in Python and advanced SQL
  • Hands-on commercial experience with Databricks
  • Strong knowledge of Apache Spark, primarily PySpark
  • Experience designing and building modern cloud-based data platforms
  • Experience developing ETL / ELT pipelines and large-scale data processing solutions
  • Hands-on experience with Delta Lake
  • Experience with Spark Declarative Pipelines
  • Experience with cluster monitoring, metrics analysis, and performance optimization
  • Strong understanding of distributed data processing architectures
  • Solid understanding of data warehousing concepts and dimensional modeling
  • Experience with Airflow or similar orchestration tools
  • Experience optimizing complex analytical SQL workloads
  • Experience implementing CI / CD practices for data engineering platforms
  • Strong troubleshooting and performance optimization skills
  • Ability to work collaboratively in cross-functional international teams
  • Upper-Intermediate or higher English level

WILL BE A PLUS

  • Experience working with GCP cloud services
  • Experience with AWS or Azure cloud platforms
  • Experience in retail analytics or pricing optimization domains
  • Experience supporting machine learning or AI-related data workloads
  • Experience with platform modernization and cloud migration initiatives

PERSONAL PROFILE

  • Strong analytical and problem-solving mindset
  • Proactive and ownership-driven approach
  • Ability to work independently and collaboratively
  • Good communication and stakeholder collaboration skills
  • Passion for scalable data engineering and modern data platforms
  • Interest in continuous learning and technology innovation

Recommended for you based on this role

Similar stack
Same company
In your city
Full-Time • 7+ year exp • Lviv
JavaScript
Node JS
AI/ML
Claude
Claude Code
Frontend
GraphQL
React.js
DevOps
Rest API
Apply
In office • Full-Time • 4+ year exp • Warsaw
C#
SQL
TypeScript
JavaScript
C#
.NET
Entity Framework Core
Databases
Apache Kafka
MS SQL
AI/ML
Claude
Copilot
Frontend
Angular
DevOps
Azure
Git
Rest API
Apply
Remote • Full-Time • 6+ year exp • Bachelor's Degree • Bucharest
C#
Java
Python
TypeScript
JavaScript
Java
Spring Boot
Databases
Apache Kafka
AI/ML
AI Agents
LLM
Frontend
Angular
DevOps
ArgoCD
Azure
Helm
Kubernetes
Terraform
Apply
Senior AI Engineer 8 days ago
Remote • Full-Time • 5+ year exp • Warsaw
Go
AI/ML
AI Agents
LLM
RAG
Semantic Search
Spark
DevOps
AWS
Azure
GCP
Kubernetes
Vector
Apply
Remote • Full-Time • 3+ year exp • Warsaw
JavaScript
Node JS
Python
TypeScript
Node JS
Nest.JS
Databases
NATS
PostgreSQL
Frontend
Next.js
React.js
DevOps
CI/CD
Docker
Git
Kubernetes
Cybersecurity
Keycloak
Apply
Remote • Full-Time • 3+ year exp • Warsaw
C++
C
C++
CMake
Qt
C
Valgrind
Design
Figma
Apply
Remote • Full-Time • 3+ year exp • Warsaw
JavaScript
Node JS
Python
TypeScript
Node JS
Nest.JS
Databases
NATS
PostgreSQL
Frontend
Next.js
React.js
DevOps
CI/CD
Docker
Git
Kubernetes
Cybersecurity
Keycloak
Apply
Remote • Full-Time • 4+ year exp • Warsaw
C++
AI/ML
Computer Vision
DevOps
RTOS
IoT
FreeRTOS
MQTT
Apply
Remote • Full-Time • 3+ year exp • Warsaw
JavaScript
Node JS
Python
TypeScript
Node JS
Nest.JS
Databases
NATS
PostgreSQL
Frontend
Next.js
React.js
DevOps
CI/CD
Docker
Git
Kubernetes
Cybersecurity
Keycloak
Apply
Remote • Full-Time • 4+ year exp • Brasília
Python
JavaScript
Databases
Amazon Redshift
Cassandra
HBase
Frontend
GraphQL
React.js
DevOps
AWS
Azure
CI/CD
GCP
Rest API
Apply
Career impact
Discover how this job can transform your career
Get a personal career forecast for this job - salary uplift, next-level role, skill boost and a 3-year financial impact, all calculated from your profile.
Personal salary uplift vs. your current pay
Your 3-year career trajectory
Skills you will level up in this role
3-year financial impact in dollars
Create free account
Free forever • Less than a minute • No credit card

Work setup

Location
Lisbon
Remote work
Remote
Employment
Full-Time

Compensation

Benefits
Continuous learning