368,910open jobs
9,449companies
47,822added this week
Browse all
Location
Remote (Ukraine)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Sigma Software Group is a global software engineering and IT consulting company founded in 2002. Part of the Swedish IT consultancy group Sigma AB (owned by Danir AB) since 2006, the company combines Nordic management practices with Eastern European software engineering talent.

Are you a Senior Data Engineer passionate about building scalable, high-performance data platforms and working with modern Lakehouse technologies? Join Sigma Software’s Data Engineering Center of Excellence and contribute to the modernization of an enterprise-scale analytics ecosystem for the retail domain.

We are looking for a Senior specialist with strong Databricks, PySpark, and cloud data engineering expertise to participate in the migration of a large-scale analytical platform from BigQuery to Databricks. You will collaborate with international teams, contribute to architectural decisions, and help shape reliable and scalable data solutions.

We at Sigma Software create opportunities for continuous learning, technology growth, and meaningful engineering impact while working on complex international projects.

CUSTOMER

Our Customer is a leading retail technology company specializing in AI-driven pricing optimization solutions for enterprise retailers. The company helps businesses improve profitability and competitiveness through advanced analytics, automation, and intelligent pricing strategies. Their platform combines business intelligence with sophisticated algorithms to support data-informed pricing decisions at scale for global retail organizations.

PROJECT

The project focuses on the strategic migration of a large-scale analytical platform from a legacy BigQuery ecosystem to a modern Databricks Lakehouse architecture. The platform processes high-volume retail datasets, machine learning workloads, analytics pipelines, and customer-specific business logic.

As part of the modernization initiative, the engineering team is implementing scalable Spark-based processing, Delta Lake architecture, medallion data layers, and modern governance practices. The role offers an opportunity to work with distributed data processing systems, optimize large-scale workloads, and contribute to the evolution of an enterprise-grade data platform.

Key Technologies: Databricks, Apache Spark, PySpark, Delta Lake, Python, SQL, Airflow, GCP, CI/CD, Unity Catalog

  • Participate in the migration of a large-scale analytical platform from BigQuery to Databricks
  • Design and implement scalable Lakehouse architectures using Databricks and Delta Lake
  • Analyze existing ETL / ELT workloads and define migration approaches
  • Develop and optimize data pipelines processing large volumes of retail and analytical data
  • Implement incremental processing strategies and scalable transformation frameworks
  • Build and maintain Spark-based data processing solutions using PySpark
  • Design and maintain medallion architecture layers including Bronze, Silver, and Gold
  • Implement data governance and security best practices using Unity Catalog
  • Collaborate with Data Science, Analytics, Product, and Customer Engineering teams
  • Participate in architecture discussions and technical solution design
  • Develop reusable data platform components and engineering standards
  • Conduct code reviews and contribute to platform reliability and maintainability
  • Troubleshoot and optimize complex SQL and Spark workloads
  • Support production deployments and platform modernization activities
  • 5+ years of professional experience as a Data Engineer
  • Strong programming skills in Python and advanced SQL
  • Hands-on commercial experience with Databricks
  • Strong knowledge of Apache Spark, primarily PySpark
  • Experience designing and building modern cloud-based data platforms
  • Experience developing ETL / ELT pipelines and large-scale data processing solutions
  • Hands-on experience with Delta Lake
  • Experience with Spark Declarative Pipelines
  • Experience with cluster monitoring, metrics analysis, and performance optimization
  • Strong understanding of distributed data processing architectures
  • Solid understanding of data warehousing concepts and dimensional modeling
  • Experience with Airflow or similar orchestration tools
  • Experience optimizing complex analytical SQL workloads
  • Experience implementing CI / CD practices for data engineering platforms
  • Strong troubleshooting and performance optimization skills
  • Ability to work collaboratively in cross-functional international teams
  • Upper-Intermediate or higher English level

WILL BE A PLUS

  • Experience working with GCP cloud services
  • Experience with AWS or Azure cloud platforms
  • Experience in retail analytics or pricing optimization domains
  • Experience supporting machine learning or AI-related data workloads
  • Experience with platform modernization and cloud migration initiatives

PERSONAL PROFILE

  • Strong analytical and problem-solving mindset
  • Proactive and ownership-driven approach
  • Ability to work independently and collaboratively
  • Good communication and stakeholder collaboration skills
  • Passion for scalable data engineering and modern data platforms
  • Interest in continuous learning and technology innovation
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,910 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Kyiv
$30k – $40k per year • Equity 0–0.2% • In office • Full-Time • 6+ years exp • Bengaluru
Databases
Apache Kafka
MySQL
PostgreSQL
Redis
AI/ML
Flink
Pandas
DevOps
AWS
CI/CD
GCP
Analytics
ETL/ELT
Apply
AI Engineer 4 days ago
$40k – $60k per year • Equity 0.1–0.3% • In office • Full-Time • 6+ years exp • Master's Degree • Bengaluru
Python
Python
Dask
Databases
Chroma
Pinecone
Weaviate
AI/ML
Anomaly Detection
Computer Vision
DeepSeek
Few-Shot Learning
Fine-tuning
In-Context Learning
Kubeflow
LangChain
Llama
LoRA
MLFlow
NLP
Ollama
PEFT
Prompt Engineering
PyTorch
QLoRA
RAG
Ray
Reinforcement Learning
Semantic Search
Spark
TensorFlow
Transformers
vLLM
Hugging Face
Semantic Search
TGI
Time Series Forecasting
DevOps
AWS
Azure
GCP
Vector
Analytics
A/B Testing
Apply
Fullstack Engineer 4 days ago
$90k – $140k per year • Remote • Full-Time • 3+ years exp • Atlanta
JavaScript
Node JS
SQL
TypeScript
Node JS
Prisma
Sequelize
TypeORM
Databases
DynamoDB
MySQL
PostgreSQL
Redis
Frontend
GraphQL
Next.js
React.js
Tailwind CSS
DevOps
AWS
Azure
Docker
GCP
GitHub Actions
WebSockets
GitHub
QA
Cypress
Jest
Playwright
Vitest
Apply
Data Analyst 4 days ago
Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Athens
Python
SQL
Databases
Databricks
DevOps
Azure
Analytics
Power BI
Apply
Android Engineer 4 days ago
$87k – $182k per year (Estimated) • Remote • Full-Time • 3+ years exp
Java
Kotlin
SQL
Databases
DynamoDB
PostgreSQL
Mobile
Jetpack Compose
DevOps
AWS
Apply
$66k – $118k per year (Estimated) • Remote • Full-Time • 5+ years exp • Warsaw
C++
Go
Objective-C
Rust
Apply
Remote • Full-Time • 1+ year exp • Dnipro
PowerShell
Python
AI/ML
LLM
DevOps
Ansible
AWS
Azure
Azure DevOps
CI/CD
Configuration Management
Datadog
GCP
GitHub Actions
GitLab CI
GitOps
Grafana
Jenkins
Kubernetes
Prometheus
TeamCity
Terraform
Vector
Amazon CloudWatch
GitHub
GitLab
Cybersecurity
Least Privilege
Apply
Remote • Full-Time • 1+ year exp • Kharkiv
PowerShell
Python
AI/ML
LLM
DevOps
Ansible
AWS
Azure
Azure DevOps
CI/CD
Configuration Management
Datadog
GCP
GitHub Actions
GitLab CI
GitOps
Grafana
Jenkins
Kubernetes
Prometheus
TeamCity
Terraform
Vector
Amazon CloudWatch
GitHub
GitLab
Cybersecurity
Least Privilege
Apply
Remote • Full-Time • 1+ year exp • Tirana
PowerShell
Python
AI/ML
LLM
DevOps
Ansible
AWS
Azure
Azure DevOps
CI/CD
Configuration Management
Datadog
GCP
GitHub Actions
GitLab CI
GitOps
Grafana
Jenkins
Kubernetes
Prometheus
TeamCity
Terraform
Vector
Amazon CloudWatch
GitHub
GitLab
Cybersecurity
Least Privilege
Apply
Remote • Full-Time • 1+ year exp • Bucharest
PowerShell
Python
AI/ML
LLM
DevOps
Ansible
AWS
Azure
Azure DevOps
CI/CD
Configuration Management
Datadog
GCP
GitHub Actions
GitLab CI
GitOps
Grafana
Jenkins
Kubernetes
Prometheus
TeamCity
Terraform
Vector
Amazon CloudWatch
GitHub
GitLab
Cybersecurity
Least Privilege
Apply
In office • Full-Time • Kyiv
Python
Python
Django
FastAPI
Flask
PyQt
Databases
MySQL
DevOps
AWS
CI/CD
Docker
Docker Compose
Kubernetes
QA
Pytest
Apply
Power BI 2 days ago
In office • Full-Time • 2+ years exp • Kyiv
SQL
DevOps
AWS
Azure
GCP
Analytics
Power BI
Apply
Remote/Hybrid • Full-Time • Kyiv • Warsaw
Apply
Remote/Hybrid • Full-Time • Kyiv • Warsaw
Apply
In office • Full-Time • 5+ years exp • Kyiv
DevOps
FinOps
GCP
Kubernetes
Terraform
Apply
See all jobs
This is one of many
368,910 more open roles from verified company boards, updated every day.