Overview
Company
Profile match
Impact
Conditions
Benefits
Hiring process
Similar jobs

Deutsche Telekom

Deutsche Telekom is a leading global telecommunications company and the largest network provider in Europe by revenue. Headquartered in Bonn, Germany, the firm offers fixed-line broadband, mobile network services, internet solutions, and enterprise IT services worldwide. It is best known internationally for its mobile operations under the T-Mobile brand and its corporate IT services division, T-Systems.

As Hungary’s most attractive employer in 2025 (according to Randstad’s representative survey), Deutsche Telekom IT Solutions is a subsidiary of the Deutsche Telekom Group. The company provides a wide portfolio of IT and telecommunications services with more than 5300 employees. We have hundreds of large customers, corporations in Germany and in other European countries. DT-ITS recieved the Best in Educational Cooperation award from HIPA in 2019, acknowledged as the the Most Ethical Multinational Company in 2019. The company continuously develops its four sites in Budapest, Debrecen, Pécs and Szeged and is looking for skilled IT professionals to join its team.

We are looking for a Senior Data Engineer to build and operate scalable data ingestion and CDC capabilities on our Azure-based Lakehouse platform. Beyond developing pipelines in Azure Data Factory and Databricks, you will help us mature our engineering approach: we increasingly deliver ingestion and CDC preparation through Python projects and reusable frameworks, and we expect this role to apply professional software engineering practices (clean architecture, testing, code reviews, packaging, CI/CD, and operational excellence).

Our platform runs batch-first processing, with streaming sources landed raw and processed in batch and selective evolution toward streaming where needed.

You will work within the Common Data Intelligence Hub, collaborating with data architects, analytics engineers, and solution designers to enable robust data products and governed data flows across the enterprise.

  • Your team owns ingestion & CDC engineering end-to-end (design, build, operate, observability, reliability, reusable components).
  • You contribute to platform standards (contracts, layer semantics, readiness criteria) and reference implementations.
  • You do not primarily own cloud infrastructure provisioning (e.g., enterprise networking, core IaC foundations), but you collaborate with the platform team by defining requirements, reviewing changes, and maintaining deployable code for pipelines and jobs.

Platform data engineering & delivery

  • Design and develop ingestion pipelines using Azure and Databricks services (ADF pipelines, Databricks notebooks/jobs/workflows).
  • Implement and operate CDC patterns (inserts, updates, deletes), including late arriving data and reprocessing strategies.
  • Structure and maintain bronze and silver Delta Lake datasets (schema enforcement, de-duplication, performance tuning).
  • Build “transformation-ready” datasets and interfaces (stable schemas, contracts, metadata expectations) for analytics engineers and downstream modeling.
  • Ingest data in a batch-first approach (raw landing, replayability, idempotent batch processing), and help evolve patterns toward true streaming where future use cases require it.

Software engineering for data frameworks

  • Develop and maintain Python-based ingestion/CDC components as production-grade software (modules/packages, versioning, releases).
  • Apply engineering best practices: code reviews, unit/integration tests, static analysis, formatting/linting, type hints, and clear documentation.
  • Establish and improve CI/CD pipelines for data engineering code and pipeline assets (build, test, security checks, deploy, rollback patterns).
  • Drive reuse via shared libraries, templates, and reference implementations; reduce “one-off notebook” solutions.

Operations, reliability & observability

  • Implement logging, metrics, tracing, and data pipeline observability (run-time KPIs, SLAs, alerting, incident readiness).
  • Troubleshoot distributed processing and production issues end-to-end.
  • Work with solution designers on event-based triggers and orchestration workflows; contribute to operational standards.
  • Implement operational and security hygiene: secure secret handling, least-privilege access patterns, and support for auditability (e.g., logs/metadata/lineage expectations).

Collaboration & leadership

  • Mentor other engineers and promote consistent engineering practices across teams.
  • Contribute to the Data Engineering Community of Practice and help define standards, patterns, and guardrails.
  • Contribute to architectural discussions (layer semantics, readiness criteria, contracts, and governance).
  • Work with architects and governance stakeholders to ensure datasets meet governance requirements (cataloging, ownership, documentation, access patterns, compliance constraints) before promotion to higher layers.
  • 3-5 years of hands-on experience building data pipelines with Databricks and Azure in production.
  • Strong knowledge of Delta Lake patterns (CDC, schema evolution, deduplication, partitioning, performance optimization).
  • Advanced Python engineering skills: building maintainable projects (packaging, dependency management, testing, tooling).
  • Solid SQL skills (complex transformations, debugging, performance tuning).
  • Proven experience with CI/CD and Git-based workflows (merge requests, branching strategies, automated testing, environment promotion).
  • Ability to diagnose and resolve issues in distributed systems (Spark execution, cluster/runtime behavior, data correctness).
  • Good understanding of data modeling principles and how they influence ingestion and performance.
  • Practical experience applying data governance and security controls in a Lakehouse environment (permissions/access patterns, secure secret handling, audit needs; Unity Catalog is a plus).
  • Proactive, reliable, and able to work independently within agile teams.
  • Strong communication skills in English (spoken and written).

Technical Core Skills

  • Databricks (Jobs/Workflows, Notebooks, Spark, Autoloader, Delta Lake)
  • Azure Functions & Durable Functions (orchestration, long-running workflows)
  • SQL (analysis + performance tuning)
  • PySpark and Python (production-grade)
  • Azure Data Factory (pipelines, triggers, linked services, monitoring)
  • ADLS Gen2 (lake storage design, folder/partition strategy, access controls, lifecycle/retention)

Software engineering toolchain

  • Git + code review workflows
  • CI/CD pipelines (e.g., GitLab CI, Azure DevOps)
  • Testing: unit/integration tests, test data strategies
  • Code quality: linting/formatting, static analysis, type hints
  • Packaging & dependency management (e.g., Poetry/pip-tools/conda - whichever you standardize on)

Governance, security & orchestration

  • Unity Catalog (cataloging, permissions/access patterns, basic governance controls)
  • Secure secret handling and service authentication patterns (Key Vault or equivalent)
  • Event Grid / Azure Functions / event-driven orchestration
  • Observability (structured logging, metrics, alerting; Log Analytics or equivalent)

* Please be informed that our remote working possibility is only available within Hungary due to European taxation regulation.

Recommended for you based on this role

Similar stack
Same company
In your city
Remote • Full-Time • 5+ year exp • Budapest
Python
Scala
SQL
Databases
Apache Kafka
AI/ML
dbt
Spark
DevOps
AWS
Azure
GCP
Analytics
Power BI
Tableau
Apply
Remote • Full-Time • 5+ year exp • PhD • Budapest
Python
Databases
OpenSearch
DevOps
Alertmanager
Ansible
ArgoCD
CI/CD
Docker
GitLab CI
GitOps
Grafana
Harbor
Helm
Jenkins
JFrog Artifactory
Kubernetes
Loki
OpenTelemetry
Prometheus
SLI/SLO/SLA
Tekton
Terraform
Cybersecurity
Cosign
Grype
SBOM
SLSA
Syft
Trivy
Apply
Remote • Full-Time • 5+ year exp • Budapest
Python
Databases
OpenSearch
Qdrant
AI/ML
Ollama
Open WebUI
vLLM
DevOps
Alertmanager
Ansible
ArgoCD
CI/CD
Configuration Management
Crossplane
GitOps
Grafana
Helm
Incident Management
Jenkins
Kubernetes
Loki
OpenTelemetry
Platform Engineering
Prometheus
Terraform
Vector
Cybersecurity
Keycloak
SBOM
Apply
Remote • Full-Time • 8+ year exp • Budapest
C++
Go
Java
Python
Rust
AI/ML
AutoGen
Claude
Claude Code
Cline
Code Llama
Cursor
DeepSeek
DeepSeek-Coder
LangGraph
Llama
LlamaIndex
LLM
Mistral
Qwen
RAG
StarCoder
Windsurf
LangChain
DevOps
ArgoCD
CI/CD
Docker
Helm
Jenkins
Kubernetes
OpenStack
Platform Engineering
Apply
Remote • Full-Time • 2+ year exp • Budapest
Python
Python
FastAPI
AI/ML
Aider
Claude
Claude Code
Cline
Code Llama
Cursor
DeepSeek
DeepSeek-Coder
LangChain
Llama
LlamaIndex
LLM
Mistral
Prompt Engineering
Qwen
RAG
StarCoder
Windsurf
DevOps
CI/CD
Docker
Git
Helm
Jenkins
Kubernetes
OpenStack
Apply
Remote • Full-Time • 5+ year exp • Budapest
C++
Go
Java
Python
Rust
AI/ML
Aider
Claude
Claude Code
Cline
Code Llama
Cursor
DeepSeek
DeepSeek-Coder
Llama
Qwen
StarCoder
Windsurf
DevOps
ArgoCD
CI/CD
Docker
Helm
Jenkins
Kubernetes
OpenStack
Platform Engineering
Cybersecurity
Falco
GitGuardian
Grype
OWASP Dependency-Check
SBOM
Semgrep
SonarQube
Syft
Trivy
Apply
Remote • Full-Time • 5+ year exp • Budapest
Python
AI/ML
Aider
AutoGen
Claude
Claude Code
Cline
Code Llama
CrewAI
Cursor
DeepSeek
DeepSeek-Coder
LangGraph
Llama
LlamaIndex
Qwen
RAG
Semantic Kernel
StarCoder
Windsurf
LangChain
DevOps
CI/CD
Git
Jenkins
Kubernetes
OpenStack
Platform Engineering
Vector
Apply
Remote • Full-Time • 5+ year exp • Budapest
C++
Go
Java
Python
Rust
Assembly
Assembly
Keystone Engine
AI/ML
Claude
Claude Code
Cline
Code Llama
Cursor
DeepSeek
DeepSeek-Coder
Llama
Qwen
StarCoder
Windsurf
DevOps
ArgoCD
CI/CD
Docker
Git
Helm
Jenkins
Kubernetes
OpenStack
Platform Engineering
Apply
Remote • Full-Time • Budapest
Java
Java
Apache Camel
Maven
Quarkus
Spring Boot
Databases
ActiveMQ
Apache Kafka
DevOps
Ansible
ArgoCD
CI/CD
Docker
Git
Grafana
Helm
Kubernetes
Kustomize
OpenShift
Prometheus
Management
Confluence
Jira
Apply
Remote • Full-Time • 5+ year exp • Budapest
Java
Java
Spring Boot
Spring Security
Databases
PostgreSQL
DevOps
CI/CD
Git
Rest API
Cybersecurity
OWASP Top 10
Threat Modeling
Apply
Career impact
Discover how this job can transform your career
Get a personal career forecast for this job - salary uplift, next-level role, skill boost and a 3-year financial impact, all calculated from your profile.
Personal salary uplift vs. your current pay
Your 3-year career trajectory
Skills you will level up in this role
3-year financial impact in dollars
Create free account
Free forever • Less than a minute • No credit card

Work setup

Location
Budapest
Remote work
Remote (Hungary)
Employment
Full-Time