409,719open jobs
14,193companies
73,646added this week
Browse all
Location
Remote (Ukraine)
Seniority
Senior · 7+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Simulmedia is a New York advertising technology company founded in 2008 that buys television audiences with data. Its platform plans and measures campaigns across linear and streaming inventory using viewership and outcome data. The company works with advertisers treating television as a performance channel.

Simulmedia is looking for an experienced and dynamic Data Engineer with a curious and creative mindset to join our Data Services team. The ideal candidate will have a strong background in Python, SQL and large-scale data pipelines. This is an opportunity to join a team of amazing engineers, data scientists, product managers and designers who are obsessed with building the most advanced TV and streaming advertising platform in the market. As a Data Engineer you will design, build and operate the data platform that powers the company: pipelines that ingest and transform very large datasets from external data partners, data models that the whole company queries, and services that make that data available to internal products. You will work on a team that empowers the other teams to use our huge amount of data efficiently. Using a large variety of technologies and tools, you will solve complicated technical problems and build solutions to make our pipelines robust and fault tolerant and our data easily accessible throughout the company.

Location:Ukraine is mandatory. Our offices are located in Kyiv and Lviv. Teams are located in Kyiv and Lviv and primarily work remotely with occasional offline meetings.

Responsibilities:

  • Design and build batch data pipelines that ingest, validate and transform multi-billion-row datasets from external data providers and internal systems
  • Model complex real-world data: dimensional models, reference data, and temporal data whose attributes change over time (e.g. slowly changing dimensions), and evolve those models safely as upstream sources change their schemas and semantics
  • Develop and operate workloads on our lakehouse platform (Databricks / Spark / Delta) and our data warehouse (Redshift), including migrating existing pipelines from the warehouse to the lakehouse
  • Orchestrate pipelines with Airflow: scheduling, dependencies, retries, backfills and alerting
  • Prove correctness, not just completion: design parity checks and reconciliation queries when replacing an existing pipeline, run large historical backfills, and investigate data discrepancies down to the row level
  • Build and maintain Python services and REST APIs that serve data to internal products
  • Optimize for performance and cost: query tuning, table design, workload management and right-sizing compute
  • Own what you ship: monitor production pipelines, participate in incident triage and root-cause analysis, and harden systems so the same failure does not happen twice
  • Collaborate cross-functionally with product managers, data scientists and stakeholders across the company to deliver on product roadmap
  • Work within an Agile team that releases cutting-edge new features regularly
  • Take a high degree of ownership and freedom to experiment with new technologies to improve our software

Qualifications:

  • Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
  • 7+ years of work experience as a data engineer
  • Proficiency in Python and using it as the primary development language in recent years
  • Expert-level SQL: comfortable writing, reading and tuning complex analytical queries against very large tables, and debugging why two result sets disagree
  • Hands-on experience with a distributed data processing platform (Spark/Databricks strongly preferred; EMR, Snowflake or BigQuery also relevant) and with a columnar data warehouse (Redshift, Snowflake, BigQuery, ClickHouse, etc)
  • Ability to design complex data models: normalized, dimensional and temporal (slowly changing dimensions, effective-dated records, point-in-time correctness)
  • Experience with workflow orchestration tools (Airflow or similar): building DAGs, managing dependencies and running backfills
  • Experience integrating third-party data feeds: handling schema drift, late or missing deliveries, vendor data-quality defects and versioned reference data
  • Experience building REST services in Python (FastAPI, Flask, etc)
  • Experience developing, maintaining, and debugging problems in large server-side code bases
  • Working knowledge of AWS (S3, IAM, ECS or similar compute) and Docker
  • Good knowledge of engineering best practices and testing (unit test, integration test, code review, CI/CD)
  • The desire to take a high level of ownership of the things you work on
  • Ability to learn new things quickly, maintain a high bar for quality, and be pragmatic
  • Must be able to communicate with U.S based teams
  • Experience with Delta Lake / medallion lakehouse architectures is a plus
  • Experience migrating legacy pipelines between platforms with strict parity requirements is a plus
  • Experience with advertising, media or measurement industry data is a plus
  • Ability to communicate effectively with the U.S.-based teams and work 11:00 AM - 8:00 PM EEST (11:00 - 20:00).

Our Tech Stack:

  • Almost everything we run is on AWS (S3, ECS, EMR, RDS and more)
  • Python is our primary language; SQL is everywhere
  • Databricks (Spark, Delta Lake) is our lakehouse platform; Redshift and Postgres are our warehouses and operational databases
  • Airflow orchestrates our pipelines
  • Docker for packaging; GitHub Actions and Jenkins for CI/CD
  • Grafana, Sentry and OpenSearch for observability
  • Datasets measured in billions of rows

Interview Process:

  • Pre-screening (30 mins)
  • Technical Interview (1h)
  • System Design (1.5h)
  • Product Interview (30 mins)
  • Offer

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
409,719 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Lviv
In office • Internship • Master's Degree
Python
SQL
Databases
Amazon Redshift
AI/ML
Jupyter Notebook
DevOps
Amazon EC2
Amazon S3
AWS
Analytics
Power BI
Tableau
Apply
Staff Engineer 10 hours ago
$40k – $86k per year (Estimated) • In office • Bengaluru
C#
JavaScript
SQL
TypeScript
C#
ASP.NET Core
Databases
MySQL
PostgreSQL
AI/ML
ChatGPT
Claude
Copilot
Frontend
Angular
React.js
DevOps
Amazon CloudWatch
Amazon EC2
Amazon ECS
Amazon EKS
Amazon S3
AWS
AWS Lambda
Azure
CI/CD
CloudFormation
Datadog
Docker
GCP
Git
GitHub
GitLab
GitLab CI
Kubernetes
Octopus Deploy
Service Mesh
Shift-Left
Splunk
TeamCity
Terraform
Cybersecurity
Shift-Left Security
Apply
Remote/Hybrid • Master's Degree
Java
Python
Scala
SQL
AI/ML
Amazon SageMaker
NumPy
PyTorch
Spark
Time Series Forecasting
DevOps
Amazon EC2
Amazon S3
AWS
Git
Apply
In office • 6+ years exp • Master's Degree
Python
Scala
SQL
Databases
MySQL
Redis
AI/ML
Airflow
Feature Store
Spark
DevOps
Amazon EC2
Amazon S3
AWS
Docker
Git
Jenkins
Kubernetes
Management
Jira
Notion
Apply
$21k – $48k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Mumbai
Python
SQL
Databases
BigQuery
Google BigQuery
Apply
AI Product Specialist 1 month ago
$96k – $194k per year (Estimated) • In office • Full-Time • Bachelor's Degree • New York
Python
Apply
Remote • Full-Time • 5+ years exp • Bachelor's Degree • Lviv
Node JS
Ruby
TypeScript
JavaScript
Frontend
Chart.js
D3.js
Next.js
React.js
Tailwind CSS
tRPC
DevOps
CI/CD
Git
Vercel
QA
Cypress
Playwright
Vitest
Apply
Operation Analyst 1 year ago
$600k – $720k per year • In office • Full-Time • New York
SQL
Management
Google Sheets
Apply
Remote • Full-Time • 5+ years exp • Bachelor's Degree • Lviv
Node JS
Python
Ruby
TypeScript
JavaScript
Ruby
Sinatra
AI/ML
Claude
Claude Code
Cursor
OpenAI Codex
Frontend
Next.js
tRPC
React.js
DevOps
Amazon EC2
Amazon ECS
Amazon EKS
Amazon S3
AWS
AWS CDK
CI/CD
Datadog
Docker
GitHub
GitHub Actions
Grafana
Jenkins
OpenTelemetry
Prometheus
Kubernetes
Analytics
ETL/ELT
QA
Sentry
Apply
In office • Full-Time • 3+ years exp • New York
Apply
Remote • Contractor • 15+ years exp • Lviv
Python
SQL
Python
FastAPI
Databases
Azure SQL Database
AI/ML
LLM
Model Context Protocol
CatBoost
LightGBM
NumPy
Prophet
Scikit-learn
Statsmodels
Time Series Forecasting
Tool Use
DevOps
AWS
Azure
Docker
Apply
Remote • Full-Time • 8+ years exp • Lviv
Java
Python
Scala
SQL
Databases
Apache Iceberg
Trino
AI/ML
Flink
Spark
DevOps
AWS
CI/CD
Platform Engineering
Apply
Manual QA 3 days ago
Remote • Full-Time • 1+ year exp • Lviv
Management
Jira
QA
TestRail
Apply
Remote/Hybrid • Full-Time • Lviv
Java
Python
SQL
TypeScript
JavaScript
Java
Apache Tomcat
Gradle
Spring Boot
Databases
PostgreSQL
Frontend
npm
DevOps
Ansible
CI/CD
Docker
Gerrit
Git
Grafana
Jenkins
JFrog Artifactory
Nginx
Prometheus
Apply
In office • Contractor • 15+ years exp • Lviv
Java
Python
Scala
TypeScript
AI/ML
Claude
Cursor
LLM
DevOps
AWS
CI/CD
Cybersecurity
GDPR
HIPAA
SOC 2
Analytics
ETL/ELT
Apply
See all jobs
This is one of many
409,719 more open roles from verified company boards, updated every day.