405,710open jobs
14,091companies
78,515added this week
Browse all
Salary
$48k – $89k per year
Location
In office (Warsaw)
Seniority
Middle · 3+ years exp
Employment
Internship
Overview
Company
Impact
Profile match
Samba TV measures television viewing through software embedded in smart TVs from major manufacturers. Advertisers use the resulting panel to see which households saw a campaign and to target follow-up advertising on other devices. The company operates across dozens of countries and publishes regular viewership research.

As a mid-level Data Scientist at Samba in Warsaw, you will own end-to-end delivery of significant data science projects with minimal guidance. You are a reliable, autonomous contributor with deep expertise in at least one of Samba's core domains - measurement, or audience modelling - and the technical range to build production-ready solutions using modern ML and AI methodologies. You'll work closely with peers, product, and engineering, and play an active role in mentoring junior data scientists on the team.

What You'll Do:

  • Own end-to-end delivery of significant data science projects - from problem scoping and approach design through to production deployment
  • Make sound, independently-reasoned decisions on methodology, model selection, and evaluation; document them clearly in technical solution documents covering problem statement, approach, metrics, and timeline
  • Lead solution design for your own initiatives; break down complex epics into well-scoped user stories with clear acceptance criteria, adopting DataOps and MLOps best practices throughout - experiment tracking, pipeline orchestration, model monitoring, and reproducibility
  • Build production-quality Python and PySpark code on Databricks - well-tested, documented, and reusable - and implement advanced ML and AI-powered workflows including entity resolution, probabilistic record linkage, embedding-based matching, semantic similarity, and LLM-augmented pipelines
  • Develop and maintain reusable tools, libraries, and documentation that improve team efficiency and technical standards; conduct code reviews with constructive, specific feedback that raises the bar
  • Mentor junior data scientists on technical execution, code quality, and career development; lead internal talks or workshops on ML topics
  • Collaborate cross-functionally with product, engineering, and operations - translate business requirements into technical specifications, partner with data engineering on scalable pipeline design, and participate in cross-functional design reviews and working groups

Who You Are:

  • Bachelor's degree required in Statistics, Data Science, Computer Science, Mathematics or a related quantitative field; Master's strongly preferred
  • 3-5 years of hands-on data science experience with demonstrated ability to own and deliver complex, multi-sprint projects independently
  • Advanced Python with production-quality code, testing, and documentation; strong SQL and PySpark for billion-row datasets
  • Databricks workflows, Delta Lake, and job orchestration; working knowledge of cloud platforms (AWS or GCP)
  • Solid command of core ML - regression, classification, clustering, model evaluation, and experimental design - applied to complex, high-volume data
  • Proficiency with MLOps practices: experiment tracking, pipeline orchestration (Airflow), and reproducible model deployment
  • Exposure to modern AI methodologies: RAG systems, LLM-augmented models, vector databases, and semantic search
  • Strong communicator - able to translate technical work into clear documentation, user stories, and cross-functional conversations
  • Demonstrated ability to mentor junior data scientists and contribute to team standards

Preferred skills:

  • Hands-on experience with knowledge graph construction, entity resolution, or semantic data modeling (RDF, OWL, SPARQL, or equivalent graph frameworks)
  • Familiarity with probabilistic record linkage, identity graph approaches, or embedding-based entity matching at scale
  • Experience with causal inference methods (A/B testing, synthetic control, uplift modeling)
  • Experience with deduplication, enrichment, or web-to-TV linkage problems
  • Background in media, ad tech, or measurement - TV viewership (ACR/STB data), digital audience modeling, cross-platform measurement (linear + CTV/OTT), or identity resolution in privacy-constrained environments
  • Familiarity with the measurement and identity vendor landscape (Nielsen, Comscore, LiveRamp, The Trade Desk
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
405,710 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Warsaw
$100k – $150k per year • Remote • 6+ years exp • Bachelor's Degree
Bash
Python
DevOps
Ansible
ArgoCD
AWS
Azure
CI/CD
GCP
GitOps
Helm
Istio
Jenkins
Kubernetes
Linkerd
OpenShift
Red Hat
Service Mesh
Tekton
Terraform
Cybersecurity
HIPAA
PCI DSS
SOC 2
Apply
$145k – $205k per year • Remote • 6+ years exp • PhD
C++
Rust
SQL
Rust
Actix Web
Axum
Databases
Apache Kafka
DynamoDB
Kafka
MySQL
NATS
PostgreSQL
RabbitMQ
Redis
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
Grafana
gRPC
Istio
Kubernetes
Linkerd
OpenTelemetry
Prometheus
Pulumi
Rest API
Service Mesh
Terraform
Apply
In office • 8+ years exp • PhD
Python
DevOps
AWS
Azure
Dynatrace
GCP
Kubernetes
Self-Healing
Splunk
Apply
$120k – $170k per year • Remote • 6+ years exp • Bachelor's Degree
PowerShell
Python
ABAP
ABAP
SAP BTP
Databases
SAP HANA
DevOps
Ansible
AWS
Azure
GCP
Platform Engineering
Terraform
Apply
$120k – $170k per year • Remote • 6+ years exp • Bachelor's Degree
JavaScript
SQL
TypeScript
Java
Java
Spring Boot
Databases
Apache Kafka
Kafka
MySQL
PostgreSQL
RabbitMQ
Frontend
Angular
React.js
Vue.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
Kubernetes
Apply
$156k – $184k per year • Remote/Hybrid • Internship • 12+ years exp • Bachelor's Degree • Warsaw
Databases
Apache Kafka
BigQuery
Databricks
Google BigQuery
Kafka
Snowflake
AI/ML
Amazon SageMaker
Embeddings
Flink
Kubeflow
MLFlow
Multimodal AI
Semantic Search
Semantic Search
DevOps
Amazon EKS
CloudFormation
Google GKE
Kubernetes
Terraform
Vector
AWS
GCP
Cybersecurity
GDPR
Apply
$67k – $102k per year • Remote/Hybrid • Internship • 5+ years exp • Bachelor's Degree • Warsaw
Python
Python
pySpark
Databases
Amazon Neptune
Databricks
GraphDB
Milvus
Pinecone
Weaviate
AI/ML
GNN
GraphRAG
Knowledge Graph
LangChain
LlamaIndex
LLM
Spark
DevOps
Vector
Apply
Data Scientist 13 days ago
$49k – $138k per year (Estimated) • In office • Internship • 2+ years exp • Bachelor's Degree • Amsterdam
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Claude
LLM
RAG
Semantic Search
Semantic Search
Spark
DevOps
AWS
GCP
Vector
Apply
$29k – $41k per year • Remote/Hybrid • Internship • Bachelor's Degree • Porto
C#
C++
Java
PowerShell
Python
DevOps
CI/CD
Git
Apply
$52k – $114k per year (Estimated) • Remote/Hybrid • Internship • 2+ years exp • London
Apply
Video Content Creator 5 hours ago
Remote • Warsaw
Marketing
Instagram
Apply
Head of Sales 5 hours ago
$76k – $113k per year (Estimated) • Remote • 5+ years exp • Warsaw
Apply
In office • 3+ years exp • Bachelor's Degree • Warsaw
Apply
In office • 5+ years exp • Warsaw
Apply
$92k – $154k per year (Estimated) • In office • 2+ years exp • Bachelor's Degree • Warsaw
DevOps
Incident Management
SLI/SLO/SLA
Apply
See all jobs
This is one of many
405,710 more open roles from verified company boards, updated every day.