658,057open jobs
38,313companies
94,964added this week
Browse all
Salary
$87k – $212k per year (Estimated)
Location
Remote (Colombia)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Jobgether is a Belgian recruitment platform built entirely around remote and flexible work, aggregating openings from thousands of employers that allow work from outside an office. Its matching engine ranks roles against a candidate's skills, seniority and stated preferences on location and flexibility, rather than leaving people to filter a keyword search, and it verifies how genuinely remote each posting is. The company also runs an AI screening layer that shortlists applicants for employers, and publishes research and guidance on distributed work practices alongside the job marketplace itself.

About Jobgether

Jobgether is building the job search platform for experienced professionals navigating the global remote job market.

We aggregate hundreds of thousands of opportunities from companies around the world, structure and enrich this data, and combine it with user profiles and preferences to help people identify the opportunities and companies where they have the strongest fit.

Data is at the core of our product. The quality of our job inventory, user data, enrichment systems and matching models directly determines the quality of the experience we provide.

We are a fully remote, international team focused on building a smarter and more effective way to search for work.

What you will own

    As our Senior Data Engineer, you will own the data infrastructure behind Jobgether's job marketplace and matching engine.

    Your responsibility will span the complete data lifecycle: from discovering and collecting jobs, to cleaning and enriching them, capturing high-quality user data, and making sure both sides of the marketplace can be matched accurately.

    This is a highly hands-on role for someone who enjoys building systems, solving messy data problems and taking ownership of data quality at scale.

    1/ Job Aggregation & Ingestion

    Own and continuously improve the systems responsible for collecting job opportunities from thousands of companies and external sources.

    You will:

    • Design and maintain scalable job aggregation and scraping pipelines.

    • Improve coverage, freshness and reliability of our job inventory.

    • Build systems to detect failed sources, missing jobs and ingestion anomalies.

    • Manage deduplication and job lifecycle management.

    • Improve the scalability and resilience of our aggregation infrastructure.

    • Monitor the health and performance of the entire ingestion ecosystem.

    • 2/ Data Quality & Enrichment

      Raw job data is only the starting point. You will be responsible for transforming heterogeneous job information into reliable, structured data that can power search, matching and product experiences.

      This includes:

      • Normalizing job titles, locations, companies, skills, seniority and employment information.

      • Improving classification and taxonomy systems.

      • Developing automated quality controls and anomaly detection.

      • Designing enrichment pipelines using deterministic systems, external data and AI/LLMs.

      • Defining and tracking data-quality metrics across the job inventory.

      • Identifying systematic quality issues and building solutions rather than manual fixes.

      • 3/ User Data & Profiles

        You will also own the data layer on the candidate side of the marketplace.

        You will work on:

        • Structuring user profiles from CVs, onboarding data, preferences and product interactions.

        • Improving how skills, experience, seniority, job preferences and career signals are represented.

        • Designing reliable data models that can evolve as Jobgether collects richer career information.

        • Ensuring user data is consistent, usable and available to our matching and personalization systems.

        • Maintaining strong privacy and data-governance standards.

        • 4/ Matching Quality

          Ultimately, our data exists to create better matches between people and opportunities.

          You will work closely with Product and AI/ML teams to continuously improve matching quality.

          This includes:

          • Building the data foundations used by our matching algorithms.

          • Identifying missing or unreliable signals affecting match quality.

          • Creating datasets and evaluation frameworks to measure matching performance.

          • Monitoring match quality across roles, geographies and user segments.

          • Helping productionize new scoring, ranking and AI/ML systems.

          • Connecting user behavior and outcomes back into the matching system to continuously improve recommendations.

          • 5/ Data Platform & Engineering

            Across these areas, you will:

            • Design scalable and maintainable data architectures.

            • Build and operate reliable ETL/ELT pipelines.

            • Improve observability, monitoring and alerting.

            • Optimize database performance and data access patterns.

            • Maintain clear documentation of data models, pipelines and architecture.

            • Contribute to technical decisions around our broader data stack.

            • Troubleshoot production issues and take ownership from diagnosis through resolution.

            • What we are looking for

              We're looking for someone with strong data-engineering fundamentals and a product mindset. Someone who's comfortable with messy, real-world data and takes ownership of the results, making sure the pipeline produces best-in-class data to power the product.

              Core Requirements

              • 5+ years of professional experience in Data Engineering, Backend Engineering or a closely related field.

              • Strong professional experience with Python.

              • Deep experience designing and operating production data pipelines.

              • Strong SQL skills and experience with relational databases such as PostgreSQL or MySQL.

              • Experience working with NoSQL databases such as MongoDB.

              • Strong understanding of data modeling, schemas and large-scale data transformation.

              • Experience with cloud infrastructure, preferably AWS.

              • Experience designing monitoring, observability and data-quality systems.

              • Strong understanding of APIs, web data ingestion and distributed systems.

              • Ability to investigate complex data problems and identify their root causes.

              • Strong written and spoken English.

              • Strong Plus

                Experience in one or several of these areas would be particularly relevant:

                • Large-scale web scraping or data aggregation.

                • Search engines, marketplaces or recommendation systems.

                • Job, talent or HR data.

                • Entity resolution and deduplication.

                • Taxonomies, classification and semantic data enrichment.

                • LLM-based data extraction and enrichment.

                • Machine Learning pipelines and MLOps.

                • Ranking or recommendation systems.

                • Docker, Kubernetes and CI/CD environments.

                • Data privacy and GDPR-compliant architectures.

                • This role will suit you particularly well if:

                  • You like owning a problem end-to-end rather than maintaining one small part of a system.

                  • Messy data problems interest you more than perfect datasets.

                  • You naturally investigate why data is wrong, not only why a pipeline failed.

                  • You think about the product impact of the infrastructure you build.

                  • You are comfortable making architectural decisions while remaining highly hands-on.

                  • You prefer automating recurring problems instead of creating manual processes.

                  • You enjoy working in an environment where there is a lot to build and improve.

                  • Why join Jobgether

                    High ownership: You will own one of the most critical systems at Jobgether and have significant influence over its architecture.

                    Real scale: Your systems will process and enrich a large, constantly changing global job inventory and millions of candidate data points.

                    Direct product impact: Improvements in your work translate directly into better job discovery, matching and recommendations for our users.

                    Technical challenges: Aggregation, entity resolution, classification, enrichment, search and matching create genuinely complex data problems to solve.

                    Remote by design: Jobgether is a fully remote company with an international team and a culture focused on accountability and outcomes.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
658,057 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$22k – $50k per year (Estimated) • Remote • Full-Time • 10+ years exp • Bachelor's Degree • Bengaluru
Python
Java
SQL
Scala
Databases
Snowflake
Databricks
Oracle
Delta Lake
Apache Kafka
AI/ML
Spark
AI Agents
Flink
Edge AI
DevOps
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Management
Agile
Scrum
Apply
$18k – $43k per year (Estimated) • Remote • Voronezh
Java
SQL
Databases
PostgreSQL
Apache Kafka
DevOps
gRPC
OpenShift
CI/CD
Management
Agile
Scrum
Kanban
Apply
$16k – $34k per year (Estimated) • Remote • Full-Time • 4+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Databases
Snowflake
Databricks
AI/ML
Spark
Airflow
AI Agents
Edge AI
DevOps
Azure
Git
AWS
Amazon S3
Analytics
ETL/ELT
Azure Data Factory
Dimensional Modeling
Management
Agile
Apply
$50k – $148k per year (Estimated) • In office • Full-Time • 8+ years exp • Leeds
JavaScript
Kotlin
SQL
C#
Dart
Swift
C#
.NET
Databases
Snowflake
MS SQL
Frontend
React.js
JQuery
Mobile
Flutter
React Native
Kotlin Multiplatform
DevOps
Azure DevOps
New Relic
PagerDuty
Azure
CI/CD
AWS
Kubernetes
Amazon EKS
Amazon EC2
Incident Management
Cybersecurity
Snyk
Auth0
Analytics
Fivetran
Management
Agile
Apply
$15k – $36k per year (Estimated) • In office • Full-Time • 5+ years exp • Ipoh
SQL
Databases
MS SQL
Analytics
SSIS
SSAS
Apply
$100k – $125k per year • Equity • Remote • Full-Time • 2+ years exp
Python
JavaScript
PHP
Ruby
PowerShell
Bash
DevOps
GCP
Azure
AWS
Cybersecurity
MITRE ATT&CK
OWASP Top 10
Apply
$121k – $221k per year (Estimated) • Equity • In office • Full-Time • 7+ years exp
Design
Figma
ProtoPie
Framer
Apply
$106k – $133k per year • Remote • Full-Time • 10+ years exp • Bachelor's Degree
Apply
$117k – $219k per year • Remote • Full-Time • 5+ years exp
JavaScript
TypeScript
SQL
Ruby
C#
Ruby
Ruby on Rails
C#
.NET
Databases
PostgreSQL
Frontend
GraphQL
RxJS
Angular
NgRx
DevOps
CI/CD
Git
AWS
Docker
Kubernetes
AWS Lambda
Amazon EC2
Amazon S3
Amazon ECS
Analytics
Power BI
ETL/ELT
SSIS
AWS Glue
Management
Jira
Agile
Scrum
Apply
$48k – $116k per year (Estimated) • Remote • Full-Time • 5+ years exp
Python
SQL
Databases
MySQL
PostgreSQL
AI/ML
LLM
Anomaly Detection
Recommender Systems
DevOps
CI/CD
AWS
Docker
Kubernetes
Cybersecurity
GDPR
Analytics
ETL/ELT
Apply
See all jobs
This is one of many
658,057 more open roles from verified company boards, updated every day.