{"id":930008,"url":"https://alion.io/job/jobgether-senior-data-engineer-12","title":"Senior Data Engineer","company":{"id":3336,"name":"Jobgether","domain":"jobgether.com","url":"https://alion.io/company/jobgether","size_band":"11-50","is_staffing_agency":true,"employer_type":"agency","is_intermediary":true,"listed_via":null,"ats_vendor":"Lever","truth_index":null},"role":"Data Science","role_family":"Data Science","seniority":"senior","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"structured","locations":[],"countries":[],"hiring_countries":["BR"],"hiring_countries_total":1,"salary":null,"salary_estimate":{"min_usd":45000,"max_usd":119000,"period":"year","method":"global_role_cell_scaled_by_country","sample_n":1536},"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Anomaly Detection","optional":false},{"name":"AWS","optional":false},{"name":"CI/CD","optional":false},{"name":"Docker","optional":false},{"name":"ETL/ELT","optional":false},{"name":"GDPR","optional":false},{"name":"Kubernetes","optional":false},{"name":"LLM","optional":false},{"name":"Machine Learning","optional":false},{"name":"MySQL","optional":false},{"name":"PostgreSQL","optional":false},{"name":"Python","optional":false},{"name":"Recommender Systems","optional":false},{"name":"SQL","optional":false}],"status":"live","first_seen_at":"2026-09-15T09:39:34Z","employer_posted_date":"2026-09-15","last_verified_at":"2026-09-18T12:30:15Z","board_verified":false,"closed_at":null,"days_open":16,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":16},"description":"About Jobgether\nJobgether is building the job search platform for experienced professionals navigating the global remote job market.\nWe aggregate hundreds of thousands of opportunities from companies around the world, structure and enrich this data, and combine it with user profiles and preferences to help people identify the opportunities and companies where they have the strongest fit.\nData is at the core of our product. The quality of our job inventory, user data, enrichment systems and matching models directly determines the quality of the experience we provide.\nWe are a fully remote, international team focused on building a smarter and more effective way to search for work.\nWhat you will own\nAs our Senior Data Engineer, you will own the data infrastructure behind Jobgether's job marketplace and matching engine.\nYour responsibility will span the complete data lifecycle: from discovering and collecting jobs, to cleaning and enriching them, capturing high-quality user data, and making sure both sides of the marketplace can be matched accurately.\nThis is a highly hands-on role for someone who enjoys building systems, solving messy data problems and taking ownership of data quality at scale.\n1/ Job Aggregation & Ingestion\nOwn and continuously improve the systems responsible for collecting job opportunities from thousands of companies and external sources.\nYou will:\nDesign and maintain scalable job aggregation and scraping pipelines.\n\nImprove coverage, freshness and reliability of our job inventory.\n\nBuild systems to detect failed sources, missing jobs and ingestion anomalies.\n\nManage deduplication and job lifecycle management.\n\nImprove the scalability and resilience of our aggregation infrastructure.\n\nMonitor the health and performance of the entire ingestion ecosystem.\n\n2/ Data Quality & Enrichment\nRaw job data is only the starting point. You will be responsible for transforming heterogeneous job information into reliable, structured data that can power search, matching and product experiences.\nThis includes:\nNormalizing job titles, locations, companies, skills, seniority and employment information.\n\nImproving classification and taxonomy systems.\n\nDeveloping automated quality controls and anomaly detection.\n\nDesigning enrichment pipelines using deterministic systems, external data and AI/LLMs.\n\nDefining and tracking data-quality metrics across the job inventory.\n\nIdentifying systematic quality issues and building solutions rather than manual fixes.\n\n3/ User Data & Profiles\nYou will also own the data layer on the candidate side of the marketplace.\nYou will work on:\nStructuring user profiles from CVs, onboarding data, preferences and product interactions.\n\nImproving how skills, experience, seniority, job preferences and career signals are represented.\n\nDesigning reliable data models that can evolve as Jobgether collects richer career information.\n\nEnsuring user data is consistent, usable and available to our matching and personalization systems.\n\nMaintaining strong privacy and data-governance standards.\n\n4/ Matching Quality\nUltimately, our data exists to create better matches between people and opportunities.\nYou will work closely with Product and AI/ML teams to continuously improve matching quality.\nThis includes:\nBuilding the data foundations used by our matching algorithms.\n\nIdentifying missing or unreliable signals affecting match quality.\n\nCreating datasets and evaluation frameworks to measure matching performance.\n\nMonitoring match quality across roles, geographies and user segments.\n\nHelping productionize new scoring, ranking and AI/ML systems.\n\nConnecting user behavior and outcomes back into the matching system to continuously improve recommendations.\n\n5/ Data Platform & Engineering\nAcross these areas, you will:\nDesign scalable and maintainable data architectures.\n\nBuild and operate reliable ETL/ELT pipelines.\n\nImprove observability, monitoring and alerting.\n\nOptimize database performance and data access patterns.\n\nMaintain clear documentation of data models, pipelines and architecture.\n\nContribute to technical decisions around our broader data stack.\n\nTroubleshoot production issues and take ownership from diagnosis through resolution.\n\nWhat we are looking for\nWe're looking for someone with strong data-engineering fundamentals and a product mindset. Someone who's comfortable with messy, real-world data and takes ownership of the results, making sure the pipeline produces best-in-class data to power the product.\nCore Requirements\n5+ years of professional experience in Data Engineering, Backend Engineering or a closely related field.\n\nStrong professional experience with Python.\n\nDeep experience designing and operating production data pipelines.\n\nStrong SQL skills and experience with relational databases such as PostgreSQL or MySQL.\n\nExperience working with NoSQL databases such as MongoDB.\n\nStrong understanding of data modeling, schemas and large-scale data transformation.\n\nExperience with cloud infrastructure, preferably AWS.\n\nExperience designing monitoring, observability and data-quality systems.\n\nStrong understanding of APIs, web data ingestion and distributed systems.\n\nAbility to investigate complex data problems and identify their root causes.\n\nStrong written and spoken English.\n\nStrong Plus\nExperience in one or several of these areas would be particularly relevant:\nLarge-scale web scraping or data aggregation.\n\nSearch engines, marketplaces or recommendation systems.\n\nJob, talent or HR data.\n\nEntity resolution and deduplication.\n\nTaxonomies, classification and semantic data enrichment.\n\nLLM-based data extraction and enrichment.\n\nMachine Learning pipelines and MLOps.\n\nRanking or recommendation systems.\n\nDocker, Kubernetes and CI/CD environments.\n\nData privacy and GDPR-compliant architectures.\n\nThis role will suit you particularly well if:\nYou like owning a problem end-to-end rather than maintaining one small part of a system.\n\nMessy data problems interest you more than perfect datasets.\n\nYou naturally investigate why data is wrong, not only why a pipeline failed.\n\nYou think about the product impact of the infrastructure you build.\n\nYou are comfortable making architectural decisions while remaining highly hands-on.\n\nYou prefer automating recurring problems instead of creating manual processes.\n\nYou enjoy working in an environment where there is a lot to build and improve.\n\nWhy join Jobgether\nHigh ownership: You will own one of the most critical systems at Jobgether and have significant influence over its architecture.\nReal scale: Your systems will process and enrich a large, constantly changing global job inventory and millions of candidate data points.\nDirect product impact: Improvements in your work translate directly into better job discovery, matching and recommendations for our users.\nTechnical challenges: Aggregation, entity resolution, classification, enrichment, search and matching create genuinely complex data problems to solve.\nRemote by design: Jobgether is a fully remote company with an international team and a culture focused on accountability and outcomes.","description_format":"text","description_chars":7112,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[{"language":"English","level":"All levels","optional":false}]},"benefits":[],"hiring_locations":[{"name":"Brazil","iso":"BR","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-09-15T10:48:36Z"}],"liveness":{"score":3,"band":"cold","label":"Long shot","p_open":0.7,"p_active":0.137,"p_room":0.35,"age_days":15,"expected_fill_days":7,"reasons":["conf:305","agency","wave","velocity","win:tail"],"computed_at":"2026-10-01T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/jobgether-senior-data-engineer-12","json_url":"https://alion.io/job/jobgether-senior-data-engineer-12.json","meta":{"generated_at":"2026-10-01T18:05:15Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":324,"day_limit":5000,"remaining_today":4676,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}