819,440open jobs
52,731companies
133,157added this week
Browse all
Salary
≈ $47k – $114k per year (Estimated)
Location
Remote (LATAM)
Seniority
Senior · 5+ years exp
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 26, 2026. First seen by Alion on Sep 23, 2026. Ryzlabs scores B on the Alion truth index.

Overview
Company
Impact
Profile match
Unlock the power of LatAm's elite nearshore talent with Ryz Labs. Our top-tier staff augmentation services provide access to the brightest minds in software development, IT, sales, ops, and CS. Elevate your team and achieve unparalleled results with our nearshore experts.

Remote position - only for professionals based in LATAM.

We are looking for a Senior Data Engineer for one of our client's teams. You will take ownership of the systems responsible for ingesting, transforming, validating, and publishing data across our platform.

In this role, you will work at the data ingestion boundary, where data from multiple external sources enters our systems. You will be responsible for building reliable pipelines, resolving identity and entity-matching challenges, detecting data-quality issues, and ensuring data flows correctly into downstream products and services.

This is a hands-on engineering role that combines AWS data engineering, Python/SQL development, event-driven architectures, data quality, and identity resolution. You will collaborate closely with engineering, product, analytics, and downstream platform teams to troubleshoot issues and improve the reliability of our data ecosystem.

What You'll Do

  • Own and evolve data ingestion pipelines, including source ingestion/scraping, data cleaning and curation using AWS Glue, entity resolution, and event-driven data publishing.
  • Design and maintain reliable batch and event-driven data workflows across our data platform.
  • Diagnose and resolve identity-resolution issues across multiple data sources, including deduplication and entity-matching challenges.
  • Develop and maintain data quality and validation frameworks, including freshness monitoring, null-rate checks, consistency validation, and schema-drift detection.
  • Monitor data at the ingestion boundary and proactively identify issues before they impact downstream systems and products.
  • Partner with engineering teams responsible for the events bridge and downstream identity services to trace data and events end-to-end.
  • Collaborate with Product, Assessments, Analytics, and other stakeholders to understand how data flows into downstream products and ensure those requirements are reflected in the data pipelines.
  • Automate infrastructure and pipeline changes using Infrastructure as Code, primarily Terraform or AWS CDK.
  • Work with AWS services including Glue, Athena, S3, and event-driven services to build and operate scalable data infrastructure.
  • Participate in production incident response, quickly identifying the scope and impact of data-quality issues and implementing remediation.
  • Continuously improve pipeline reliability, observability, maintainability, and operational processes.

What You'll Bring

  • 5+ years of experience building and operating production-grade data pipelines.
  • Strong experience working with both batch data processing and event-driven architectures.
  • Hands-on experience with AWS Glue, Athena, and S3-based data lakes, including layered/medallion-style data transformations.
  • Strong proficiency in Python and SQL for data transformation, processing, and analysis.
  • Experience with identity resolution, entity matching, or deduplication, including exact and fuzzy matching approaches.
  • Understanding of the challenges and tradeoffs involved in maintaining durable identifiers and first-seen/locked identity mappings.
  • Experience working with event schemas and schema-registry-backed contracts, such as Protobuf, and an understanding of the risks associated with schema changes.
  • Strong troubleshooting and production incident-response skills, including the ability to determine impact, identify root causes, and implement fixes quickly.
  • Ability to understand and debug code written in a functional or concurrent programming language, such as Elixir.
  • Strong communication and collaboration skills, with the ability to work effectively across engineering, product, analytics, and other technical teams.

Nice to Have

  • Experience with Elixir/Phoenix or another BEAM-based concurrent processing framework.
  • Experience with DynamoDB-backed identity, lookup, or matching services.
  • Experience with the Snowflake ecosystem, including data modeling, Snowpipe, Streams, and Tasks.
  • Experience working with sports data providers, such as MLB.com or Sportradar, or with other licensed-content/data provider ecosystems.
  • Experience implementing data observability, including freshness/staleness alerts, null-rate monitoring, schema-drift detection, and data-quality dashboards.
  • Experience working with Infrastructure as Code using Terraform or AWS CDK.

Technologies

Languages: Python, SQL, familiarity with Elixir

AWS: Glue, Athena, S3, DynamoDB

Data & Streaming: Data Lakes, Event-Driven Architecture, Protobuf, Schema Registries

Infrastructure: Terraform, AWS CDK

Data Platforms: Snowflake

Engineering Practices: Data Quality, Data Observability, Identity Resolution, Entity Matching, Incident Response

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
819,440 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
In your city
Data Engineer 3 days ago
≈ $77k – $172k per year (Estimated) • Remote (Sweden) • Full-Time • 4+ years exp
Python
Python
pySpark
AI/ML
Spark
Machine Learning
DevOps
AWS
Docker
Amazon ECS
Apply
Data Engineer LATAM 10 hours ago
≈ $34k – $85k per year (Estimated) • Remote (Mexico) • Full-Time • 3+ years exp • Mexico City
Python
Java
SQL
Scala
Databases
ClickHouse
Apache Iceberg
AI/ML
Copilot
Cursor
Spark
DevOps
CI/CD
AWS
Docker
Kubernetes
Amazon S3
Analytics
Tableau
ETL/ELT
Metabase
Looker
Superset
Apply
Data Engineer LATAM 10 hours ago
≈ $37k – $94k per year (Estimated) • Remote (Brazil) • Full-Time • 3+ years exp • Rio de Janeiro
Python
Java
SQL
Scala
Databases
ClickHouse
Apache Iceberg
AI/ML
Copilot
Cursor
Spark
DevOps
CI/CD
AWS
Docker
Kubernetes
Amazon S3
Analytics
Tableau
ETL/ELT
Metabase
Looker
Superset
Apply
Data Engineer LATAM 10 hours ago
≈ $73k – $183k per year (Estimated) • Remote (Colombia) • Full-Time • 3+ years exp
Python
Java
SQL
Scala
Databases
ClickHouse
Apache Iceberg
AI/ML
Copilot
Cursor
Spark
DevOps
CI/CD
AWS
Docker
Kubernetes
Amazon S3
Analytics
Tableau
ETL/ELT
Metabase
Looker
Superset
Apply
Data Scientist 4 hours ago
≈ $78k – $174k per year (Estimated) • Remote (United States, US time zones hours) • 4+ years exp
Python
AI/ML
Claude
Prompt Engineering
Function Calling
AI Agents
Gemini
LLM
OpenAI
Anthropic
LLM Guardrails
Agentic Workflows
Machine Learning
DevOps
Azure
Apply
≈ $19k – $48k per year (Estimated) • In office • Moscow
Python
SQL
AI/ML
NumPy
DevOps
Prometheus
Grafana
Analytics
Tableau
Power BI
Seaborn
Matplotlib
ETL/ELT
Looker
Management
Confluence
Jira
Apply
≈ $28k – $58k per year (Estimated) • In office • 5+ years exp • Bengaluru
Python
SQL
Databases
Snowflake
AI/ML
dbt
Analytics
ETL/ELT
Apply
In office • 15+ years exp • Bachelor's Degree • Hyderabad
Python
Go
JavaScript
Java
SQL
Ruby
Node JS
Python
Flask
Django
Java
Spring Boot
DevOps
GCP
CircleCI
Azure
CI/CD
Jenkins
AWS
Docker
Kubernetes
AWS Lambda
Amazon EC2
GitLab
Amazon S3
Management
Trello
Jira
Agile
Scrum
Kanban
Apply
$87k – $165k per year • Hybrid • Secret • Full-Time • 5+ years exp • Lowell
Python
JavaScript
TypeScript
SQL
C#
Perl
SAS
Databases
MS SQL
MariaDB
Frontend
Angular
React.js
DevOps
Rest API
CI/CD
Management
Agile
Apply
≈ $19k – $48k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • Moscow
Python
SQL
Apply
≈ $22k – $65k per year (Estimated) • Remote (Argentina) • Full-Time • 3+ years exp • Bachelor's Degree
Management
QuickBooks
Microsoft Office
Apply
Asset Titler 3 days ago
≈ $16k – $42k per year (Estimated) • Remote (Argentina) • Full-Time
Apply
≈ $55k – $125k per year (Estimated) • Remote (Argentina) • Full-Time • 5+ years exp • Bachelor's Degree
Python
SQL
AI/ML
Model Context Protocol
Management
ServiceNow
Apply
Webflow Developer 12 days ago
≈ $32k – $90k per year (Estimated) • Remote (Argentina) • Part-Time • 3+ years exp • Buenos Aires
JavaScript
DevOps
AWS
Design
Webflow
Apply
Sr QA Engineer 15 days ago
≈ $37k – $92k per year (Estimated) • Remote (Argentina) • Full-Time • Buenos Aires
JavaScript
Java
SQL
Java
Spring Boot
Databases
PostgreSQL
Frontend
Next.js
React.js
Mobile
React Native
Maestro
DevOps
Rest API
GitHub Actions
Datadog
CI/CD
QA
Playwright
Postman
Apply
See all jobs
This is one of many
819,440 more open roles from verified company boards, updated every day.