658,057open jobs
38,313companies
94,964added this week
Browse all
Salary
$44k – $107k per year (Estimated)
Location
Remote (Brazil)
Seniority
Senior
Overview
Company
Impact
Profile match
CI&T is a Brazilian digital services company founded in Campinas in 1995 that builds software, data platforms and artificial intelligence systems for large enterprises. It works mainly with banks, retailers, consumer goods companies and healthcare providers in Brazil, the United States and Europe, taking on product engineering, cloud modernisation and experience design as long-running delivery engagements. The company listed on the New York Stock Exchange in 2021 and employs several thousand people across delivery centres in Latin America, North America, Europe and Asia.

Nosso cliente está conduzindo uma modernização massiva de seus pipelines de dados, migrando cargas de trabalho legadas em Databricks para uma nova arquitetura nativa no Google Cloud. Esta posição é central para essa transformação: o profissional atuará como um dos pilares hands-on da refatoração, transformando lógicas complexas de processamento distribuído em padrões modernos de ELT, sem comprometer a continuidade das integrações já existentes.

O Data Developer Senior atua na conversão direta de notebooks legados, na construção de pipelines de ingestão escaláveis e na garantia de que a arquitetura em camadas (Raw, Trusted Core e Gold) preserve os contratos de dados que sustentam sistemas e dashboards em produção. O papel combina profundidade técnica na reengenharia de código acoplado com colaboração próxima aos times de negócio para validar cada etapa da migração.

Responsibilities:

  • Code Refactoring: Translates legacy Databricks notebook logic into modern ELT patterns, prioritizing declarative SQL-based transformation for standard cases and developing Python-based distributed processing routines for intricate logic (such as RDD-based operations).
  • Data Contract Preservation: Ensures backward compatibility during migration by implementing trusted-core layering and reverse-view strategies, preventing breaking changes for downstream consuming systems and dashboards.
  • Ingestion Pipeline Development: Builds scalable ingestion pipelines by parameterizing YAML configurations that feed an automated DAG generation pipeline, orchestrated through a workflow scheduler and consuming standardized processing templates.
  • Data Quality Implementation: Applies synchronous and unit-level assertions within the transformation layers to prevent null values, enforce key uniqueness, and validate domain conformity.
  • Migration Validation: Executes technical validation projects by reconciling data output between the legacy environment and the new platform, confirming record-level and field-level parity ("cara-a-crachá" checks).
  • GitOps Development: Follows strict CI/CD workflows through feature branches, pull requests, and automated deployment pipelines to development and production environments.
  • Cross-Team Alignment: Collaborates consultatively with business Data Stewards to align integration dependencies, negotiate refactoring scope, and validate migrated outputs side by side with stakeholders.

Requirements:

  • Strong, hands-on technical depth in Python, PySpark, and advanced SQL for large-scale distributed data processing
  • Robust experience migrating data lakes and refactoring legacy code, including reverse-engineering tightly coupled logic (such as Databricks or Azure Data Factory pipelines) into modular cloud-native components
  • Practical mastery of the Google Cloud data engineering ecosystem, particularly BigQuery, Dataform, and Dataproc Serverless
  • Solid understanding of Medallion Architecture (Raw/Bronze, Silver/Trusted Core, Gold) and analytical data modeling strategies, including Star Schema and One Big Table approaches
  • Hands-on experience with modern GitOps development flows (code review, pull request submission) and orchestration automation using Airflow/Composer
  • Maturity in data quality practices, with testing applied directly to transformation pipelines
  • Consultative, collaborative communication style to align cross-team dependencies, negotiate refactoring scope, and validate outcomes alongside business Data Stewards

Nice to Have:

  • Experience applying Generative AI agents (such as Gemini, Vertex AI, or Claude) to automate PySpark-to-SQL code refactoring, dependency graph analysis, or automated documentation generation
  • Prior experience with Change Data Capture (CDC) architectures and event-driven ingestion patterns
  • Strong performance tuning skills in BigQuery, including query optimization, partitioning strategies, and execution cost control
  • Deep knowledge of the boundaries between Data Vault and Star Schema modeling approaches for trusted/silver layers
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
658,057 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
Decision Scientist II 5 hours ago
$54k – $118k per year (Estimated) • Equity • In office • Full-Time • 6+ years exp • Bachelor's Degree • Charlotte • Atlanta • Winston-Salem
Python
SQL
SAS
Databases
Db2
Oracle
Analytics
Tableau
MicroStrategy
Apply
$104k – $202k per year • In office • Full-Time • 3+ years exp • Bachelor's Degree • United States
Python
JavaScript
Rust
TypeScript
C++
DevOps
CI/CD
Platform Engineering
Incident Management
Apply
(USA) Data Analyst II 5 hours ago
$59k – $121k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • United States
Python
SQL
Databases
Google BigQuery
BigQuery
AI/ML
Copilot
AI Agents
Anomaly Detection
QA
Playwright
Apply
(USA) Data Analyst II 5 hours ago
$59k – $123k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Bentonville
Python
SQL
Scala
Databases
Google BigQuery
BigQuery
AI/ML
Spark
Analytics
Tableau
Power BI
Apply
$25k – $56k per year (Estimated) • In office • Full-Time • 9+ years exp • Chennai
Python
Kotlin
MATLAB
Kotlin
StateFlow
MATLAB
Simulink
DevOps
Git
Apply
$38k – $91k per year (Estimated) • Remote
Python
SQL
Python
FastAPI
Databases
PostgreSQL
Snowflake
AI/ML
dbt
Pandas
DevOps
Terraform
GitHub Actions
CloudFormation
Prometheus
GitLab CI
CI/CD
Git
AWS
Kubernetes
Grafana
Amazon EKS
AWS Lambda
SLI/SLO/SLA
Amazon CloudWatch
Analytics
ETL/ELT
A/B Testing
Apply
Remote
JavaScript
Java
TypeScript
SQL
AI/ML
AI Agents
Frontend
Angular
Mobile
JUnit
DevOps
AWS
Apply
$71k – $144k per year (Estimated) • Remote
DevOps
GitLab CI
CI/CD
AWS
Apply
$50k – $124k per year (Estimated) • Remote • PhD
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Spark
MLFlow
MLlib
RAG
LLMOps
Feature Store
DevOps
Azure DevOps
Azure
CI/CD
Analytics
ETL/ELT
Apply
$65k – $118k per year (Estimated) • Remote/Hybrid • Full-Time • London
Management
Agile
Apply
See all jobs
This is one of many
658,057 more open roles from verified company boards, updated every day.