368,634open jobs
9,437companies
50,578added this week
Browse all
Location
Remote (Colombia)
Overview
Company
Impact
Profile match
Ciandt is a leading provider of digital transformation services. We help organizations transform their businesses by leveraging the latest technologies. Our solutions provide comprehensive digital strategy, design, development, and implementation capabilities.

As CI&T continues to expand its data and analytics capabilities, we are seeking a talented and experienced Data Developer to join our team and drive the evolution of modern data platforms for our clients. This role is critical in designing, building, and optimizing scalable data pipelines and lake architectures that empower data-driven decision-making across the organization.

The Data Developer will work with cloud-native solutions to support the entire data lifecycle-from ingestion and transformation to storage optimization and analytics enablement. This position requires strong technical expertise in distributed data processing, deep SQL proficiency, and a solid understanding of cloud infrastructure, particularly within the AWS ecosystem. The ideal candidate will balance performance, cost, and maintainability while contributing to reusable, well-architected data solutions.

Responsibilities:

Data Pipeline Development & Optimization:

  • Design, build, and maintain robust ETL/ELT processes to ingest, transform, and deliver data across a modern Data Lake architecture

  • Develop and optimize distributed data processing workflows using Python and PySpark to handle large-scale datasets efficiently

  • Implement and refine partitioning strategies for data lake storage frameworks (such as Delta Lake or Apache Iceberg) to balance query performance with storage costs

Data Transformation & Modeling:

  • Write, optimize, and translate complex SQL queries involving CTEs, window functions, conditional expressions, and aggregations

  • Migrate and modernize data pipelines from legacy RDBMS platforms to cloud-native analytics environments

  • Leverage object-oriented programming principles to contribute to in-house libraries for code reusability and standardization

Cloud Infrastructure & Orchestration:

  • Work confidently with AWS-native services including Glue (Jobs, Catalog, Triggers, Workflows), Athena, Redshift, S3, Lambda, EventBridge, and related data services

  • Collaborate with infrastructure and DevOps teams to provision and manage data resources using Infrastructure as Code (IaC) tools such as CloudFormation, CDK, or Terraform

  • Monitor data pipeline health and performance using CloudWatch and other observability tools, proactively addressing issues and improving reliability

Data Governance & Quality:

  • Ensure data integrity, consistency, and compliance across pipelines and storage layers

  • Implement metric tracking and observability frameworks to provide transparency into data workflows and SLAs

  • Support data catalog management and metadata governance practices

Collaboration & Continuous Improvement:

  • Partner with data analysts, scientists, and business stakeholders to understand requirements and translate them into scalable technical solutions

  • Contribute to technical documentation, code reviews, and knowledge sharing within the team

  • Stay current with emerging data engineering practices, tools, and cloud-native innovations

Requirements:

  • Solid experience working with ETL processes and data pipeline development with AWS

  • Strong proficiency in Python as the primary programming language, with demonstrated experience writing and optimizing PySpark code for distributed data processing

  • Thorough understanding of SQL, including complex queries (CTEs, window functions, aggregations, conditional expressions) and experience translating workloads from legacy RDBMS platforms

  • Hands-on experience with AWS Glue (Jobs, Catalog, Triggers, Workflows), Athena, and Redshift

  • Solid understanding of Data Lake architectures and partitioning strategies to optimize performance and cost

  • Good understanding of object-oriented programming (OOP) principles and experience working with reusable code libraries

  • Comfortable working with Git, Shell scripts, and Linux environments

  • Familiarity with observability, monitoring, and metric tracking practices

  • English Advanced/Fluent

Nice to Have:

  • Experience with open table formats such as Delta Lake or Apache Iceberg

  • Knowledge of TypeScript

  • Practical experience with Infrastructure as Code (IaC) tools such as CloudFormation (preferred), CDK, or Terraform

  • Familiarity with additional AWS services such as SageMaker AI, ECS, RDS, DynamoDB, IAM, or EventBridge

  • Experience working with Pandas for data manipulation and analysis

  • Exposure to machine learning workflows or AI-driven data initiatives

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$18k – $46k per year (Estimated) • Remote • Full-Time • Tula
C#
JavaScript
Node JS
SQL
TypeScript
C#
.NET
Node JS
InversifyJS
Databases
DynamoDB
MySQL
AI/ML
Claude
Copilot
Cursor
OpenAI Codex
Frontend
Angular
React.js
Tailwind CSS
Mobile
Dependency Injection
DevOps
AWS
AWS Lambda
CI/CD
OpenTelemetry
Rest API
Terraform
Amazon CloudWatch
Amazon S3
API Gateway
GitHub
Cybersecurity
HIPAA
Apply
$170k – $318k per year (Estimated) • Remote/Hybrid • Full-Time • 12+ years exp • Associate's Degree • Chicago • Milwaukee • Dallas • Columbus • Kirkland
JavaScript
Python
TypeScript
Python
pySpark
AI/ML
Prompt Engineering
Spark
DevOps
AWS
Azure
CI/CD
GCP
Git
Jenkins
GitHub
GitLab
Analytics
ETL/ELT
Apply
In office • Full-Time • 3+ years exp • Indonesia
DevOps
AWS
Azure
GCP
IAM
Apply
$19k – $53k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Pune
Bash
JavaScript
Python
TypeScript
Frontend
Angular
React.js
DevOps
AWS
Azure
Datadog
Docker
GCP
Grafana
Kubernetes
Prometheus
Splunk
IAM
Cybersecurity
Keycloak
Apply
SDET 1 day ago
$12k – $39k per year (Estimated) • In office • 4+ years exp • Gurgaon
Java
DevOps
AWS
Azure
CI/CD
Jenkins
QA
Appium
JMeter
Playwright
Rest-Assured
Selenium
Apply
$39k – $111k per year (Estimated) • Remote
JavaScript
SQL
Java
Java
Hibernate
Spring Boot
Spring Security
Databases
Apache Kafka
MySQL
Frontend
React.js
Storybook
styled-components
DevOps
Kibana
Apply
$67k – $140k per year (Estimated) • Remote
C#
SQL
C#
ASP.NET Core
Entity Framework Core
MediatR
Databases
Apache Kafka
Frontend
GraphQL
Mobile
Dependency Injection
DevOps
Azure
CI/CD
Datadog
Docker
Grafana
Kubernetes
Apply
$30k – $101k per year (Estimated) • Remote
Go
Java
Kotlin
Node JS
Python
TypeScript
JavaScript
Java
Spring Boot
Node JS
Nest.JS
Databases
Apache Kafka
DynamoDB
ElasticSearch
PostgreSQL
Redis
Frontend
GraphQL
DevOps
AWS
CI/CD
Datadog
Docker
Helm
Kong
Kubernetes
Service Mesh
API Gateway
IAM
Apply
$38k – $107k per year (Estimated) • Remote
C#
JavaScript
TypeScript
C#
.NET
AutoMapper
Entity Framework Core
Databases
PostgreSQL
Frontend
React.js
Mobile
Clean Architecture
DevOps
Azure
Azure DevOps
CI/CD
Docker
Git
Gitflow
Apply
$39k – $112k per year (Estimated) • Remote
C#
SQL
C#
.NET
Databases
Apache Kafka
PostgreSQL
RabbitMQ
AI/ML
Copilot
Anthropic
OpenAI
DevOps
Datadog
Docker
Grafana
GitHub
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.