368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$15k – $58k per year (Estimated)
Location
Remote (India)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Blend360 is a global provider of data science, artificial intelligence, and enterprise technology solutions that help organizations modernize their business operations. The company offers end-to-end services, including advanced analytics strategy, custom AI model development, cloud architecture, and marketing technology deployment. Headquartered in Columbia, Maryland, Blend360 combines domain expertise with scalable tech talent to drive measurable revenue growth and digital transformation for Fortune 500 enterprises.

Blend360 is a data and AI services company specializing in data engineering, data science, MLOps, and governance to build scalable analytics solutions. It partners with enterprise and Fortune 1000 clients across industries including financial services, healthcare, retail, technology, and hospitality to drive data-driven decision making. Headquartered in Columbia, Maryland, the company is recognized for rapid growth and global delivery of AI solutions through the integration of people, data, and technology.

We are seeking a hands-on Data Engineer with deep expertise in distributed systems, ETL/ELT development, and enterprise-grade database management. The engineer will design, implement, and optimize ingestion, transformation, and storage workflows to support the MMO platform. The role requires technical fluency across big data frameworks (HDFS, Hive, PySpark), orchestration platforms (NiFi), and relational systems (Postgres), combined with strong coding skills in Python and SQL for automation, custom transformations, and operational reliability.

We are implementing a Media Mix Optimization (MMO) platform designed to analyze and optimize marketing investments across multiple channels. This initiative requires a robust on-premises data infrastructure to support distributed computing, large-scale data ingestion, and advanced analytics. The Data Engineer will be responsible for building and maintaining resilient pipelines and data systems that feed into MMO models, ensuring data quality, governance, and availability for Data Science and BI teams. The environment integrates HDFS for distributed storage, Apache NiFi for orchestration, Hive and PySpark for distributed processing, and Postgres for structured data management.

This role is central to enabling seamless integration of massive datasets from disparate sources (media, campaign, transaction, customer interaction, etc.), standardizing data, and providing reliable foundations for advanced econometric modeling and insights.

Responsibilities:

Data Pipeline Development & Orchestration

o Design, build, and optimize scalable data pipelines in Apache NiFi to

automate ingestion, cleansing, and enrichment from structured, semi-structured, and unstructured sources.

Ensure pipelines meet low-latency and high-throughput requirements for distributed processing.

Data Storage & Processing

o Architect and manage datasets on HDFS to support high-volume,

fault-tolerant storage.

o Develop distributed processing workflows in PySpark and Hive to

handle large-scale transformations, aggregations, and joins across

petabyte-level datasets.

o Implement partitioning, bucketing, and indexing strategies to

optimize query performance.

Database Engineering & Management

o Maintain and tune Postgres databases for high availability, integrity,

and performance.

o Write advanced SQL queries for ETL, analysis, and integration with

downstream BI/analytics systems.

Collaboration & Integration

o Partner with Data Scientists to deliver clean, reliable datasets for

model training and MMO analysis.

o Work with BI engineers to ensure data pipelines align with reporting

and visualization requirements.

Monitoring & Reliability Engineering

o Implement monitoring, logging, and alerting frameworks to track

data pipeline health.

o Troubleshoot and resolve issues in ingestion, transformations, and

distributed jobs.

Data Governance & Compliance

o Enforce standards for data quality, lineage, and security across

systems.

o Ensure compliance with internal governance and external

regulations.

Documentation & Knowledge Transfer

o Develop and maintain comprehensive technical documentation for

pipelines, data models, and workflows.

o Provide knowledge sharing and onboarding support for cross-

functional teams.

  • Bachelor’s degree in Computer Science, Information Technology, or related field (Master’s preferred).

  • Proven experience as a Data Engineer with expertise in HDFS, Apache NiFi, Hive, PySpark, Postgres, Python, and SQL.

  • Strong background in ETL/ELT design, distributed processing, and relational database management.

  • Experience with on-premises big data ecosystems supporting distributed computing.

  • Solid debugging, optimization, and performance tuning skills.

  • Ability to work in agile environments, collaborating with multi-disciplinary

    teams.

  • Strong communication skills for cross-functional technical discussions.

    Preferred Qualifications:

  • Familiarity with data governance frameworks, lineage tracking, and data cataloging tools.

  • Knowledge of security standards, encryption, and access control in on- premises environments.

  • Prior experience with Media Mix Modeling (MMM/MMO) or marketing analytics projects.

  • Exposure to workflow schedulers (Airflow, Oozie, or similar).

  • Proficiency in developing automation scripts and frameworks in Python for

    CI/CD of data pipelines.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Hyderabad
$20k – $45k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Gurgaon
Python
SQL
Python
pySpark
Databases
Databricks
Microsoft Fabric
AI/ML
Hadoop
Spark
DevOps
AWS
Azure
Analytics
Power BI
Tableau
Apply
$27k – $62k per year (Estimated) • In office • Full-Time • Gurgaon
Python
SQL
AI/ML
Hadoop
Spark
Analytics
Power BI
Tableau
Apply
$19k – $45k per year (Estimated) • In office • Full-Time • 8+ years exp • Master's Degree • Gurgaon
Python
SQL
Analytics
Power BI
Tableau
A/B Testing
Apply
$18k – $46k per year (Estimated) • Remote • Full-Time • Tula
C#
JavaScript
Node JS
SQL
TypeScript
C#
.NET
Node JS
InversifyJS
Databases
DynamoDB
MySQL
AI/ML
Claude
Copilot
Cursor
OpenAI Codex
Frontend
Angular
React.js
Tailwind CSS
Mobile
Dependency Injection
DevOps
AWS
AWS Lambda
CI/CD
OpenTelemetry
Rest API
Terraform
Amazon CloudWatch
Amazon S3
API Gateway
GitHub
Cybersecurity
HIPAA
Apply
Java Developer 1 day ago
$31k – $54k per year (Estimated) • Remote • 5+ years exp • Tula
C#
JavaScript
TypeScript
Java
C#
.NET
Java
Spring Boot
Databases
Apache Kafka
ElasticSearch
PostgreSQL
RabbitMQ
Frontend
Angular
Bootstrap
React.js
DevOps
Docker
Git
Jenkins
Kubernetes
Rest API
GitLab
QA
Swagger
Apply
Snowflake Field CTO 4 days ago
$250k – $300k per year • Remote • Full-Time • 15+ years exp • New York
Databases
Snowflake
AI/ML
dbt
Embeddings
Feature Store
DevOps
AWS
Apply
$27k – $53k per year (Estimated) • In office • Full-Time • 3+ years exp • Hyderabad
Python
SQL
Python
pySpark
Databases
Databricks
AI/ML
Spark
DevOps
AWS
Azure
Analytics
A/B Testing
Apply
$20k – $46k per year (Estimated) • In office • Full-Time • 5+ years exp • Hyderabad
Python
SQL
AI/ML
NumPy
Pandas
Analytics
Power BI
Tableau
A/B Testing
Marketing
Marketo
Salesforce
Apply
$41k – $105k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Guadalajara
Databases
Databricks
Snowflake
DevOps
AWS
GitHub
Apply
$27k – $53k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Master's Degree • Hyderabad
Python
SQL
AI/ML
NumPy
Scikit-learn
SciPy
SHAP
Amazon SageMaker
DevOps
AWS
Analytics
A/B Testing
Apply
$16k – $36k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Hyderabad
Apply
Data Engineer 1 hour ago
$22k – $55k per year (Estimated) • In office • Full-Time • 3+ years exp • Hyderabad
Databases
Databricks
AI/ML
AI Agents
Analytics
ETL/ELT
Apply
LLM Model Developer 1 hour ago
$28k – $77k per year (Estimated) • In office • Full-Time • 3+ years exp • Hyderabad
AI/ML
AI Agents
Fine-tuning
LLM
Apply
Web Developer 1 hour ago
$25k – $67k per year (Estimated) • In office • Full-Time • 5+ years exp • Hyderabad
JavaScript
SQL
Java
Java
Spring Boot
Apply
$26k – $69k per year (Estimated) • In office • Full-Time • 9+ years exp • Bachelor's Degree • Bengaluru • Hyderabad • Chennai • Noida
Databases
Db2
IMS
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.