368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$21k – $54k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Middle · 3+ years exp
Overview
Company
Impact
Profile match
Accelerating Research ZAGENO (pronounced ZAH-GEE-NOH) exists to make lab supply ordering easier, so research scientists can focus on accelerating innovation.

ZAGENO is hiring a Data Engineer to own the operational data pipelines powering our life sciences catalogue. This is a high-ownership role: you'll be the primary engineer responsible for the reliability, correctness, and scalability of the pipelines our operations teams depend on daily. The work spans production pipeline reliability, reducing accumulated technical debt in existing systems, building observability and testing infrastructure, and partnering with Data Science and Analytics on clean data delivery. You'll work directly with CatalogOps and business stakeholders - turning evolving, often underspecified business rules into architectures that stay flexible without compromising data quality. You'll make architectural decisions, push back on requests that introduce heuristic debt, and own incident response end-to-end.

Responsibilities:

  • Own reliability and performance of operational pipelines across our product catalogue infrastructure.
  • Identify and reduce technical debt, replacing reactive patches with designed, testable logic.
  • Build and maintain low-latency data APIs that serve downstream operational and analytics consumers.
  • Implement CDC patterns to keep catalogue data synchronised across systems with minimal lag.
  • Build monitoring, observability, and automated testing so failures surface before stakeholders report them.
  • Design and implement unit standardisation and master data logic at catalogue scale.
  • Translate business requirements from non-technical stakeholders into durable pipeline logic.
  • Own code versioning, deployment, and incident response for your layer.
  • Leverage your expertise in the tech stack: BigQuery, Databricks, Spark/PySpark, AWS/GCP.

Requirements:

  • 3+ years as a Data Engineer, including solo or primary ownership of production pipelines.
  • Strong Python - data engineering, transformation logic, testing discipline.
  • Strong SQL with the ability to write correct queries, identify and refactor anti-patterns.
  • Databricks, Delta Lake, Airflow for production orchestration.
  • Experience with CDC patterns for real-time or near-real-time data synchronisation.
  • Experience building low-latency APIs serving operational or analytical consumers.
  • Test-driven development discipline - unit tests, integration tests, regression coverage as standard practice, not an afterthought.
  • Operates independently under ambiguity; designs systems to be maintained, not just to run.
  • Preferred: Kafka or equivalent event streaming platform experience.
  • Experience with entity matching, deduplication, or master data management.
  • Exposure to ML pipeline support in production.
  • Familiarity with NLP techniques for entity resolution or text normalisation (tokenisation, similarity matching, named entity recognition).

What success looks like:

  • Engineers ship data products: Outputs are documented, versioned, and designed for reuse across Analytics, Data Science, and operational consumers.
  • Reliability: Pipelines run reliably with minimal manual intervention.
  • Performance: Data latency and downtime decrease measurably over time.
  • Data Quality: Analytics and Data Science teams receive clean, trustworthy data without ad hoc fixes.
  • Scalability: Infrastructure scales with volume growth without proportional cost increase.
  • Reduction of Debt: Technical debt in existing pipelines decreases measurably as heuristic patches are replaced with designed logic.
  • Synchronisation: CDC-driven synchronisation eliminates the class of cross-pipeline identity drift bugs.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$220k – $325k per year • Remote/Hybrid • Full-Time • 15+ years exp • New York
DevOps
AWS
Azure
CI/CD
GCP
Kubernetes
Platform Engineering
Cybersecurity
Threat Modeling
Apply
Remote/Hybrid • 7+ years exp • Bachelor's Degree
C#
JavaScript
Python
SQL
TypeScript
C#
ASP.NET Core
Blazor
Dapper
Databases
Oracle
Frontend
Angular
Bootstrap
JQuery
Vue.js
DevOps
AWS
Azure
Azure DevOps
CI/CD
Git
Apply
$30k – $101k per year (Estimated) • Remote/Hybrid • 4+ years exp • Madrid
Java
Java
Spring Boot
Databases
Apache Kafka
DevOps
Dynatrace
Apply
$54k – $122k per year (Estimated) • Remote/Hybrid • Madrid
Java
Java
Spring Boot
Databases
Apache Kafka
DevOps
Dynatrace
Apply
$26k – $90k per year (Estimated) • Remote/Hybrid • 2+ years exp • Madrid
Python
Python
Celery
Django
Databases
Apache Kafka
MySQL
PostgreSQL
RabbitMQ
Redis
Apply
$34k – $71k per year (Estimated) • In office • 5+ years exp • Bengaluru
AI/ML
Recommender Systems
Analytics
A/B Testing
Apply
$28k – $63k per year (Estimated) • In office • Bengaluru
Design
Figma
Apply
Senior Data Engineer 14 days ago
$31k – $60k per year (Estimated) • In office • 3+ years exp • Bengaluru
Python
SQL
Python
pySpark
Databases
Apache Kafka
Databricks
Delta Lake
Google BigQuery
AI/ML
NLP
Spark
Tokenization
DevOps
AWS
GCP
Apply
DevOps Engineer 26 days ago
$18k – $82k per year (Estimated) • In office • Bengaluru
C++
JavaScript
Python
Ruby
Databases
Apache Kafka
DevOps
ArgoCD
AWS
CI/CD
Configuration Management
Datadog
Docker
FluxCD
GCP
Git
GitHub Actions
Kubernetes
Terraform
GitHub
Apply
$34k – $71k per year (Estimated) • In office • 4+ years exp • Bengaluru
SQL
AI/ML
Claude
Hallucination
LLM
AI Agents
Apply
$16k – $34k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Mumbai • Bengaluru
JavaScript
PowerShell
SQL
C#
C#
.NET
Databases
Azure SQL Database
MS SQL
DevOps
Azure
Rest API
Cybersecurity
Microsoft Entra ID
QA
Postman
Swagger
Apply
$41k – $89k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bengaluru
C#
TypeScript
JavaScript
C#
.NET
Databases
Apache Kafka
AI/ML
Copilot
LLM
OpenAI
Frontend
Angular
GraphQL
DevOps
Azure
Azure AKS
Azure DevOps
CI/CD
Docker
GitHub
GitHub Actions
Grafana
Kubernetes
Prometheus
Rest API
Apply
$38k – $83k per year (Estimated) • In office • Full-Time • 12+ years exp • Bachelor's Degree • Bengaluru
Databases
Oracle
DevOps
AWS
Platform Engineering
Apply
Data Architect 3 hours ago
$38k – $91k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru • Pune
Node JS
Python
SQL
JavaScript
Databases
Databricks
MongoDB
Redis
Apply
$28k – $71k per year (Estimated) • In office • Full-Time • 5+ years exp • Bengaluru
DevOps
CI/CD
Platform Engineering
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.