369,078open jobs
9,456companies
47,986added this week
Browse all
Salary
$30k – $59k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Senior
Overview
Company
Impact
Profile match
Apna is an Indian recruitment and professional networking platform founded in 2019 that focuses on connecting frontline, blue-collar, and grey-collar workers with job opportunities. The company operates a mobile-first marketplace that leverages AI-driven matching algorithms to streamline hiring for small businesses, enterprises, and job seekers across hundreds of Indian cities. Beyond job listings, the platform enables users to build professional profiles, upgrade their skills, and engage with professional communities to advance their careers.

Company: Apna

Role: Senior Data Engineer / SSE

Team: Data Platform / Engineering

Location: Work from Office - Domlur, Bangalore (5 days / week)

Experience : 4-6 Years of Experience

Why Join Apna

At Apna, data is central to how we build products, understand users, improve employer outcomes, power recommendations, and scale decision-making. This role gives you the opportunity to build the backbone of Apna’s data platform and influence how data is used across the company.

You will work on real-world, high-scale problems across jobs, users, employers, communities, matching, growth, and AI-driven systems.

About the Role

Apna is looking for a Senior Software Engineer to build and scale our core data platform. This role will work on large-scale data pipelines, lakehouse architecture, query platforms, workflow orchestration, and data reliability systems that power analytics, product intelligence, machine learning, business dashboards, experimentation, and operational decision-making across Apna.

We are looking for someone who can think deeply about data architecture, design reliable pipelines, improve data quality, and help build a platform that can scale with Apna’s growth.

Requirements

What You’ll Own:

You will be responsible for designing, building, and operating critical parts of Apna’s data platform, including:

  • Building scalable batch and near-real-time data pipelines across product, business, growth, and ML use cases.
  • Designing and improving our lakehouse architecture using technologies likeApache Hudi.
  • Working with query engines such asPresto / Trino for large-scale analytical workloads.
  • Building and maintaining orchestration workflows usingApache Airflow.
  • Creating reusable data models, curated datasets, and reliable data marts for analytics and product teams.
  • Improving data platform reliability, observability, SLA tracking, lineage, and data quality checks.
  • Optimizing storage, compute, query performance, and pipeline costs.
  • Partnering with product, analytics, ML, and backend engineering teams to understand data needs and convert them into scalable platform solutions.
  • Driving engineering standards around data modeling, schema evolution, partitioning, deduplication, backfills, replayability, and pipeline ownership.
  • Mentoring data engineers and influencing architecture decisions across teams.

What We’re Looking For

Must Have

  • Strong experience indata engineering, preferably at scale.
  • Hands-on experience withApache Airflow or similar orchestration systems.
  • Strong knowledge ofPresto / Trino or other distributed query engines.
  • Good understanding ofApache Hudi concepts such as:
    • Copy-on-write vs merge-on-read
    • Upserts and deletes
    • Incremental reads
    • Compaction
    • Clustering
    • Timeline and commits
    • Schema evolution
    • Partitioning strategy
  • Strong knowledge of distributed data processing and storage systems.
  • Ability to design and build reliable ETL / ELT pipelines.
  • Strong SQL skills and ability to debug complex data issues.
  • Good understanding of different data architectures, including:
    • Data warehouse
    • Data lake
    • Lakehouse
    • Lambda architecture
    • Kappa architecture
    • Medallion architecture
    • Event-driven data architecture
  • Experience with data modeling for analytics and reporting.
  • Strong programming skills in at least one language such asPython, Java, or Scala.
  • Ability to reason about trade-offs between freshness, cost, reliability, latency, and complexity.
  • Strong debugging and production ownership mindset.

Good to Have

  • Experience with Kafka, Spark, Flink, Hive, Iceberg, Delta Lake, or BigQuery.
  • Experience building internal data platforms or self-serve data infrastructure.
  • Experience with data quality frameworks such as Great Expectations, Deequ, Soda, or custom validation systems.
  • Exposure to ML feature pipelines or feature stores.
  • Experience with metadata management, data catalogs, lineage, and governance.
  • Experience with cloud infrastructure such as AWS, GCP, or Azure.
  • Understanding of privacy, compliance, PII handling, and access control in data systems.

What Success Looks Like

In this role, success means:

  • Critical business and product datasets are reliable, discoverable, and trusted.
  • Pipelines are observable, recoverable, and have clear SLAs.
  • Query performance improves across major analytical workloads.
  • Data freshness and quality issues reduce significantly.
  • Teams can build on top of the data platform faster without reinventing pipelines.
  • The platform can scale with Apna’s user, job, employer, and engagement data.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
369,078 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$121k – $147k per year • Remote • Full-Time • 5+ years exp • PhD • Columbus
Java
Java
Spring Boot
Databases
Apache Kafka
DevOps
Azure
Azure AKS
Azure DevOps
Bicep
CI/CD
Dynatrace
GitLab
Kubernetes
Splunk
Terraform
Management
Jira
Apply
Data Analyst 6 hours ago
$66k – $93k per year • Remote/Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Overland Park
Java
SQL
AI/ML
Hadoop
Marketing
Salesforce
Apply
$128k – $173k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • United States
JavaScript
Node JS
TypeScript
Frontend
React.js
DevOps
Amazon EC2
AWS
AWS CDK
AWS Lambda
CI/CD
GitHub Actions
Jenkins
Terraform
Amazon CloudWatch
Amazon S3
API Gateway
GitHub
GitLab
IAM
Apply
ARG SSR Data Engineer 6 hours ago
$38k – $95k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Argentina
Python
SQL
Databases
Trino
AI/ML
Airflow
DevOps
Amazon S3
AWS
Azure
CI/CD
GCP
Git
Analytics
ETL/ELT
Apply
$35k per year • In office • Internship • Bachelor's Degree • Freiburg im Breisgau
C++
Java
Node JS
Python
C#
JavaScript
C#
.NET
AI/ML
AI Agents
DevOps
CI/CD
GitLab
Apply
$23k – $65k per year (Estimated) • In office • Bengaluru
Python
SQL
Databases
Apache Kafka
AI/ML
Claude
DeepEval
Gemini
Hallucination
LangSmith
LLM
Prompt Engineering
Promptfoo
RAG
Ragas
Semantic Search
TruLens
OpenAI
Recommender Systems
Semantic Search
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
Git
GitHub Actions
Grafana
Jenkins
Kibana
Kubernetes
GitHub
Management
Jira
QA
Appium
JMeter
k6
Locust
Playwright
Postman
Pytest
Rest-Assured
Robot Framework
Selenium
Apply
Product Analyst 8 days ago
$19k – $46k per year (Estimated) • In office • 3+ years exp • Bengaluru
Python
SQL
Analytics
A/B Testing
Marketing
Amplitude
Mixpanel
Apply
$35k – $72k per year (Estimated) • In office • Internship • 5+ years exp • Master's Degree • Bengaluru
AI/ML
AI Agents
LLM
RAG
Apply
$34k – $71k per year (Estimated) • In office • Internship • 5+ years exp • Master's Degree • Bengaluru
AI/ML
AI Agents
RAG
Apply
Product Manager 18 days ago
$26k – $60k per year (Estimated) • In office • 3+ years exp • Bengaluru
SQL
AI/ML
ChatGPT
Claude
Gemini
Prompt Engineering
OpenAI
Recommender Systems
Analytics
Power BI
A/B Testing
Marketing
Amplitude
Mixpanel
Apply
$27k – $69k per year (Estimated) • In office • 7+ years exp • Bachelor's Degree • Bengaluru
Apex
Apex
Aura Components
Lightning Web Components
MuleSoft
Visualforce
Marketing
Salesforce
Apply
$31k – $82k per year (Estimated) • In office • Full-Time • 3+ years exp • Hyderabad • Bengaluru
Apply
$31k – $73k per year (Estimated) • In office • Full-Time • 5+ years exp • Bengaluru
Apply
$16k – $34k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Mumbai • Bengaluru
JavaScript
PowerShell
SQL
C#
C#
.NET
Databases
Azure SQL Database
MS SQL
DevOps
Azure
Rest API
Cybersecurity
Microsoft Entra ID
QA
Postman
Swagger
Apply
$37k – $73k per year (Estimated) • In office • Internship • 4+ years exp • Bachelor's Degree • Bengaluru
Python
Scala
SQL
Databases
Apache Kafka
Databricks
AI/ML
ChatGPT
Copilot
Cursor
Spark
DevOps
AWS
Azure
CI/CD
GCP
Git
GitHub
Terraform
Apply
See all jobs
This is one of many
369,078 more open roles from verified company boards, updated every day.