996,562open jobs
59,446companies
165,965added this week
Browse all
Salary
≈ $16k – $45k per year (Estimated)
Location
Hybrid (Roorkee, India)
Seniority
Intern
Employment
Internship

First seen by Alion on Sep 30, 2026.

Overview
Company
Impact
Profile match

Fam

Share stories. Meet new people. Make new friends.

About Fam (previously FamPay)

Fam is India's first payments app for everyone above 11. FamApp helps make online and offline payments through UPI and FamCard. We are on a mission to raise a new, financially aware generation, and drive 250 million+ young users in India to kickstart their financial journey super early in their life.

We're reimagining how the next generation experiences fintech-going beyond payments to build a lifestyle brand that blends money, identity, and everyday experiences into one seamless, intuitive journey.

Founded in 2019 by IIT Roorkee alumni, Fam is backed by some of the most respected investors around the world like Elevation Capital, Y-Combinator, Peak XV (Sequoia Capital) India, Venture Highway, Global Founder's Capital and the likes of Kunal Shah, Amrish Rao as angel investors.

About the role

We are looking for Data Engineering Interns to work closely with the Fam Data Team in building and operating our data platform. You will contribute to our lakehouse, data pipelines and data quality checks, which power analytics, product and compliance reporting for a UPI/fintech platform serving millions of users. You will begin with well-scoped tasks under the guidance of a mentor and progressively take end-to-end ownership of individual pipeline tasks.

On the Job

  • Platform Understanding: Learn how our OLTP source systems feed the OLAP lakehouse. Read table schemas, including column types and partition columns, and trace the end-to-end flow of an individual Airflow DAG task.
  • SQL & Transforms: Write and modify SQL queries and simple Spark/SQL transformations under guidance, and validate the correctness of their output.
  • Pipeline Quality: Add basic row-count and null checks, along with meaningful logging, to the pipeline tasks you own. Understand the importance of idempotent pipelines and apply this principle in your changes.
  • Pipeline On-call (Shadow): Participate in the pipeline on-call rotation alongside an experienced engineer. Identify failed runs and data freshness breaches, escalate with the relevant DAG and run details, and execute existing backfill playbooks under guidance.
  • Cost-aware Querying: Apply partition filters instead of full table scans, and understand that storage grows with data volume and every query incurs compute cost.
  • Analyst Support: Collaborate with product analysts as a peer by fixing queries, directing them to the right datasets in the data catalog, and publishing simple, reusable views.
  • AI-assisted Data Work: Use the tools provided, such as the Trino MCP, Text-to-SQL and data discovery tools, in day-to-day work, while ensuring that PII is never included in prompts.
  • AI-assisted Development: Use AI coding assistants (Cursor, Claude Code, Copilot or similar) to accelerate development tasks such as writing SQL, transformations, tests and scripts. Review, test and fully understand all generated code before raising a pull request; accountability for the code remains with you.
  • Experimentation: Identify small improvement ideas, build prototypes and present them during sprint reviews. The focus is on learning rather than measurable impact.

Must-haves:

  • Final-year student or recent graduate in Computer Science, Information Technology or a related field, or equivalent practical experience.
  • Clear understanding of OLTP vs OLAP systems and their respective use cases.
  • Ability to write correct, readable SQL, including joins, aggregations, filters and basic window functions.
  • Working knowledge of Python; familiarity with Scala or Java is a plus.
  • Ability to read a table schema and explain the data it represents.
  • Basic understanding of how LLM / GPT models work, including tokens, context windows, prompting and the causes of hallucination.
  • Practical experience using AI coding assistants for development tasks, with the judgement to verify their output.
  • At least one academic project or internship involving data or backend engineering that you can explain in detail.
  • Proficiency with Git, Linux and the command line.
  • A curious mindset: you question unexpected data patterns and escalate issues early.

Good to have

  • Exposure to Apache Spark, Airflow or any data processing or orchestration framework.
  • Understanding of idempotency, partitioning and columnar file formats such as Parquet.
  • Ability to read a simple query plan and identify full table scans.
  • Experience running an open-weight LLM locally (Ollama, llama.cpp or similar) or building a small embedding and similarity search prototype.
  • Basic understanding of Retrieval-Augmented Generation (RAG), including chunking, embeddings, retrieval and grounding responses in source data.
  • Exposure to vector databases such as pgvector, Qdrant, Chroma or OpenSearch.
  • Familiarity with AWS fundamentals such as S3 and IAM.
  • Exposure to BI tools such as Superset, Metabase or Power BI.
  • Interest in fintech, payments or data privacy.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
996,562 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
Roorkee
≈ $21k – $63k per year (Estimated) • Remote (likely EAEU) • Internship • Bachelor's Degree • Moscow
Python
SQL
Python
pySpark
AI/ML
Spark
MLFlow
PyTorch
Machine Learning
DevOps
Docker
Kubernetes
Grafana
Apply
Sr. Data Architect 4 months ago
≈ $127k – $252k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Atlanta
Databases
Snowflake
Databricks
Microsoft Fabric
AI/ML
dbt
Embeddings
AI Agents
LLM
LLMOps
Feature Store
GraphRAG
DevOps
Azure
CI/CD
AWS
FinOps
Analytics
Data Vault
Dimensional Modeling
Apply
≈ $26k – $46k per year (Estimated) • Hybrid • Full-Time • 14+ years exp • Master's Degree • Hyderabad
Python
SQL
Databases
Snowflake
Databricks
AI/ML
Embeddings
Scikit-learn
Computer Vision
NLP
TensorFlow
PyTorch
Feature Store
OCR
Recommender Systems
Machine Learning
DevOps
GCP
CI/CD
AWS
Apply
$228k – $423k per year • Hybrid • Full-Time • 8+ years exp • Master's Degree • San Francisco • New York • Bellevue
Python
SQL
Databases
Snowflake
AI/ML
Spark
Reinforcement Learning
NLP
TensorFlow
Keras
PyTorch
Recommender Systems
Machine Learning
DevOps
GCP
AWS
Apply
Staff Data Scientist 2 months ago
≈ $32k – $64k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Master's Degree • Bengaluru
AI/ML
Computer Vision
Speech Recognition
Text-to-Speech
Machine Learning
Apply
BI Engineer 1 day ago
$77k – $85k per year • Remote (Canada) • Full-Time • 2+ years exp • Toronto
Python
JavaScript
SQL
Scala
Databases
Snowflake
Amazon Redshift
AI/ML
Machine Learning
Chips/EDA
OpenLane
Analytics
ETL/ELT
A/B Testing
Domo
Management
Google Sheets
Apply
≈ $15k – $39k per year (Estimated) • Hybrid • Master's Degree • Hanoi
Python
SQL
DevOps
GCP
Azure
AWS
Analytics
Power BI
ETL/ELT
Informatica
Collibra
Apply
$203k – $268k per year • Equity • Hybrid • 10+ years exp • Bachelor's Degree • San Francisco
Python
SQL
Databases
Databricks
AI/ML
AI Agents
DevOps
AWS
Analytics
Fivetran
Looker
Apply
$135k – $243k per year • Equity • In office • Full-Time • 7+ years exp • Bachelor's Degree • Bellevue • Overland Park • Herndon
SQL
Databases
Databricks
AI/ML
Copilot
Claude
ChatGPT
DevOps
Azure
Analytics
Tableau
Power BI
ETL/ELT
A/B Testing
Apply
$142k – $147k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • Sunrise
Python
SQL
Analytics
Tableau
Power BI
Apply
≈ $16k – $34k per year (Estimated) • In office • Full-Time • Roorkee
Apply
Sales Manager 2 months ago
≈ $13k – $36k per year (Estimated) • In office • Full-Time • Roorkee
Apply
≈ $16k – $34k per year (Estimated) • In office • Full-Time • Roorkee
Apply
In office • Full-Time • Bachelor's Degree • Roorkee
Apply
See all jobs
This is one of many
996,562 more open roles from verified company boards, updated every day.