937,957open jobs
57,076companies
155,728added this week
Browse all
Salary
≈ $20k – $49k per year (Estimated)
Location
Hybrid (Pune, India)
Seniority
Senior
Employment
Full-Time

Confirmed on the employer's own hiring board on Sep 29, 2026. First seen by Alion on Sep 28, 2026. Citi scores A on the Alion truth index.

Overview
Company
Impact
Profile match
Citi is one of the largest banks in the world, tracing its lineage to the City Bank of New York founded in 1812 and taking its modern shape through the 1998 merger that created Citigroup. Its most distinctive asset is a cross-border payments and treasury network unmatched by any competitor, moving trillions of dollars a day for multinational corporations, governments and other banks across roughly ninety countries. Alongside that institutional franchise it runs markets and investment banking, wealth management and a United States personal bank, and has spent recent years simplifying itself by exiting consumer operations across Asia, Europe and Latin America.

Technology

Join a small, high-impact engineering team in Citi Markets Technology building the data foundation for greenfield Generative AI products across asset classes. We are looking for an experienced Data Engineer who combines deep Python expertise with strong engineering judgement and a practical understanding of how to build reliable, high-performance data platforms.

The role goes beyond constructing ETL pipelines. You will help define how billions of records from diverse Markets data sources are collected, validated, transformed, governed, and made available to production AI applications with consistently low retrieval latency.

If you want to work on ambitious data engineering problems, shape a platform from the ground up, and help define how data is engineered for production AI in Markets, this is the role for you.

The Team

We are a fast-moving team specialising in Generative AI within Markets Technology. We build greenfield products that span multiple asset classes and solve real business problems using modern AI and data engineering approaches.

Our applications depend on a robust and well-designed data foundation. That means creating pipelines and serving layers that can process billions of records while preserving data quality, provenance, security, and operational reliability.

The team is still small enough for every engineer to have genuine influence. Data engineering is a core part of the product architecture, not a downstream support function.

The Role

As a Vice President in the team, you will be a hands-on senior engineer responsible for designing and building the data foundation for greenfield Generative AI products across Markets, including our conversational AI platform.

You will develop production-grade data pipelines and services, primarily using Python, to ingest, validate, transform, enrich, and serve data from a wide range of internal and external sources. You will help create a curated, high-performance data-serving layer that enables low-latency retrieval by our AI and application services.

You will contribute directly to code while also shaping architecture, engineering standards, data models, quality controls, and operational practices. We are looking for someone who can make pragmatic technology choices, challenge assumptions, and take long-term ownership of the platform.

What You’ll Do

  • Design, build, and evolve the data architecture supporting Generative AI products across Markets
  • Develop scalable Python pipelines that ingest, validate, transform, enrich, and integrate billions of records from diverse sources
  • Build curated, queryable datasets and high-performance serving layers for low-latency application access
  • Design and optimise the application’s data storage strategies, initially centred on PostgreSQL and Parquet
  • Select suitable processing approaches for each workload, from efficient in-process and columnar processing to distributed frameworks where required
  • Optimise ingestion, transformation, storage, indexing, and query performance
  • Establish robust controls for data quality, reconciliation, lineage, schema evolution, idempotency, and recovery
  • Build production observability into data pipelines, including metrics, logging, alerting, and operational diagnostics
  • Design solutions that respect data classification, entitlements, access controls, and security requirements
  • Contribute directly to code, architecture reviews, technical standards, and the wider engineering direction of the team
  • Build automated tests and CI/CD pipelines that enable reliable and repeatable delivery
  • Collaborate closely with AI engineers, software engineers, architects, product partners, and Markets stakeholders
  • Use AI-assisted engineering tools, including Devin and GitHub Copilot, to improve development quality and productivity

What We’re Looking For

  • Extensive hands-on experience in data engineering, software engineering, or a closely related discipline
  • Deep practical expertise in Python and the ability to build maintainable, production-grade software
  • A proven track record of designing and delivering large-scale data platforms or data-intensive applications
  • Strong SQL skills and extensive experience with relational databases, particularly PostgreSQL
  • Experience with data modelling, indexing, partitioning, query optimisation, and database performance tuning
  • Strong knowledge of the Python data ecosystem, including libraries such as pandas, PyArrow, SQLAlchemy, and NumPy
  • Practical experience working with columnar formats such as Apache Parquet and selecting efficient storage and serialisation strategies
  • Experience processing large-scale datasets using technologies such as Apache Spark, Dask, Polars, or equivalent frameworks
  • Strong understanding of ETL and ELT architecture, including incremental processing, idempotency, failure recovery, and schema evolution
  • Experience implementing automated data quality controls, reconciliation, lineage, monitoring, and operational alerting
  • Strong software engineering fundamentals, including design, testing, maintainability, code review, and CI/CD
  • Experience deploying and operating services on container platforms such as Kubernetes or OpenShift
  • Understanding of data governance and security practices, including encryption, masking, classification, entitlements, and fine-grained access control
  • The ability to operate in ambiguity, take ownership, and influence the technical direction of a product
  • Strong communication skills and a collaborative approach to engineering

What Makes This Role Different

This is not a role focused on maintaining legacy ETL jobs or moving data between systems without understanding how it will be used.

You will help build the data foundation of a new generation of AI products in Markets. The engineering challenges include integrating complex datasets, processing billions of records, delivering consistently low retrieval latency, and meeting the quality, security, and reliability standards expected of production financial systems.

The role offers the opportunity to influence the architecture from an early stage, work closely with AI and application engineers, and take genuine ownership of a platform that is central to the product.

Preferred Experience

  • Experience with workflow orchestration technologies such as Apache Airflow, Dagster, or Prefect
  • Familiarity with Generative AI applications, retrieval architectures, LLMs, agentic systems, or structured evaluation
  • Experience building data foundations for search, retrieval, analytics, machine learning, or AI applications
  • Knowledge of financial instruments, trading concepts, and data structures within FX, Equities, or other capital markets domains
  • Experience working with temporal, reference, market, or transactional data
  • A bachelor’s or master’s degree in Computer Science, Engineering, or another relevant quantitative discipline, or equivalent professional experience

------------------------------------------------------

Job Family Group:

Technology

------------------------------------------------------

Job Family:

Applications Development

------------------------------------------------------

Time Type:

Full time

------------------------------------------------------

Most Relevant Skills

Please see the requirements listed above.

------------------------------------------------------

Other Relevant Skills

For complementary skills, please see above and/or contact the recruiter.

------------------------------------------------------

Citi is an equal opportunity employer, and qualified candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, status as a protected veteran, or any other characteristic protected by law.

If you are a person with a disability and need a reasonable accommodation to use our search tools and/or apply for a career opportunity review Accessibility at Citi.

View Citi’s EEO Policy Statement and the Know Your Rights poster.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
937,957 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
Pune
≈ $28k – $50k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Databases
Snowflake
Apache Kafka
Amazon Redshift
AI/ML
Streamlit
DevOps
AWS
AWS Lambda
Amazon S3
Amazon Kinesis
Cybersecurity
PCI DSS
SOC 2
GDPR
HIPAA
Analytics
Tableau
Power BI
Looker
Data Vault
Apply
≈ $27k – $47k per year (Estimated) • Remote (India) • Full-Time • 5+ years exp • Bachelor's Degree • India
Python
SQL
Databases
Microsoft Fabric
AI/ML
Copilot
DevOps
CI/CD
Git
GitHub
Analytics
Power BI
ETL/ELT
Azure Data Factory
Dimensional Modeling
Apply
Data Engineer 1 hour ago
≈ $20k – $41k per year (Estimated) • In office • Full-Time • 5+ years exp • Hyderabad
Python
SQL
Python
pySpark
Databases
MySQL
PostgreSQL
Databricks
Presto
Amazon Redshift
AI/ML
Spark
Pandas
PyTorch
DevOps
GCP
CloudFormation
Azure
AWS
AWS Lambda
GitHub
Amazon S3
Analytics
Tableau
ETL/ELT
Management
Agile
Apply
≈ $14k – $32k per year (Estimated) • In office • Full-Time • Bengaluru • Chennai • Gurgaon
AI/ML
LangChain
LlamaIndex
NLP
Sentiment Analysis
Apply
≈ $22k – $44k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru • Gurgaon
Python
Databases
PostgreSQL
Weaviate
Neo4j
Chroma
Milvus
pgvector
Pinecone
AI/ML
LangGraph
LangChain
LlamaIndex
Prompt Engineering
Multimodal AI
AI Agents
NLP
Llama
Mistral
Transformers
Gemini
LLM
RAG
OpenAI
Anthropic
Knowledge Graph
DevOps
GCP
Azure
AWS
Apply
$187k – $229k per year • In office • Full-Time • 8+ years exp • Chicago
Python
JavaScript
TypeScript
SQL
C#
Node JS
Databases
MySQL
PostgreSQL
ElasticSearch
Azure Cosmos DB
AI/ML
Prompt Engineering
Function Calling
AI Agents
LLM
Structured Outputs
Context Engineering
Tool Use
Frontend
Vue.js
Angular
React.js
DevOps
Azure
CI/CD
Octopus Deploy
Apply
In office • Contractor • Bachelor's Degree • Bengaluru
Python
SQL
Python
pySpark
Databases
MySQL
PostgreSQL
Oracle
AI/ML
Spark
DevOps
Rest API
GitHub Actions
CI/CD
Jenkins
Git
AWS
AWS Lambda
GitLab
Amazon S3
IAM
Amazon CloudWatch
Amazon EventBridge
AWS Step Functions
SOAP
Cybersecurity
Least Privilege
Analytics
ETL/ELT
AWS Glue
Apply
$132k – $165k per year • Remote (United States) • Full-Time • 3+ years exp • Bachelor's Degree • United States
JavaScript
TypeScript
SQL
Node JS
Node JS
Nest.JS
BullMQ
TypeORM
Databases
PostgreSQL
Redis
Apache Kafka
Frontend
GraphQL
Next.js
React.js
pnpm
Turborepo
DevOps
Rest API
GCP
AWS
Management
Slack
Apply
≈ $14k – $32k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
Databases
Databricks
AI/ML
Spark
DevOps
Azure
CI/CD
Analytics
Azure Data Factory
Management
Agile
Apply
≈ $87k – $171k per year (Estimated) • Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Portland
SQL
Databases
Snowflake
Databricks
Azure SQL Database
DevOps
CI/CD
Git
Analytics
Power BI
ETL/ELT
SSIS
Dimensional Modeling
Apply
≈ $14k – $33k per year (Estimated) • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Pune
Python
SQL
Python
pySpark
Databases
Apache Kafka
AI/ML
Spark
Airflow
Pandas
DevOps
CI/CD
Git
Analytics
ETL/ELT
Dimensional Modeling
Management
Agile
Apply
Data Scientist 1 day ago
≈ $15k – $37k per year (Estimated) • Hybrid • Full-Time • 4+ years exp • Master's Degree • Gurgaon
Python
SQL
Databases
Teradata
AI/ML
Spark
AI Agents
Pandas
NumPy
Machine Learning
Analytics
Microsoft Excel
Apply
≈ $24k – $42k per year (Estimated) • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Pune
Python
SQL
Databases
Oracle
AI/ML
Copilot
DevOps
CI/CD
Git
Analytics
Tableau
ETL/ELT
Management
Agile
Apply
≈ $19k – $38k per year (Estimated) • Hybrid • Full-Time • 5+ years exp • Pune • Chennai
SQL
Apply
≈ $14k – $33k per year (Estimated) • Hybrid • Full-Time • 4+ years exp • Pune
Java
SQL
DevOps
Splunk
Ansible
Docker
Kubernetes
Grafana
AppDynamics
TeamCity
GitHub
Unix
Apply
≈ $13k – $30k per year (Estimated) • Remote (India) • Full-Time • Bengaluru • Hyderabad • Pune • Greater Noida
SQL
Databases
PostgreSQL
Db2
Oracle
Cassandra
MS SQL
DevOps
Ansible
Dynatrace
Azure
Git
Configuration Management
Cybersecurity
Active Directory
LDAP
Management
ServiceNow
ITSM
Apply
≈ $32k – $70k per year (Estimated) • In office • Full-Time • Pune
Apply
Hybrid • Full-Time • Bachelor's Degree • Pune
Analytics
Alteryx
Management
Microsoft Office
Apply
≈ $18k – $39k per year (Estimated) • Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Pune
Java
SQL
C#
C#
.NET
Mobile
JUnit
DevOps
Azure DevOps
GitHub Actions
Azure
CI/CD
Jenkins
Git
AWS
Docker
GitHub
Management
Jira
Agile
QA
TestNG
Selenium
Cucumber
JMeter
Playwright
Swagger
Postman
Rest-Assured
Apply
In office • Full-Time • Pune • Chennai
Apply
See all jobs
This is one of many
937,957 more open roles from verified company boards, updated every day.