823,562open jobs
53,068companies
134,459added this week
Browse all
Salary
≈ $23k – $48k per year (Estimated)
Location
Hybrid (Hyderabad, India)
Seniority
Senior · 5+ years exp

First seen by Alion on Sep 24, 2026.

Overview
Company
Impact
Profile match
HiringBlaze is a recruitment firm in India that sources technology candidates for client companies and advertises their openings on job boards such as hirist. Its postings describe unnamed clients, among them a US enterprise software firm with a development centre in Noida, a supply chain and fintech SaaS company and a long-established medical devices maker in Agra, with roles in Gurgaon, Hyderabad, Bengaluru, Mumbai, Chennai and Pune. It recruits data engineers, database administrators, backend, Java and full stack developers, DevOps engineers, test automation engineers, data scientists, AI engineers and solution architects.

Description :

Role : AI Data Engineer

- Location : Rai Durg, Hyderabad


- Work mode- Hybrid model working (3 days work from office)


- Experience : 5 - 8 Years (Minimum 5 years- AI Data Engineer)


- Mandatory Skills : DVC (Data Version Control) and Airflow, Apache Spark, Flink, and Kafka, Advanced level Python and AI logic and Rust (or C++), Vector Database Mastery like configuration of HNSW indexes, scalar quantization, and metadata filtering strategies


- Qualification : Bachelor of Engineering - Bachelor of Technology (B.E./B.Tech.)


- Notice period : Immediate / early joiners (Max. 15-30 days)


- Interview Process : 2 - 3 Technical rounds

Important Note :

- We are currently prioritizing immediate / early joiners (maximum 15-30 days- notice period above 30 days will be automatically rejected.).


- All mandatory technical skills must be clearly highlighted within the project descriptions in your resume, not just listed in the Skills or Roles & Responsibilities sections .

Position Overview :

We are seeking a hardcore, hands-on AI Data Engineer to build the high-performance data infrastructure required to power autonomous AI agents. You won't just be moving data from A to B; you will be architecting Dynamic Context Windows, managing Real-time Semantic Indexes, and building Self-Cleaning Data Pipelines that feed our "Super Employee" agents.

Key Responsibilities :

- Vector & Graph ETL : Design and maintain pipelines that transform unstructured data (PDFs, emails, logs, chats) into optimized embeddings for Vector Databases (Pinecone, Weaviate, Milvus).


- Semantic Data Modeling : Engineer data structures that optimize for Retrieval-Augmented Generation (RAG), ensuring agents find the "needle in the haystack" in milliseconds.


- Knowledge Graph Construction : Build and scale Knowledge Graphs (Neo4j) to represent complex relationships in our trading and support data that standard vector search misses.


- Automated Data Labeling & Synthetic Data : Implement pipelines using LLMs to auto-label datasets or generate synthetic edge cases for agent training and evaluation.


- Stream Processing for Agents : Build real-time data "listeners" (Kafka/Flink) that feed live context to agents, allowing them to react to market or support events as they happen.


- Data Reliability & "Drift" Detection : Build monitoring for "Embedding Drift", identifying when the statistical distribution of your data changes and the agent's "knowledge" becomes stale.

Qualifications :

- Vector Database Mastery : Expert-level configuration of HNSW indexes, scalar quantization, and metadata filtering strategies within Pinecone, Milvus, or Qdrant.


- Advanced Python & Rust : Proficiency in Python for AI logic and Rust (or C++) for high-performance data processing and custom embedding functions.


- Big Data Ecosystem : Hands-on experience with Apache Spark, Flink, and Kafka in a high-throughput environment (Trading/FinTech preferred).


- LLM Data Tooling : Deep experience with Unstructured.io, LlamaIndex, or LangChain for document parsing and chunking strategy optimization.


- MLOps & DataOps : Mastery of DVC (Data Version Control) and Airflow/Prefect for managing complex, non-linear AI data workflows.


- Embedding Models : Understanding of how to fine-tune embedding models (e.g., BGE, Cohere, or OpenAI) to better represent domain-specific (Trading) terminology.

Additional qualifications :

- Chunking Strategy Architect : You don't just "split text." You implement Semantic Chunking and Parent-Child retrieval strategies to maximize LLM context relevance.


- Cold/Warm/Hot Storage Strategy : Managing cost and latency by tiering data between Vector DBs (Hot), SQL/NoSQL (Warm), and S3/Data Lakes (Cold).


- Privacy & Redaction Pipelines : Building automated PII (Personally Identifiable Information) redaction into the ingestion layer to ensure agents never "see" or "leak" sensitive user data.

Why Join ?

- Opportunity to lead transformative initiatives, modernizing legacy systems and shaping the future of trading technology.


- Work with cutting-edge technologies in a dynamic, fast-paced environment.


- Competitive compensation, professional growth opportunities, and the chance to work with industry-leading experts.


Skills

Data Engineering, Artificial Intelligence, Data Infrastructure, Data Pipeline, ETL, LLM, MLOps, LangChain, Python

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
823,562 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Data Science
Similar stack
Same company
Hyderabad
≈ $74k – $165k per year (Estimated) • Equity • Hybrid • Full-Time • Bachelor's Degree • Munich
Python
SQL
AI/ML
LLM
Machine Learning
DevOps
GCP
Azure
AWS
Analytics
ETL/ELT
Apply
≈ $79k – $172k per year (Estimated) • Hybrid • Bachelor's Degree • London
Python
SQL
Analytics
Tableau
Power BI
Apply
≈ $44k – $106k per year (Estimated) • Remote (Portugal) • Full-Time • 5+ years exp • Lisbon
Python
SQL
AI/ML
Spark
DevOps
GCP
Azure
CI/CD
AWS
Analytics
ETL/ELT
Azure Data Factory
Apply
≈ $53k – $116k per year (Estimated) • In office • Full-Time • 6+ years exp • PhD • Lisbon
Python
SQL
Python
pySpark
Databases
Databricks
Delta Lake
AI/ML
Spark
DevOps
Azure DevOps
Azure
CI/CD
Git
Analytics
ETL/ELT
Management
Agile
Apply
Data Engineer 10 days ago
≈ $41k – $94k per year (Estimated) • In office • Full-Time • Lisbon
SQL
Databases
Google BigQuery
BigQuery
AI/ML
dbt
DevOps
GCP
Analytics
Collibra
Apply
$150k – $250k per year • Equity 0.5–2% • In office • Full-Time • 3+ years exp • New York
Python
C++
AI/ML
AI Agents
Embodied AI
Robotics
ROS2
Model Predictive Control
Apply
≈ $67k – $140k per year (Estimated) • Hybrid • Full-Time • Bachelor's Degree • United States
Python
PowerShell
DevOps
Rest API
Terraform
Ansible
Red Hat
VMWare
Windows Server
Hyper-V
Linux
Management
Agile
Apply
≈ $70k – $145k per year (Estimated) • In office • Full-Time • 5+ years exp • Lynchburg
Python
Java
PowerShell
Databases
ElasticSearch
DevOps
Prometheus
CI/CD
Grafana
Linux
Windows
Apply
≈ $39k – $117k per year (Estimated) • Remote (United Kingdom, Ireland) • 4+ years exp
Python
Java
Ruby
PowerShell
C#
Bash
C#
.NET
DevOps
Puppet
Ansible
Chef
CI/CD
AWS
Configuration Management
AWS Lambda
API Gateway
Linux
Windows
DNS
Management
Agile
Apply
DevOps Engineer 1 day ago
Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Bengaluru
Python
Java
Groovy
AI/ML
Prompt Engineering
DevOps
GCP
Azure
CI/CD
Jenkins
Git
AWS
Docker
GitHub
GitLab
Apply
≈ $25k – $52k per year (Estimated) • Hybrid • 7+ years exp • Pune
Python
Databases
PostgreSQL
Snowflake
AI/ML
dbt
Anomaly Detection
Feature Store
Machine Learning
DevOps
Terraform
CI/CD
AWS
Apply
≈ $20k – $49k per year (Estimated) • Hybrid • 3+ years exp • Hyderabad
Python
Java
Databases
Redis
Cassandra
Apache Kafka
DevOps
Linux
Management
Agile
Apply
Data Scientist II 4 days ago
$31k – $37k per year (gross) • Hybrid • 3+ years exp • Mumbai
Python
SQL
AI/ML
Machine Learning
Analytics
A/B Testing
Apply
$37k – $57k per year (gross) • In office • 8+ years exp • Bengaluru
Python
DevOps
GitHub Actions
CI/CD
Jenkins
Git
Cybersecurity
ISO 27001
SOC 2
Management
Agile
Scrum
QA
Selenium
JMeter
Appium
Postman
Rest-Assured
Pytest
Apply
AI Product Manager 3 days ago
≈ $20k – $55k per year (Estimated) • In office • 2+ years exp • Bengaluru
Python
SQL
AI/ML
Edge AI
Machine Learning
Analytics
Tableau
Management
Jira
Apply
Shift Incharge 11 hours ago
≈ $20k – $44k per year (Estimated) • In office • Full-Time • 5+ years exp • Hyderabad
DevOps
SLI/SLO/SLA
Apply
Senior Data Analyst 11 hours ago
≈ $32k – $73k per year (Estimated) • Hybrid • 6+ years exp • Bachelor's Degree • Hyderabad
SQL
AI/ML
AI Agents
DevOps
Windows
Analytics
Power BI
Microsoft Excel
Apply
Procurement Analyst 11 hours ago
In office • 6+ years exp • Bachelor's Degree • Hyderabad
Analytics
Microsoft Excel
Management
Microsoft Office
Apply
Procurement Analyst 11 hours ago
In office • 6+ years exp • Bachelor's Degree • Hyderabad
Analytics
Microsoft Excel
Management
Microsoft Office
Apply
AI Security Architect 11 hours ago
In office • 7+ years exp • Hyderabad
AI/ML
Copilot
Cursor
LangChain
Claude
Model Context Protocol
Function Calling
AI Agents
Promptfoo
LLM
OpenAI Codex
Red Teaming
LLM Guardrails
Agentic Workflows
Tool Use
DevOps
Azure
CI/CD
AWS
GitHub
IAM
Cybersecurity
Least Privilege
Threat Modeling
Apply
See all jobs
This is one of many
823,562 more open roles from verified company boards, updated every day.