1,437,834open jobs
84,810companies
219,155added this week
Browse all
Salary
≈ $12k – $29k per year (Estimated)
Location
In office (Hyderabad)
Seniority
Junior · 2+ years exp

First seen by Alion on Oct 8, 2026.

Overview
Company
Impact
Profile match
SMARTnCODE provides AI, Digital Transformation, Cloud, Data Analytics, Hyper Automation and Software Development services.

Position : AI/GenAI Engineer (LLM Integration Specialist)

Experience : 2 - 4 years

CTC : 12 to 16 LPA

Location : Hyderabad

About the Role :

We're building a high-performance chat application and looking for an AI/GenAI Engineer to lead the integration and optimization of Large Language Models (LLMs). You'll be responsible for connecting LLM APIs, implementing domain-specific fine-tuning strategies, prompt engineering, and ensuring optimal performance for production use.

Key Responsibilities :

LLM Integration & Architecture :

- Integrate multiple LLM APIs (OpenAI, Anthropic Claude, Google Gemini, or open-source models)

- Design and implement robust API wrapper services with retry logic, fallback mechanisms, and error handling

- Implement streaming responses for real-time chat experience

- Build rate limiting and quota management systems

- Handle token counting, context window management, and cost optimization

Domain Customization & Fine-tuning :

- Develop domain-specific prompt engineering strategies

- Implement RAG (Retrieval Augmented Generation) pipelines using vector databases

- Fine-tune or adapt models for specific use cases using techniques like LoRA, prompt tuning

- Create and maintain knowledge bases for domain-specific responses

Performance & Optimization :

- Optimize API response times and reduce latency

- Implement caching strategies for common queries

- Monitor and optimize token usage to control costs

Safety & Quality :

- Implement content moderation and safety filters

- Build guardrails to prevent prompt injection and jailbreaking

- Develop evaluation frameworks to measure response quality

Technical Stack :

- Languages : Python (primary), JavaScript/TypeScript (basic understanding)

- LLM APIs : OpenAI, Anthropic, Google Gemini, Cohere

- Frameworks : LangChain, LlamaIndex, FastAPI

- Vector DBs : Pinecone, Weaviate, Qdrant, or ChromaDB

- Infrastructure : Docker, Redis, PostgreSQL, Message Queues

- Cloud : AWS/GCP/Azure


Infrastructure :


- Design scalable architecture for handling concurrent LLM requests


- Implement queue systems for managing high-volume API calls


- Set up monitoring and logging for LLM interactions


- Work with DevOps to deploy models (if self-hosted)


Required Skills & Experience :


Must Have :


- Experience working with LLMs and GenAI technologies


- Strong experience with OpenAI API, Anthropic Claude, or similar LLM APIs


- Proficiency in Python (FastAPI, LangChain, LlamaIndex preferred)


- Strong understanding of prompt engineering techniques and best practices


- Experience with vector databases (Pinecone, Weaviate, Qdrant, ChromaDB)


- Knowledge of RAG (Retrieval Augmented Generation) implementation


- Understanding of transformer architecture and attention mechanisms


- Experience with API integration, webhooks, and streaming responses


- Strong problem-solving skills and ability to debug complex AI systems


- Experience with LangChain, LlamaIndex, or similar LLM frameworks


- Knowledge of fine-tuning techniques (LoRA, QLoRA, PEFT)


- Experience with embedding models and semantic search


- Familiarity with HuggingFace Transformers library


- Experience deploying models using vLLM, TGI (Text Generation Inference)


- Knowledge of function calling/tool use with LLMs


- Experience with model evaluation metrics (BLEU, ROUGE, BERTScore)


- Understanding of token economics and cost optimisation


- Experience with open-source models (Llama, Mistral, Falcon)


- Knowledge of model quantization and optimization techniques


- Experience with multi-modal models (vision, audio)


- Familiarity with MLOps practices and experiment tracking (Weights & Biases, MLflow)


- Experience with AWS SageMaker, Google Vertex AI, or Azure ML


- Understanding of chain-of-thought prompting, ReAct, agents


- Experience building chatbots or conversational AI systems


- Publications or contributions to AI/ML community


Skills

Generative AI, LLM, Python, LangChain, Prompt Engineering, RAG, AI Integration, VectorDB, Cloud, Artificial Intelligence

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,437,834 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Hyderabad
≈ $150k – $332k per year (Estimated) • In office • Full-Time • San Francisco
Apply
$108k per year • In office • Part-Time • Zurich
Python
AI/ML
Recommender Systems
Chips/EDA
PoC Library
Apply
AI Agent Developer 1 hour ago
≈ $96k – $216k per year (Estimated) • In office • Contractor • Bachelor's Degree • Wallisellen
AI/ML
Copilot
AI Agents
LLM
Apply
Remote (Germany) • Part-Time
AI/ML
LangChain
Management
n8n
Apply
≈ $101k – $227k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Zurich
AI/ML
LangGraph
LangChain
NLP
LLM
RAG
Multi-Agent Systems
Cybersecurity
GDPR
Management
Agile
Apply
≈ $23k – $51k per year (Estimated) • Hybrid • Full-Time • 2+ years exp • Hyderabad
Python
SQL
Databases
Snowflake
AI/ML
Copilot
Cursor
LangGraph
LangChain
Claude
Claude Code
LoRA
Model Context Protocol
Fine-tuning
AI Agents
Cline
Langfuse
LangSmith
PEFT
QLoRA
Llama
AWS Bedrock
Gemini
LLM
RAG
TruLens
Hallucination
Hybrid Search
OpenAI
AWS Bedrock AgentCore
A2A
OCR
Human-in-the-Loop
LLM Guardrails
Copilot Studio
DevOps
Rest API
GCP
Azure
CI/CD
AWS
Management
Power Automate
Apply
≈ $28k – $116k per year (Estimated) • Hybrid • Full-Time • Bengaluru • Hyderabad
Python
SQL
Python
FastAPI
Pydantic
Databases
SAP HANA
Snowflake
Pinecone
FAISS
OpenSearch
AI/ML
Copilot
Cursor
LangGraph
LangChain
Claude
Claude Code
Model Context Protocol
Embeddings
AI Agents
Cline
DeepEval
Promptfoo
Llama
AWS Bedrock
Gemini
LLM
RAG
Hallucination
OpenAI
AWS Bedrock AgentCore
A2A
OCR
LLM Guardrails
Agentic Workflows
Copilot Studio
DevOps
GCP
Azure
AWS
Management
Power Automate
Apply
Management Trainee 9 hours ago
≈ $7k – $14k per year (Estimated) • In office • Hyderabad
Analytics
Microsoft Excel
Apply
≈ $8k – $18k per year (Estimated) • In office • Full-Time • 4+ years exp • Hyderabad
Cybersecurity
CAPA
Apply
In office • Internship • Hyderabad
Marketing
X (Twitter)
Instagram
Apply
See all jobs
This is one of many
1,437,834 more open roles from verified company boards, updated every day.