881,486open jobs
54,986companies
148,057added this week
Browse all
Salary
≈ $26k – $57k per year (Estimated)
Location
Hybrid (Bengaluru, India)
Seniority
Staff · 5+ years exp

First seen by Alion on Sep 22, 2026.

Overview
Company
Impact
Profile match
SPi Global, formed in 1980 deals in the fields of Data Intelligence, Knowledge, Analytics, UI/UX/App Dev, Digitization, Marketing & Support Solutions, Integrated Workflow Solutions, E-learning, Content/Design Dev.

Principal AI Architect - Generative AI & Enterprise Delivery.

About Us :


Straive is a global leader in data analytics and AI operationalization, helping enterprises embed advanced AI and data capabilities into core business workflows to deliver measurable business outcomes and ROI.

With a workforce of ~20,000 professionals serving 350+ clients across 30+ markets, Straive combines technical scale with deep domain expertise.

A key differentiator is its network of over 6,000 subject matter experts who specialize in managing and enriching complex, unstructured data enabling organizations to build AI systems grounded in accuracy, context, and business relevance.

The company continues to earn recognition from leading industry analysts and was recently named a Leader in AIM's 2026 Generative AI and Data Engineering PeMa Quadrants.

Backed by EQT - a purpose-driven global investment organization which was ranked among the world's leading private equity firms by PEI in 2025 - Straive is positioned as a high-value alternative to traditional IT services providers, combining domain-led intelligence with AI execution at scale.

About the Role :


Straive is seeking a hands-on Principal AI Architect & Senior Developer to serve as the technical authority for our enterprise Generative AI delivery practice.

In this role, you will lead the end-to-end architectural design, backend engineering, and production deployment of enterprise Generative AI platforms, multi-agent systems, document intelligence pipelines, and real-time streaming LLM applications.

This is an intensely hands-on role requiring expert backend proficiency in Python and SQL alongside deep expertise in Generative AI architectures.

You will move beyond simple API orchestration to fine-tune open-weight models, build custom agentic workflows, engineer advanced multimodal RAG pipelines, write high-concurrency streaming backend microservices (FastAPI/AsyncIO), enforce AI security/PII guardrails, and optimize LLM unit economics (inference cost, latency, accuracy) for production deployments.

Key Responsibilities :


- Hands-on Gen AI Architecture & System Design : Architect scalable enterprise Gen AI solutions - including multi-agent orchestration (LangGraph or AutoGen or CrewAI), advanced RAG retrieval pipelines, real-time streaming interfaces (SSE, WebSockets), prompt security, guardrails, and automated evaluation suites.

- High-Concurrency Backend Engineering & SQL : Write clean, async, high-performance Python backend code (FastAPI or AsyncIO or PyDantic) and write complex, optimized SQL queries to build robust microservices, caching layers (Redis), queueing systems (RabbitMQ/Celery), and vector search integrations.

- Open-Weight Fine-Tuning & AI Economics : Own open-weight model fine-tuning (Llama or Mistral or Qwen or Phi) using LoRA, QLoRA, and full tuning.

- Inference Optimization : Drive inference optimization via quantization (AWQ, GPTQ), model distillation, prompt caching, and high-throughput GPU serving (vLLM/Triton).

- Agentic Workflows & Tool Integration (MCP) : Architect autonomous multi-agent workflows, planner-worker execution patterns, and custom Model Context Protocol (MCP) servers connecting LLMs securely to enterprise databases, APIs, and legacy systems of record.

- Multimodal RAG & Document Intelligence : Build complex document intelligence and multimodal ingestion pipelines to parse, OCR, chunk, and embed structured/unstructured documents (PDFs, tables, images) using Textract, Unstructured, and Vision-Language Models (VLMs).

- LLMOps, Telemetry, Guardrails & Governance : Establish end-to-end LLMOps infrastructure - automated CI/CD eval gates (Ragas, TruLens), hallucination detection, telemetry/tracing (LangSmith, OpenTelemetry, Arize), and AI safety controls (NeMo Guardrails, Llama Guard, PII redaction).

- Technical Leadership, RFPs & Client Advisory : Lead and mentor 30+ AI engineers and developers through architecture reviews and hands-on code reviews.

- Commercial Collaboration : Partner with client CTOs, CDOs, and commercial teams on RFPs, technical proposals, discovery workshops, and effort estimations.

Required Experience and Qualifications :


- 5-8 years of professional software engineering experience across system design, distributed backend development, and AI/ML, with a minimum of 4+ years dedicated to hands-on Generative AI and LLM application architecture.

- Expert-level hands-on proficiency in Python (FastAPI or AsyncIO or PyDantic or Celery) and SQL (PostgreSQL/pgvector, MySQL, complex joins, CTEs, query optimization) for building production microservices.

- Proven track record architecting and shipping production Gen AI systems using multi-agent frameworks (LangGraph or AutoGen or CrewAI or LangChain) and advanced RAG architectures (hybrid semantic + BM25 retrieval/re-ranking models).

- Hands-on experience fine-tuning open-weight LLMs/SLMs (Llama/Mistral/Qwen/Phi), dataset curation, LoRA/QLoRA execution, and serving models under production load using vLLM, Triton, or Ollama.

- Strong practical depth with Vector Databases (OpenSearch Vector Search/ Qdrant/ Milvus/ FAISS/ Pinecone) and document parsing/multimodal pipelines (AWS Textract, OCR, Vision-Language Models).

- Demonstrated ability to optimize LLM inference cost and latency via quantization (AWQ/GPTQ, GGUF), prompt engineering/caching, model routing, and batching.

- Hands-on expertise in LLMOps telemetry and AI security : evaluation frameworks (Ragas/TruLens), tracing (LangSmith/OpenTelemetry), prompt injection protection, PII masking, and guardrails (NeMo Guardrails/Llama Guard).

- Solid command of containerization (Docker, Kubernetes), streaming APIs (SSE, WebSockets), GPU serving, and cloud AI platforms across either AWS (Bedrock/SageMaker), Azure (Azure OpenAI), or GCP (Vertex AI).

Preferred Experience & Education :


- Hands-on exposure to Model Context Protocol (MCP) servers, Claude Skills, and agentic tool-use protocols.

- Working command of enterprise AI risk and compliance frameworks.

- Familiarity with enterprise data platforms (Databricks MLflow/Snowflake) or classical ML (XGBoost, scikit-learn).

- Bachelor's or Master's degree in Computer Science, Software Engineering, Artificial Intelligence, or a related quantitative field from Tier 1/Tier 2 colleges.

Skills

Artificial Intelligence, Generative AI, Enterprise AI Systems, Python, System Design, Distributed Systems, AI Integration, FastAPI, VectorDB

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
881,486 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Backend
Similar stack
Same company
Bengaluru
≈ $20k – $42k per year (Estimated) • In office • 1+ year exp • Moscow
Python
JavaScript
SQL
Databases
PostgreSQL
Frontend
JQuery
DevOps
Ubuntu
Management
ITSM
Apply
≈ $75k – $180k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Peterborough
Apply
≈ $75k – $180k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Waterlooville
Python
SQL
Apply
Architect Java 1 hour ago
≈ $82k – $127k per year (Estimated) • In office • Full-Time • Warsaw
Java
Management
Confluence
Jira
SharePoint
Apply
≈ $23k – $56k per year (Estimated) • Hybrid • 3+ years exp • Moscow
Python
AI/ML
LangGraph
LangChain
AI Agents
LLM
OpenAI
DevOps
Docker Compose
CI/CD
Docker
Apply
In office
Python
SQL
Analytics
ETL/ELT
Apply
≈ $127k – $277k per year (Estimated) • In office • PhD • Novato
Python
Python
Flask
AI/ML
Embeddings
DevOps
Linux
Windows
Apply
$80k – $130k per year • In office • PhD • Novato
Python
AI/ML
Machine Learning
Apply
≈ $59k – $141k per year (Estimated) • In office • Full-Time • 12+ years exp • Bachelor's Degree • Seoul
SQL
AI/ML
Copilot
AI Agents
Google AI Studio
LLM Guardrails
DevOps
Rest API
GCP
Azure
AWS
Cybersecurity
GDPR
HIPAA
Analytics
Tableau
Power BI
Apply
QA Team Lead 3 hours ago
In office • Full-Time • Limassol
SQL
Databases
Apache Kafka
DevOps
Rest API
CI/CD
Git
Management
Confluence
Jira
Agile
Scrum
QA
TestRail
Selenium
Cypress
Playwright
Postman
Apply
≈ $15k – $38k per year (Estimated) • In office • 6+ years exp • Mumbai
Java
SQL
Java
Spring Boot
Databases
Snowflake
DevOps
Rest API
CI/CD
AWS
Docker
Kubernetes
Apply
≈ $9k – $27k per year (Estimated) • In office • 6+ years exp • Pune
JavaScript
TypeScript
Databases
Oracle
Frontend
Angular
React.js
Apply
Sr. Java Developer 7 days ago
≈ $15k – $38k per year (Estimated) • In office • Mumbai
Java
SQL
Java
Spring Boot
Databases
Snowflake
AI/ML
AI Agents
DevOps
Rest API
CI/CD
AWS
Docker
Kubernetes
Apply
≈ $15k – $39k per year (Estimated) • In office • Mumbai
Python
Java
SQL
Databases
Snowflake
AI/ML
AI Agents
DevOps
Rest API
CI/CD
AWS
Docker
Kubernetes
Apply
≈ $17k – $45k per year (Estimated) • In office • Bengaluru
Python
Java
SQL
Databases
Snowflake
AI/ML
AI Agents
DevOps
Rest API
CI/CD
AWS
Docker
Kubernetes
Apply
≈ $18k – $46k per year (Estimated) • In office • 6+ years exp • Bengaluru
Python
AI/ML
LLM
Agentic Workflows
DevOps
Rest API
gRPC
Docker
Kubernetes
Platform Engineering
Apply
≈ $17k – $37k per year (Estimated) • In office • Full-Time • 7+ years exp • Bachelor's Degree • Pune • Chennai • Bengaluru
JavaScript
Java
TypeScript
SQL
Java
Maven
Spring Boot
Databases
MySQL
PostgreSQL
Oracle
Frontend
React.js
DevOps
Rest API
CI/CD
Git
AWS
Docker
Kubernetes
Management
Agile
Scrum
Apply
≈ $30k – $75k per year (Estimated) • In office • 7+ years exp • Bengaluru
JavaScript
TypeScript
SQL
Databases
Databricks
Azure Cosmos DB
Frontend
Angular
React.js
Mobile
React Native
Clean Architecture
DevOps
Terraform
Azure DevOps
GitHub Actions
OpenTelemetry
Datadog
FluxCD
Prometheus
Azure
CI/CD
GitOps
ArgoCD
Docker
Kubernetes
Blue-Green Deployment
Chaos Engineering
Azure AKS
Cryptography
Vault
Apply
$10k – $26k per year • Equity 0–2% • In office • Full-Time • 1+ year exp • Bengaluru
Python
JavaScript
Python
Django
AI/ML
Claude Code
AI Agents
OpenAI Codex
Agentic Workflows
Frontend
React.js
DevOps
Azure
AWS
Docker
Apply
AI / ML Engineer 1 day ago
≈ $23k – $47k per year (Estimated) • In office • 5+ years exp • Bengaluru
AI/ML
LangGraph
LangChain
Claude
AI Agents
AutoGPT
TensorFlow
CrewAI
Gemini
LLM
LLMOps
GPT-4
LLM Guardrails
Multi-Agent Systems
Machine Learning
DevOps
GCP
Azure
CI/CD
AWS
Apply
See all jobs
This is one of many
881,486 more open roles from verified company boards, updated every day.