368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$30k – $70k per year (Estimated)
Location
In office (Bengaluru)
Seniority
Staff · 10+ years exp
Overview
Company
Impact
Profile match
Sequoia Group is an innovative brokerage firm providing strategic benefits, HR, and risk management services to hyper-growth companies. Based in the US, it has over 1000 corporate clients internationally.

We are seeking a highly experienced and motivated Lead DevOps-AI Platform Engineer to join our technology organisation. This role is responsible for ensuring the reliability, security, scalability, and operational excellence of our AWS-based infrastructure while also supporting next-generation agentic AI platform capabilities. The ideal candidate will have deep expertise in Linux, AWS, automation, CI/CD, troubleshooting, and production operations, along with hands-on experience in multi-agent frameworks, retrieval-augmented generation, observability, evaluation, and AI infrastructure tooling. This position requires a strong blend of infrastructure engineering, DevOps leadership, platform automation, and familiarity with modern AI agent orchestration and governance technologies.

Responsibilities:

  • Ensure the organisation's AWS infrastructure is available, secure, and operating reliably 24x7
  • Design, build, maintain, and optimise cloud infrastructure across load balancers, containers, databases, and related platform services.
  • Lead infrastructure migration, deployment, and launch activities, ensuring smooth transitions and minimal disruption to business operations.
  • Partner closely with DevOps, Security, and Engineering teams to maintain application integrity, infrastructure security, and operational resilience.
  • Build and maintain deployment tooling, operational services, and microservices that improve efficiency, reduce manual effort, and enhance customer service standards.
  • Troubleshoot and resolve incidents across development, testing, staging, and production environments in a timely and proactive manner.
  • Validate system integrity, infrastructure designs, application deployments, and operational processes, recommending improvements where appropriate.
  • Develop, update, and document operational procedures, technical workflows, and platform standards.
  • Automate operational and infrastructure processes with an emphasis on accuracy, repeatability, compliance, and security.
  • Specify, document, and support the development of new platform features, scripts, and operational capabilities.
  • Manage code deployments, releases, fixes, updates, and related operational processes.
  • Work with open-source technologies, cloud services, and modern CI/CD tooling to support reliable delivery pipelines.
  • Collaborate effectively with teams using Git, Agile methodologies, and workflow management tools such as Jira, Workfront, Scrum, Kanban, or SAFe.
  • Support the design, implementation, and operationalisation of multi-agent AI systems and related enterprise workflows.
  • Build integrations for autonomous agents that interact with external APIs, databases, and enterprise systems through function calling and secure tool usage.
  • Contribute to the orchestration of multi-agent workflows using frameworks such as CrewAI, LlamaIndex, LangGraph, AutoGen, Semantic Kernel, Haystack, and DSPy.
  • Support retrieval-augmented generation (RAG) systems and scalable vector database architectures, including Qdrant, Chroma, pgvector, Pinecone, and Weaviate.
  • Implement observability, tracing, and evaluation for AI and agentic systems using tools such as LangSmith, Langfuse, Arize AI, Phoenix by Arize, Helicone, and MLflow.
  • Support Model Context Protocol (MCP) integrations to enable secure and standardised agent connectivity.
  • Define and enforce agent permissions, guardrails, and governance controls for autonomous systems.
  • Support AI gateway and routing solutions such as Portkey, LiteLLM, and Envoy AI Gateway.
  • Contribute to GPU scheduling and infrastructure planning for model-serving and agentic workloads.
  • Promote operational best practices for autonomous systems management, reliability, evaluation, and continuous improvement.

Requirements:

  • 10+ years of experience as a DevOps Engineer, Platform Engineer, Infrastructure Engineer, or in a similar role, with additional software or infrastructure development experience preferred.
  • Strong experience with Linux-based infrastructures, Linux/Unix administration, and AWS.
  • Strong experience with databases and data platforms such as SQL, MySQL, NoSQL, Elasticsearch, Redis, and MongoDB.
  • Proficiency in one or more scripting or programming languages such as Java, JavaScript, Perl, Ruby, Python, Golang, PHP, Groovy, or Bash.
  • Experience with CI/CD tools, Git, deployment automation, and release engineering practices.
  • Experience with open-source technologies and cloud services.
  • Experience with configuration management and automation tools such as Puppet or Chef.
  • Familiarity with Agile delivery methodologies and workflow tools such as Jira, Workfront, Scrum, Kanban, or SAFe.
  • Strong communication skills, with the ability to explain technical protocols, processes, and decisions to both technical and non-technical stakeholders.
  • Strong troubleshooting and analytical skills, with the ability to identify issues before they become problems.
  • Demonstrated ability to stay current with industry trends, IT operations best practices, and emerging platform technologies.
  • Experience with agentic AI platforms, multi-agent frameworks, observability, RAG, and related infrastructure is strongly preferred.
  • A bachelor's or master's degree in computer science, engineering, software engineering, or a related field.

Preferred Qualifications:

  • Experience with multi-agent orchestration frameworks such as CrewAI, LangGraph, AutoGen, or Semantic Kernel.
  • Experience with AI observability and evaluation platforms such as LangSmith, Langfuse, Arize AI, Phoenix by Arize, Helicone, or MLflow.
  • Experience with Model Context Protocol (MCP) and secure agent-tool integration patterns.
  • Experience with vector databases and retrieval architectures used in RAG-based applications.
  • Experience with AI gateways, policy enforcement, and operational guardrails for agentic systems.
  • Exposure to GPU scheduling, distributed inference, or AI platform operations.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Bengaluru
$25k – $42k per year • Equity 0–0.2% • Remote • Full-Time • 3+ years exp
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
$100k – $210k per year • Equity 0–0.5% • Remote • Full-Time • 3+ years exp • San Francisco
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
Team Lead DevOps 1 day ago
$23k – $62k per year (Estimated) • Remote • 5+ years exp • Moscow
Bash
Python
Erlang
Erlang
EMQX
Databases
Apache Kafka
ClickHouse
PostgreSQL
RabbitMQ
Redis
Redpanda
Trino
DevOps
Ansible
AWS
AWX
FinOps
HAProxy
Hetzner
Kubernetes
SLI/SLO/SLA
Terraform
Yandex Cloud
Amazon S3
Apply
$100k – $200k per year • Equity 0.5–5% • In office • Full-Time • 1+ year exp • New York
Python
TypeScript
JavaScript
Python
FastAPI
Databases
DynamoDB
PostgreSQL
AI/ML
Claude
LLM
OpenAI
AI Agents
Frontend
Next.js
Tailwind CSS
React.js
DevOps
AWS
Docker
Vercel
GitHub
Management
Slack
Apply
$120k – $180k per year • Equity 0.2–1.5% • In office • Full-Time • 1+ year exp • San Francisco
Python
AI/ML
AI Agents
ChatGPT
DevOps
Error Budget
Apply
$30k – $59k per year (Estimated) • In office • 5+ years exp • Bachelor's Degree • Bengaluru
C++
Java
Python
SQL
Databases
Aerospike
DynamoDB
MySQL
Redis
DevOps
SLI/SLO/SLA
Terraform
Analytics
ETL/ELT
Management
Jira
Apply
$21k – $67k per year (Estimated) • In office • 4+ years exp • Bachelor's Degree • Bengaluru
JavaScript
Frontend
React.js
Rspack
Webpack
DevOps
Git
Design
Figma
Apply
Staff Data Scientist 1 month ago
$33k – $80k per year (Estimated) • In office • Bengaluru
Python
AI/ML
NLP
Apply
$30k – $76k per year (Estimated) • In office • 7+ years exp • Bengaluru
Go
Java
Python
TypeScript
Python
Django
AI/ML
AI Agents
AutoGen
CrewAI
Embeddings
LangChain
LlamaIndex
Prompt Engineering
RAG
Semantic Search
Anthropic
LLM Guardrails
OpenAI
Semantic Search
DevOps
AWS
Azure
GCP
Vector
Apply
$22k – $59k per year (Estimated) • In office • 4+ years exp • Bengaluru
Go
Java
Python
TypeScript
AI/ML
AI Agents
AutoGen
CrewAI
Embeddings
LangChain
LlamaIndex
Prompt Engineering
RAG
Semantic Search
Anthropic
LLM Guardrails
OpenAI
Semantic Search
DevOps
AWS
Azure
GCP
Vector
Apply
$31k – $82k per year (Estimated) • In office • Full-Time • 3+ years exp • Hyderabad • Bengaluru
Apply
$31k – $73k per year (Estimated) • In office • Full-Time • 5+ years exp • Bengaluru
Apply
$16k – $34k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Mumbai • Bengaluru
JavaScript
PowerShell
SQL
C#
C#
.NET
Databases
Azure SQL Database
MS SQL
DevOps
Azure
Rest API
Cybersecurity
Microsoft Entra ID
QA
Postman
Swagger
Apply
$41k – $89k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Bengaluru
C#
TypeScript
JavaScript
C#
.NET
Databases
Apache Kafka
AI/ML
Copilot
LLM
OpenAI
Frontend
Angular
GraphQL
DevOps
Azure
Azure AKS
Azure DevOps
CI/CD
Docker
GitHub
GitHub Actions
Grafana
Kubernetes
Prometheus
Rest API
Apply
$38k – $83k per year (Estimated) • In office • Full-Time • 12+ years exp • Bachelor's Degree • Bengaluru
Databases
Oracle
DevOps
AWS
Platform Engineering
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.