368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$200k – $260k per year
Location
In office (San Francisco)
Seniority
Senior · 1+ year exp
Employment
Full-Time
Overview
Company
Impact
Profile match
LiteLLM is the open-source AI gateway that puts your full AI stack behind one OpenAI-compatible key. Track and cap LLM spend, route to the right model, and self-host anywhere — even air-gapped. 140+ providers, 1,892 models.

Senior Backend Engineer

LiteLLM is the world's most popular AI Gateway, trusted by top companies like Adobe, Netflix, and NASA. Our platform empowers developers by providing secure, reliable access to LLMs and adjacent services, and we're looking for a Senior Backend Engineer to help us build rock-solid guardrails and observability tooling at scale.

About The Role

You’ll focus on owning our guardrails and logging world-class. You will be in charge of the backend code that ensures all guardrail calls are consistently logged, errors are surfaced to users (not silently swallowed), and our observability instrumentation works for real-world, high-volume traffic. Your attention to detail in areas like latency metrics, logging traceability, and backend guardrail registration will directly impact user trust in our security and compliance features.

Responsibilities

  • Build and scale our product, ensuring performance, reliability, and continuous improvement.

  • Ensure all guardrail and policy enforcement calls (e.g., applyguardrail) are properly logged and traceable through our SpendLogs and relevant database tables

  • Build and design CPU-level guardrails to cover common attacks on LLM API's / MCP servers / Agents

  • Identify and fix areas where silent failures occur in guardrail creation, registration, and policy application-ensuring robust error handling and transparency to end users

  • Work with observability integrations, including Datadog, Splunk, Prometheus, and OpenTelemetry, to maintain accurate, configurable, and usable monitoring and logging for backend systems

  • Enhance observability integrations to work for 1B+ requests/mo., with minimal latency overhead and no memory leaks (e.g. due to cardinality of Prometheus metrics)

  • Collaborate cross-functionally on backend engineering priorities (performance, reliability, security)

What We’re Looking For

  • Bachelor’s or Master’s in Computer Science or related field

  • 1+ years of experience with Python and backend frameworks (e.g. FastAPI, Flask)

  • Understanding of logging best practices, error handling, and secure backend development

  • Exposure to monitoring, logging, or metrics platforms (Datadog, Splunk, Prometheus, OpenTelemetry)

  • Familiarity with database integration and troubleshooting (PostgreSQL, Redis, etc.)

  • Driven to deliver high-quality backend code with strong guardrails, auditing, and debugging capabilities

  • Eagerness to tackle hard bugs and ensure system transparency for end users

Why Join LiteLLM?

  • High-impact, mission-critical work on the core of compliance and reliability

  • Contribute directly to features used by enterprise customers at global scale

  • Fast-paced growth environment with room for technical ownership

  • Competitive salary, health, dental, and vision benefits

About LiteLLM

LiteLLM (https://github.com/BerriAI/litellm) is a Python SDK and Proxy Server enabling seamless calls to 100+ LLM APIs in the OpenAI format, trusted by industry leaders worldwide.

Ready to shape the future of secure, observable AI infrastructure? Apply now!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$17k – $49k per year (Estimated) • Remote/Hybrid • Kaliningrad
Node JS
Python
JavaScript
Node JS
BullMQ
Fastify
Python
FastAPI
Flask
Databases
Apache Kafka
Chroma
Pinecone
PostgreSQL
Qdrant
RabbitMQ
Redis
AI/ML
Claude
Copilot
Cursor
LangChain
LlamaIndex
LLM
Model Context Protocol
Prompt Engineering
RAG
Anthropic
OpenAI Codex
Structured Outputs
Function Calling
Frontend
Next.js
React.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
Gitflow
Rest API
WebSockets
GitHub
Management
Jira
Apply
$34k – $82k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru • Pune
Java
Python
Python
Asyncio
FastAPI
Databases
Amazon Aurora
AI/ML
AWS Bedrock
Google ADK
LangChain
LangGraph
LLM
RAG
A2A
LLM Guardrails
AI Agents
Model Context Protocol
DevOps
Amazon EKS
AWS
Envoy
Kubernetes
API Gateway
IAM
Cybersecurity
Zero Trust
Apply
NLP / LLM Engineer 12 hours ago
$18k – $24k per year (net) • In office • Full-Time • 3+ years exp • Tashkent
Python
Python
FastAPI
Databases
ElasticSearch
Milvus
Pinecone
Qdrant
Weaviate
AI/ML
AI Agents
ChatGPT
DeepEval
Embeddings
Gemini
Hybrid Search
LangChain
Langfuse
LangGraph
LlamaIndex
LLM
LoRA
NLP
PEFT
Prompt Engineering
PyTorch
QLoRA
RAG
Reranking
Semantic Search
Synthetic Data
Tokenization
Triton
vLLM
Transformers
Anthropic
DPO
GraphRAG
Hugging Face
OCR
OpenAI
Semantic Search
SFT
Structured Outputs
Function Calling
TGI
DevOps
CI/CD
Docker
Git
GitHub
Analytics
A/B Testing
Apply
$71k – $154k per year (Estimated) • In office • Full-Time • Dublin
Java
Databases
Apache Kafka
AI/ML
Copilot
LLM
LLM Guardrails
DevOps
AWS
CI/CD
Kubernetes
GitHub
Apply
Lead AI Engineer 11 hours ago
$30k – $73k per year (Estimated) • In office • Full-Time • Pune
Python
AI/ML
Fine-tuning
LLM
Reinforcement Learning
LLM Guardrails
AI Agents
DevOps
CI/CD
Docker
GitOps
Helm
Kubernetes
OpenShift
Platform Engineering
Vector
Apply
$50k – $90k per year • Remote • Full-Time
AI/ML
LiteLLM
Apply
$150k – $210k per year • In office • Full-Time
AI/ML
LiteLLM
Apply
$150k – $220k per year • In office • Full-Time • San Francisco
AI/ML
LiteLLM
Management
Slack
Apply
$150k – $250k per year • In office • Full-Time
AI/ML
LiteLLM
Management
Slack
Apply
AI Engineer 1 month ago
$200k – $250k per year • In office • Full-Time • 1+ year exp • San Francisco
Python
Python
FastAPI
Databases
PostgreSQL
Redis
AI/ML
LiteLLM
LLM
Anthropic
OpenAI
Model Context Protocol
Apply
$293k – $385k per year • In office • Full-Time • San Francisco
AI/ML
OpenAI
Apply
$180k – $260k per year • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco
Python
AI/ML
ChatGPT
OpenAI
OpenAI Codex
Cybersecurity
FedRAMP
Apply
$180k – $210k per year • Equity • In office • Full-Time • San Francisco
Node JS
JavaScript
Databases
PostgreSQL
DevOps
PagerDuty
Web3
TRM Labs
Management
Slack
Apply
$252k – $335k per year • Remote/Hybrid • Full-Time • 8+ years exp • San Francisco
AI/ML
ChatGPT
Human-in-the-Loop
OpenAI
OpenAI Codex
DevOps
SLI/SLO/SLA
Apply
$223k – $424k per year (Estimated) • In office • Bachelor's Degree • San Francisco
AI/ML
AI Agents
LLM
Recommender Systems
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.