368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$56k – $141k per year (Estimated)
Location
Remote (Argentina)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Jeeves is an all-in-one global financial operating platform and corporate card provider headquartered in Orlando, Florida. Founded in 2019 by Dileep Thazhmon and Sherwin Gandhi, the venture-backed fintech unicorn has raised over $380 million from investors including Andreessen Horowitz, Y Combinator, and CRV.

Jeeves is building AI into the core of its financial platform - from intelligent spend categorization and anomaly detection to LLM-powered workflows that help finance teams move faster. We're looking for a Senior AI Engineer who is obsessed with building AI systems that actually work in production: reliable, observable, cost-efficient, and genuinely useful.

This is not a research role. You will ship AI-powered features that process real financial data for real businesses. You'll work alongside backend engineers, data scientists, and product teams to take AI from prototype to production - and you'll help define how Jeeves builds with AI as the company scales.

If you've built LLM pipelines, designed RAG architectures, operated ML systems in production, and care deeply about what happens when your AI makes a wrong call in a financial context - we want to meet you.

Location: This is a full-time remote position. #LI-REMOTE

What You'll Do:

    LLM & AI Pipeline Engineering

  • Design, build, and maintain production-grade LLM integration pipelines - including retrieval-augmented generation (RAG), prompt engineering, output parsing, and chain orchestration.

  • Develop and operate AI features within Jeeves's core financial products: spend categorization, document extraction, anomaly detection, financial Q&A, and automated reconciliation.

  • Implement structured output validation, fallback handling, and confidence scoring to ensure AI decisions meet reliability standards for financial use cases.

  • Evaluate and integrate AI frameworks and tools (LangChain, LlamaIndex, OpenAI API, Anthropic API, HuggingFace, vector databases) and advocate for the right tool for the job.

  • Establish prompt versioning and evaluation practices to ensure AI outputs remain accurate and consistent as models and data evolve.

  • Retrieval & Vector Search

  • Design and maintain vector search pipelines using databases such as Pinecone, Weaviate, or pgvector to power semantic search and RAG-based features.

  • Build document ingestion and chunking pipelines for Jeeves's financial data - processing invoices, receipts, policy documents, and transaction records.

  • Optimize retrieval quality through embedding model selection, chunk strategy, metadata filtering, and re-ranking techniques.

  • ML Model Serving & Operations

  • Collaborate with data scientists to take trained ML models from experimental notebooks to production serving infrastructure.

  • Build and maintain model serving endpoints with appropriate latency SLOs, input validation, and output monitoring.

  • Implement model performance monitoring and data drift detection to ensure production models remain accurate over time.

  • Support model retraining workflows by designing clean data pipelines and feature engineering that can be continuously updated.

  • Backend Integration & Reliability

  • Integrate AI services cleanly with Jeeves's backend microservices - designing clear API contracts, circuit breakers, and graceful degradation patterns.

  • Write high-quality, testable backend code in Python or Go/Node.js to power AI-integrated features.

  • Instrument AI components with structured logging, distributed tracing, latency dashboards, and alerting to ensure operational visibility.

  • Build human-in-the-loop review workflows for AI decisions that require oversight - particularly for high-value financial actions.

  • Collaboration & Growth

  • Partner with Product, Backend Engineering, and Data Science to define the AI roadmap and translate requirements into reliable systems.

  • Contribute to a culture of quality by writing design docs, reviewing peers' AI system designs, and sharing learnings openly.

  • Help grow the AI engineering practice at Jeeves by establishing patterns, tooling, and best practices that the broader team can build on.

Requirements:

    Minimum Requirements

  • Bachelor's degree in Computer Science, Engineering, or a related field - or equivalent practical experience.

  • 5+ years of professional software engineering experience, with at least 3 years focused on AI/ML systems in production.

  • Hands-on experience building and deploying LLM-powered applications using APIs such as OpenAI, Anthropic, or Cohere in a production environment.

  • Experience designing and operating RAG pipelines, including chunking strategies, embedding models, and vector database integration (Pinecone, Weaviate, pgvector, or similar).

  • Strong proficiency in Python for AI/ML workloads; familiarity with at least one AI orchestration framework (LangChain, LlamaIndex, or equivalent).

  • Experience with ML model serving infrastructure: REST or gRPC inference endpoints, input/output validation, latency budgeting, and monitoring.

  • Solid backend engineering fundamentals: REST APIs, relational databases (PostgreSQL preferred), async patterns, and cloud infrastructure (AWS, GCP, or Azure).

  • Experience with observability tooling: structured logging, distributed tracing, and building dashboards for AI system health.

  • Preferred Qualifications

  • Experience in fintech, financial services, or any regulated industry where AI reliability and auditability are critical.

  • Familiarity with prompt evaluation frameworks, A/B testing AI outputs, and tracking model performance degradation in production.

  • Experience with ML lifecycle management tools: MLflow, Weights & Biases, Vertex AI, or SageMaker.

  • Knowledge of real-time data streaming (Kafka, Kinesis) for event-driven AI pipelines.

  • Contributions to open-source AI tooling, published technical writing, or talks at AI/ML conferences.

  • Prior startup or scale-up experience - comfortable with ambiguity and building foundational systems from scratch.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$230k – $260k per year • Equity • Remote • Internship • Bachelor's Degree
Python
DevOps
Amazon EKS
AWS
Azure
CI/CD
GCP
Helm
Kubernetes
Terraform
Cybersecurity
FedRAMP
Orca Security
Apply
$22k – $52k per year (Estimated) • In office • 8+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
DevOps
AWS
Azure
GCP
Analytics
Power BI
Tableau
ETL/ELT
Apply
$25k – $42k per year • Equity 0–0.2% • Remote • Full-Time • 3+ years exp
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
$100k – $210k per year • Equity 0–0.5% • Remote • Full-Time • 3+ years exp • San Francisco
Bash
Go
JavaScript
Python
TypeScript
DevOps
AWS
Azure
CI/CD
Datadog
Docker
GCP
GitHub Actions
GitLab CI
Grafana
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Amazon CloudWatch
GitHub
GitLab
IAM
Cybersecurity
Least Privilege
Apply
Team Lead DevOps 1 day ago
$23k – $62k per year (Estimated) • Remote • 5+ years exp • Moscow
Bash
Python
Erlang
Erlang
EMQX
Databases
Apache Kafka
ClickHouse
PostgreSQL
RabbitMQ
Redis
Redpanda
Trino
DevOps
Ansible
AWS
AWX
FinOps
HAProxy
Hetzner
Kubernetes
SLI/SLO/SLA
Terraform
Yandex Cloud
Amazon S3
Apply
$25k – $57k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • São Paulo
Python
SQL
AI/ML
Claude
Management
Slack
Marketing
HubSpot
Apply
$47k – $109k per year (Estimated) • Remote • Full-Time • 3+ years exp
SQL
Apply
Senior QA Engineer 1 month ago
$20k – $49k per year (Estimated) • Remote • Full-Time • 4+ years exp • India
SQL
QA
Cypress
Playwright
Apply
Senior Designer 1 month ago
$85k – $194k per year (Estimated) • Remote • Full-Time • 5+ years exp
AI/ML
LLM
Apply
Senior Designer 1 month ago
$50k – $113k per year (Estimated) • Remote • Full-Time • 5+ years exp
AI/ML
LLM
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.