368,746open jobs
9,444companies
47,506added this week
Browse all
Salary
$200k – $400k per year
Location
In office (San Francisco, New York)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Decagon is an enterprise AI company that builds autonomous customer service agents. Founded by Jesse Zhang and Ashwin Sreenivas and headquartered in San Francisco, California, the platform leverages large language models (LLMs) and advanced AI orchestration to automate complex customer support workflows end-to-end.

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.

Our technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that power personalized, deeply satisfying interactions across voice, chat, email, SMS, and every other channel.

We’re building a future where customer experiences are being redefined from support tickets and hold music to faster resolutions, richer conversations, and deeper relationships. We’re proud to be backed by world-class investors who share that vision, including a16z, Accel, Bain Capital Ventures, Coatue, and Index Ventures, along with many others.

We’re an in-office company, driven by a shared commitment to excellence and velocity. Our values - Just Get It Done, Invent What Customers Want, Winner’s Mindset, and The Polymath Principle - shape how we work and grow as a team.

About the Team

The Infrastructure team builds and operates the foundations that power Decagon: networking, data, ML serving, developer platform, and real-time voice. We partner closely with product, data, and ML to deliver high-scale, low-latency systems with clear SLOs and great developer ergonomics.

Read more about the infra team's work here: https://decagon.ai/blog/what-an-air-gapped-ai-deployment-actually-requires

About the Role

Decagon builds agentic AI that resolves customer support conversations end to end, for companies ranging from fast-growing startups to some of the largest financial institutions in the world. Keeping that system fast, reliable, and secure (across our multi-tenant cloud and inside customers' own locked-down environments) is an infrastructure problem, and that's the problem this role owns.

You'll build the platforms and abstractions our product teams ship on, and you'll architect and operate the deployments that run our agents inside enterprise customer clouds, where security, compliance, and operational rigor matter as much as speed. The work is core infrastructure at heart: reliability, CI/CD, deployment automation, on-call. You'll be joining a team with real infrastructure and momentum already in place, but there's no established playbook for running agentic systems at this scale, so a big part of the job is figuring it out as the technology shifts and usage grows by orders of magnitude.

What you'll do

  • Build the platform

    • Design the development and production platforms that power our products, and the abstractions over cloud infrastructure, Kubernetes, and networking that let engineers ship without becoming infrastructure experts.

    • Make sure it all scales to the next order of magnitude as usage grows.

  • Own enterprise deployments

    • Take end-to-end ownership of deployment architecture in customer-owned cloud environments (VPC configuration, permissioning, networking, provisioning) and the full lifecycle that follows: setup, upgrades, scaling, and incident support.

    • Build the runbooks and automation that make it repeatable.

  • Keep agentic workloads reliable

    • Treat monitoring, alerting, and rollback as first-class parts of anything you ship, not afterthoughts.

    • Own the reliability of the systems our AI agents depend on in production, where latency, availability, and graceful degradation directly shape the customer experience.

  • Partner across boundaries

    • Work directly with customers' platform, security, and DevOps teams to navigate their infrastructure and compliance constraints, and with our Product, Security, Sales, and Customer Success teams to turn customer requirements into concrete deployment plans.

Your background looks something like this

  • 4+ years building and operating core infrastructure, platform engineering, or infrastructure/DevOps, ideally with some customer-facing deployment experience.

  • Deep experience with a major cloud provider (GCP, AWS, or Azure), along with Terraform and Kubernetes at scale.

  • Strong grasp of cloud networking fundamentals (VPCs, IAM, DNS, load balancing) and how they surface as real deployment constraints.

  • A track record operating production systems reliably: monitoring, on-call, incident response, and reasoning about failure modes up front.

  • Comfort navigating ambiguity across a range of stakeholders, from engineers to security and compliance teams, and turning those conversations into actionable plans.

  • Clear technical writing and a track record of driving adoption across teams.

  • Comfortable in a fast-moving environment with rapid change.

Even better if you have

  • Experience managing deployments in customer-owned cloud environments, including security reviews, compliance requirements, and change management.

  • Experience building internal platforms or paved roads: service templates, self-serve environments, CI/CD pipeline design, deployment automation.

  • Familiarity with observability and incident management in distributed systems (Prometheus, Grafana, Datadog, or similar).

  • Infrastructure-as-code with a security-minded approach to supply chain (provenance, secrets, least privilege).

  • Experience operating latency-sensitive or ML/AI-serving workloads in production.

  • Experience using AI-assisted tooling to make yourself and your team dramatically more effective.

Compensation

$200K - $400K + Offers Equity

This range reflects the expected compensation for this role. Compensation within the range is determined based on experience, skills, and the scope of responsibilities, with flexibility for candidates who demonstrate exceptional impact.

In addition to base salary, we offer competitive equity. Final compensation may vary based on location within the United States.

Benefits

We proudly offer the following benefits for our full-time employees:

  • Medical, Dental, and Vision benefits for you and your family

  • Life Insurance and Disability Benefits

  • Retirement Plan (e.g., 401K, pension)

  • Parental Leave

  • Fertility and family building benefits through Carrot

  • Monthly stipend to support your wellness, lifestyle, and work-life balance

  • Daily lunches and snacks in the office to keep you at your best

  • Take what you need vacation policy (subject to local requirements; UK employees receive 25 days of statutory leave)

These benefits are described in more detail in Decagon’s policies, may vary by location, and can change at any time according to applicable compensation and benefits plans.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$68k – $85k per year • In office • Full-Time • Master's Degree • San Jose
Python
AI/ML
AI Agents
DevOps
Amazon EC2
AWS
AWS Lambda
Bitbucket
CI/CD
CloudFormation
Docker
Git
Kubernetes
Terraform
Amazon S3
IAM
HPC
Cybersecurity
Least Privilege
Apply
$185k – $260k per year • Remote • Full-Time • 8+ years exp • Bachelor's Degree
DevOps
AWS
CI/CD
GCP
Kubernetes
GitHub
Cybersecurity
Clair
Dependabot
OWASP Top 10
OWASP ZAP
Snyk
Trivy
Apply
$19k – $47k per year (Estimated) • Remote • Full-Time • Perm
C#
C#
ASP.NET Core
Dapper
Entity Framework Core
Databases
Apache Kafka
ClickHouse
ElasticSearch
PostgreSQL
RabbitMQ
Redis
DevOps
CI/CD
Docker
Docker Compose
GitHub Actions
Kubernetes
TeamCity
GitHub
GitLab
Apply
$17k – $43k per year (Estimated) • In office • Full-Time • Tomsk
C#
Java
Python
DevOps
CI/CD
Docker
Git
Grafana
Graylog
Kubernetes
Prometheus
Splunk
Apply
$19k – $47k per year (Estimated) • Remote • Full-Time • Tomsk
C#
C#
ASP.NET Core
Dapper
Entity Framework Core
Databases
Apache Kafka
ClickHouse
ElasticSearch
PostgreSQL
RabbitMQ
Redis
DevOps
CI/CD
Docker
Docker Compose
GitHub Actions
Kubernetes
TeamCity
GitHub
GitLab
Apply
$180k – $225k per year • In office • Full-Time • 5+ years exp • San Francisco
AI/ML
AI Agents
LLM
Design
Webflow
Marketing
Salesforce
Apply
$212k – $265k per year • In office • Full-Time • 6+ years exp • San Francisco
AI/ML
AI Agents
Claude
Claude Code
Cursor
Design
Figma
Apply
$200k – $400k per year • In office • Full-Time • 5+ years exp • San Francisco • New York
Databases
Amazon Redshift
Apache Kafka
ClickHouse
Databricks
Google BigQuery
RabbitMQ
Snowflake
AI/ML
AI Agents
Dagster
dbt
Flink
LLM
Prefect
DevOps
Amazon EKS
AWS
Azure
Azure AKS
CI/CD
Datadog
GCP
GitOps
Google GKE
Grafana
Kubernetes
OpenTelemetry
Prometheus
Terraform
Apply
Agent Data Scientist 11 days ago
$165k – $215k per year • In office • Full-Time • 5+ years exp • San Francisco
Python
SQL
AI/ML
AI Agents
LLM
Apply
$160k – $200k per year • In office • Full-Time • 6+ years exp • New York
AI/ML
AI Agents
Apply
$170k – $220k per year • Equity 1–2.8% • In office • Full-Time • 3+ years exp • San Francisco
Python
SQL
Python
Django
AI/ML
AI Agents
Context Engineering
LLM
LLM Evaluation
RAG
Apply
$173k – $314k per year • In office • Full-Time • 12+ years exp • Bachelor's Degree • San Francisco
Apex
JavaScript
Node JS
Python
SQL
TypeScript
Apex
Lightning Web Components
AI/ML
Agentforce
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
RAG
DevOps
AWS
Azure
CI/CD
Docker
GCP
GitHub
Grafana
gRPC
Kubernetes
New Relic
Prometheus
Splunk
Marketing
Salesforce
QA
Cypress
JMeter
k6
Locust
Playwright
Postman
Rest-Assured
Selenium
Apply
Senior ML Engineer 2 hours ago
$149k – $224k per year • In office • Full-Time • 5+ years exp • Master's Degree • San Francisco • Washington • Palo Alto
Python
Python
pySpark
Databases
Apache Kafka
AI/ML
AI Agents
Agentforce
Airflow
Anomaly Detection
Feature Store
Flink
Ray
Red Teaming
Spark
DevOps
CI/CD
Docker
Kubernetes
Cybersecurity
MITRE ATT&CK
Marketing
Salesforce
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • San Francisco
Go
JavaScript
Ruby
Scala
Apply
$360k – $530k per year • In office • Full-Time • Bachelor's Degree • San Francisco
MATLAB
Python
MATLAB
Simulink
AI/ML
OpenAI
Robotics
Digital Twin
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.