405,710open jobs
14,091companies
78,515added this week
Browse all
Salary
$180k – $300k per year
Location
In office (New York)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Traversal is an AI site reliability engineer for large distributed systems, triaging alerts and finding root cause across petabyte-scale telemetry. It reconstructs the causal chain between a symptom and the change that caused it. The founders came from causal inference research in academia.

About Traversal

Traversal is the AI Site Reliability Engineer (AI SRE) for the enterprise.

Production complexity was already outpacing what engineering teams could manage manually, and AI-generated code is accelerating that gap. Traversal is built for that challenge, autonomously understanding and reasoning across even the largest, most complex production environments to diagnose, fix, and prevent incidents. Our mission is to free engineers from endless firefighting and give them more time to focus on creative, high-impact work.

Today, Traversal operates in mission-critical environments at some of the world’s largest enterprises. Our roots remain deeply embedded in AI research, and we’ve brought together researchers from institutions including MIT, Harvard, Berkeley, Columbia, and Cornell with world-class technical staff and operators from companies like Google, Meta, Datadog, ServiceNow, and Citadel Securities to take on one of the hardest problems for AI to solve. Traversal is backed by Sequoia Capital, Kleiner Perkins, Hanabi, NFDG, and American Express Ventures.

The Role

As an AI Research Engineer at Traversal, you'll build the systems our research team runs on. Our agents autonomously diagnose and resolve production incidents for some of the world's largest enterprises, and improving them is an experimental problem: form a hypothesis about agent behavior, test it against real incident data, and ship what works. The bottleneck is rarely ideas. It's the infrastructure to run the experiment, trust the result, and scale the prototype.

That infrastructure is what you own. You'll build the harnesses that let researchers evaluate agents against real customer incidents, the tooling that makes agent behavior legible enough to debug, and the systems that take a promising prototype from one run to thousands. You'll design experiments alongside researchers rather than only implementing them, and you'll carry the results across the line into production.

This role sits on the research team, and it is an engineering role. We are looking for a strong backend engineer with real experimentation instincts - someone who has always run experiments on the side and wants that to be the job. No PhD required. If you want to work at the frontier of LLM agents while still building serious systems, this is the seat.

Responsibilities

  • Research Infrastructure: Build and scale the infrastructure behind the team's research - experiment harnesses, evaluation pipelines, and trajectory tooling - so that ideas can be tested quickly and results can be trusted.

  • Scaling Research Prototypes: Take promising research prototypes and make them fast, reliable, and production-ready, closing the gap between a one-off experiment and a capability that ships to enterprise customers.

  • Experiment Design: Partner with researchers to design experiments - framing the hypothesis, choosing the measurement, and making sure the result is real before it drives a decision.

  • Cross-Team Collaboration: Work closely with AI engineers, infrastructure teams, and product leads to bring research into production and close the loop between experimentation and impact.

  • Stay on the Frontier: Track developments in LLMs, agent architectures, and evaluation methodology, translating insights into actionable improvements for Traversal's domain.

  • LLM & Agent Research: Prototype and evaluate prompting strategies, reasoning workflows, and tool-use policies for agents operating on large-scale observability data and complex troubleshooting workflows. Ship improvements to production.

Requirements

  • 5+ years of software engineering experience, with a strong focus on backend development and experience building scaled systems.

  • Hands-on experimentation experience - running experiments, evaluating results, and iterating on them. Side projects and personal work count.

  • Proven ability to lead in fast-paced startup environments, with limited resources, shifting priorities, and minimal structure.

  • Strong experience with Python and web frameworks such as FastAPI.

  • Experience deploying applications on AWS, working with ECS or Kubernetes, Postgres for data storage, and S3 for large-scale object storage.

  • Familiarity with observability and monitoring tools to ensure the system is running smoothly and efficiently.

Nice to Have

  • A research engineer background, or prior experience as the engineer embedded in a research team.

  • Experience building or operating agentic AI systems.

  • Background in large-scale, complex, data-driven applications.

  • Familiarity with AI or LLM-powered products, and with evaluation or benchmarking tooling.

Compensation

We offer competitive compensation, startup equity, health insurance, and additional benefits. The U.S. base salary range for this full-time, in-person role in New York is $180,000-$300,000, plus equity and benefits. Our salary ranges are based on location, level, and role. Individual compensation is determined by experience, skills, and job-related knowledge.

Why You Should Join Us

Traversal is a place to take on hard, meaningful problems with real ownership from day one. You’ll work alongside people who challenge you to grow, learn constantly, and help define a new category of infrastructure software. We think long term, move quickly, and hold a high bar without taking ourselves too seriously.

We offer competitive salary and equity packages, health insurance, fertility benefits, a great tech setup stipend and flexible time off. Plus in-office snacks, team happy hours and outings, an annual company offsite, and plenty of built in time to collaborate across teams.

Traversal is fully in-office, 5 days a week, based in New York near Madison Square Park. We have a collaborative, hard-working culture and are energized by building the future of AI-powered software maintenance.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
405,710 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
$120k – $170k per year • Remote • 6+ years exp • Bachelor's Degree
Python
SQL
Databases
DynamoDB
MS SQL
Oracle
PostgreSQL
AI/ML
NumPy
Pandas
PyTorch
Scikit-learn
SciPy
TensorFlow
DevOps
Amazon S3
AWS
CI/CD
Apply
$100k – $150k per year • Remote • 6+ years exp • Bachelor's Degree
Bash
Python
DevOps
Ansible
ArgoCD
AWS
Azure
CI/CD
GCP
GitOps
Helm
Istio
Jenkins
Kubernetes
Linkerd
OpenShift
Red Hat
Service Mesh
Tekton
Terraform
Cybersecurity
HIPAA
PCI DSS
SOC 2
Apply
$145k – $205k per year • Remote • 6+ years exp • PhD
C++
Rust
SQL
Rust
Actix Web
Axum
Databases
Apache Kafka
DynamoDB
Kafka
MySQL
NATS
PostgreSQL
RabbitMQ
Redis
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
Grafana
gRPC
Istio
Kubernetes
Linkerd
OpenTelemetry
Prometheus
Pulumi
Rest API
Service Mesh
Terraform
Apply
$140k – $200k per year • Remote • 6+ years exp • PhD
Java
Node JS
Solidity
SQL
TypeScript
JavaScript
Solidity
Truffle
Frontend
Angular
GraphQL
React.js
Remix
Vue.js
DevOps
AWS
Azure
CI/CD
Docker
Git
Kubernetes
Rest API
Terraform
Web3
Arbitrum
Avalanche
Ethereum
Foundry
Hardhat
Hyperledger Fabric
Layer 2
Optimism
Polygon
Remix
Smart Contracts
Token Standards
Apply
$120k – $170k per year • Remote • 6+ years exp • Bachelor's Degree
C#
JavaScript
SQL
TypeScript
C#
ASP.NET Core
Frontend
Angular
React.js
Mobile
Dependency Injection
DevOps
AWS
Azure
Azure DevOps
CI/CD
Docker
Git
Kubernetes
Apply
$150k – $300k per year • Equity • In office • Full-Time • 3+ years exp
AI/ML
AI Agents
DevOps
Datadog
Kubernetes
Terraform
Management
ServiceNow
Apply
$150k – $300k per year • Equity • In office • Full-Time • 8+ years exp
AI/ML
AI Agents
DevOps
Datadog
Platform Engineering
Management
ServiceNow
Apply
$150k – $300k per year • Equity • In office • Full-Time • New York
Python
TypeScript
JavaScript
Python
FastAPI
Databases
PostgreSQL
Redis
AI/ML
AI Agents
LLM
Time Series Forecasting
Frontend
React.js
DevOps
Amazon ECS
AWS
Datadog
Kubernetes
Management
ServiceNow
Apply
GTM Recruiter 27 days ago
Equity • In office • Full-Time • 4+ years exp • New York
DevOps
Datadog
Management
ServiceNow
Marketing
LinkedIn
Apply
$125k – $150k per year • Equity • In office • Full-Time • 3+ years exp • Bachelor's Degree
DevOps
Datadog
Design
Canva
Figma
Management
ServiceNow
Marketing
HubSpot
Marketo
Salesforce
Apply
$180k – $230k per year • Equity 0.8–2% • In office • Full-Time • 3+ years exp • Bachelor's Degree • New York
AI/ML
AI Agents
Apply
$115k – $125k per year • Remote • Full-Time • New York
AI/ML
Claude
Marketing
LinkedIn
Apply
$220k – $350k per year • Remote • Full-Time • 1+ year exp • New York
AI/ML
Claude
Marketing
LinkedIn
Apply
$109k – $144k per year • In office • 3+ years exp • Master's Degree • New York
Apply
$148k – $196k per year • In office • 6+ years exp • Bachelor's Degree • New York
Apply
See all jobs
This is one of many
405,710 more open roles from verified company boards, updated every day.