410,350open jobs
14,253companies
73,765added this week
Browse all
Salary
$160k – $300k per year
Location
In office (New York)
Employment
Full-Time
Overview
Company
Impact
Profile match
Traversal is an AI site reliability engineer for large distributed systems, triaging alerts and finding root cause across petabyte-scale telemetry. It reconstructs the causal chain between a symptom and the change that caused it. The founders came from causal inference research in academia.

About Traversal

Traversal is the AI Site Reliability Engineer (AI SRE) for the enterprise.

Production complexity was already outpacing what engineering teams could manage manually, and AI-generated code is accelerating that gap. Traversal is built for that challenge, autonomously understanding and reasoning across even the largest, most complex production environments to diagnose, fix, and prevent incidents. Our mission is to free engineers from endless firefighting and give them more time to focus on creative, high-impact work.

Today, Traversal operates in mission-critical environments at some of the world’s largest enterprises. Our roots remain deeply embedded in AI research, and we’ve brought together researchers from institutions including MIT, Harvard, Berkeley, Columbia, and Cornell with world-class technical staff and operators from companies like Google, Meta, Datadog, ServiceNow, and Citadel Securities to take on one of the hardest problems for AI to solve. Traversal is backed by Sequoia Capital, Kleiner Perkins, Hanabi, NFDG, and American Express Ventures.

The Role

As an AI Researcher at Traversal, you'll work on improving the accuracy and speed of our agents - systems that autonomously diagnose and resolve production incidents for some of the world’s largest enterprises.

This is a hands-on, production-oriented research role. You will design experiments, run them on real data, and ship improvements. Not flag them - ship them.

You'll work end-to-end: identify a failure mode in agent reasoning, design an intervention, evaluate it against real customer traces, and get it into production. The tooling and infrastructure are in place. The research problem is hard. The feedback loop is fast. This is not a publish-papers role. It is a make-the-agent-work role - build-and-ship cutting edge AI.

Responsibilities

  • LLM & Agent Research: Prototype and evaluate prompting strategies, reasoning workflows, and tool-use policies for agents operating on large-scale observability data and complex troubleshooting workflows. Ship improvements to production.

  • Evaluation Design: Build and maintain eval harnesses that measure real accuracy improvements on actual customer incident types - not just benchmark scores. Own the loop from hypothesis to production measurement.

  • Cross-Team Collaboration: Work closely with AI engineers, infrastructure teams, and product leads to bring research into production and close the loop between experimentation and impact.

  • Stay on the Frontier: Track developments in LLMs, agent architectures, and AI alignment, translating insights into actionable improvements for Traversal’s domain.

  • Training & Alignment: Apply fine-tuning, reinforcement learning, and reward modeling techniques to align AI behavior with real-world SRE workflows.

  • Synthetic Data & Experimentation: Design pipelines to generate synthetic incidents and observability signals, enabling scalable training and testing in data-scarce environments.

Requirements

  • PhD in Computer Science, Electrical Engineering, Statistics, or a related technical field; demonstrated depth in LLMs, agents, or applied machine learning

  • Deep applied AI expertise, including strong working knowledge of LLMs, transformers, reinforcement learning, or neural networks in agentic systems

  • Strong judgment in model evaluation and experimental iteration to improve product accuracy and behavior

  • Strong software engineering depth, with the ability to work effectively in a complex production codebase and ship production-quality code

  • Some experience shipping AI or ML systems to production

  • Ability to run rigorous experiments, interpret results, and quickly translate learnings into product improvements

  • Startup or early-team experience, with comfort operating in ambiguous environments and building without mature infrastructure

Nice to Have

  • Experience in SRE, observability, or backend systems, especially when paired with strong AI/ML depth

  • Experience with RLHF, synthetic data pipelines, or LLM evaluation tooling

  • Contributions to open-source agent frameworks such as LangGraph, DSPy, or similar

  • Research experience in LLMs, agents, or reinforcement learning, including publications in venues such as NeurIPS, ICML, or ICLR; top-tier conference publications are a plus

Compensation

We offer competitive compensation, startup equity, health insurance, and additional benefits. The U.S. base salary range for this full-time, in-person role in New York is $160,000-$300,000, plus equity and benefits. Our salary ranges are based on location, level, and role. Individual compensation is determined by experience, skills, and job-related knowledge.

Why You Should Join Us

Traversal is a place to take on hard, meaningful problems with real ownership from day one. You’ll work alongside people who challenge you to grow, learn constantly, and help define a new category of infrastructure software. We think long term, move quickly, and hold a high bar without taking ourselves too seriously.

We offer competitive salary and equity packages, health insurance, fertility benefits, a great tech setup stipend and flexible time off. Plus in-office snacks, team happy hours and outings, an annual company offsite, and plenty of built in time to collaborate across teams.

Traversal is fully in-office, 5 days a week, based in New York near Madison Square Park. We have a collaborative, hard-working culture and are energized by building the future of AI-powered software maintenance.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
410,350 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
$28k – $67k per year (Estimated) • In office • 3+ years exp • Ufa
JavaScript
Node JS
Python
SQL
TypeScript
Python
Django
FastAPI
Flask
Databases
Chroma
MySQL
pgvector
Pinecone
PostgreSQL
Qdrant
Weaviate
AI/ML
AI Agents
Anthropic
ChatGPT
Embeddings
Fine-tuning
Function Calling
Human-in-the-Loop
Hybrid Search
LLM
LLM Guardrails
OpenAI
Prompt Engineering
RAG
Reranking
Structured Outputs
Tool Use
DevOps
Azure
CI/CD
Docker
Git
Kubernetes
Rest API
Apply
$35k – $62k per year (Estimated) • Remote/Hybrid • Moscow
AI/ML
AI Agents
LLM
Marketing
LinkedIn
X (Twitter)
Apply
In office • 7+ years exp • Bachelor's Degree
Databases
Databricks
AI/ML
AI Agents
LLM
LLM Guardrails
Model Context Protocol
DevOps
AWS
Azure
Platform Engineering
Cybersecurity
Crowdstrike
Wiz
Apply
$27k – $71k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Indore
C#
Python
TypeScript
JavaScript
C#
ASP.NET Core
Databases
Apache Kafka
Kafka
AI/ML
AI Agents
Amazon SageMaker
OpenAI
Prompt Engineering
PyTorch
TensorFlow
Frontend
Angular
GraphQL
DevOps
AWS
Azure
CI/CD
Docker
GCP
IAM
Kubernetes
Rest API
Terraform
Apply
Corporate Lawyer 1 hour ago
In office • Internship • 3+ years exp • Almaty
AI/ML
LLM
Management
Outlook
Apply
$150k – $300k per year • Equity • In office • Full-Time • 3+ years exp
AI/ML
AI Agents
DevOps
Datadog
Kubernetes
Terraform
Management
ServiceNow
Apply
$150k – $300k per year • Equity • In office • Full-Time • 8+ years exp
AI/ML
AI Agents
DevOps
Datadog
Platform Engineering
Management
ServiceNow
Apply
AI Research Engineer 20 days ago
$180k – $300k per year • Equity • In office • Full-Time • 5+ years exp • PhD • New York
Python
Python
FastAPI
Databases
PostgreSQL
AI/ML
AI Agents
LLM
Tool Use
DevOps
Amazon ECS
Amazon S3
AWS
Datadog
Kubernetes
Management
ServiceNow
Apply
$150k – $300k per year • Equity • In office • Full-Time • New York
Python
TypeScript
JavaScript
Python
FastAPI
Databases
PostgreSQL
Redis
AI/ML
AI Agents
LLM
Time Series Forecasting
Frontend
React.js
DevOps
Amazon ECS
AWS
Datadog
Kubernetes
Management
ServiceNow
Apply
GTM Recruiter 27 days ago
Equity • In office • Full-Time • 4+ years exp • New York
DevOps
Datadog
Management
ServiceNow
Marketing
LinkedIn
Apply
$180k – $230k per year • Equity 0.8–2% • In office • Full-Time • 3+ years exp • Bachelor's Degree • New York
AI/ML
AI Agents
Apply
$115k – $125k per year • Remote • Full-Time • New York
AI/ML
Claude
Marketing
LinkedIn
Apply
$220k – $350k per year • Remote • Full-Time • 1+ year exp • New York
AI/ML
Claude
Marketing
LinkedIn
Apply
$109k – $144k per year • In office • 3+ years exp • Master's Degree • New York
Apply
$148k – $196k per year • In office • 6+ years exp • Bachelor's Degree • New York
Apply
See all jobs
This is one of many
410,350 more open roles from verified company boards, updated every day.