368,941open jobs
9,452companies
47,951added this week
Browse all
Salary
$164k – $307k per year (Estimated)
Location
In office (New York, San Francisco)
Seniority
Middle · 4+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
LangChain is an AI technology company that provides open-source frameworks and developer tools for building applications powered by large language models and AI agents. Its ecosystem includes LangChain for agent development, LangGraph for workflow orchestration, and LangSmith for testing, monitoring, evaluation, and deployment. The platform helps developers and enterprises create reliable generative AI applications by connecting models, tools, data sources, and business workflows.

About Us

At LangChain, our mission is to make intelligent agents ubiquitous. We build the foundation for agent engineering in the real world, helping developers move from prototypes to production-ready AI agents that teams can rely on. We began as widely adopted open-source tools and have grown to also offer a platform for building, evaluating, deploying, and operating agents at scale.

With $125M raised at Series B from IVP, Sequoia, Benchmark, CapitalG, and Sapphire Ventures, we’re at a stage where we’re continuing to develop new products, growth is accelerating, and all team members have meaningful impact on what we build and how we work together. LangChain is a place where your contributions can shape how this technology shows up in the real world.

Today, our platform includes LangSmith (Observability, Evaluation, Deployment, Fleet, and Sandboxes), our open source frameworks (LangChain, LangGraph, and Deep Agents), and the newly launched LangSmith Engine for autonomous agent improvement. We have 100M+ monthly open source downloads, 6,000+ active LangSmith customers, and 5 of the Fortune 10 use LangSmith in production (+ 35% of the Fortune 500 overall), including teams at Klarna, Clay, Coinbase, Workday, Lyft, Cloudflare, Harvey, Rippling, Vanta, LinkedIn, Monday.com, Nvidia, and Bridgewater.

About the Team:

The LangSmith Engine team is building a proactive agent engineer that analyzes production traces, identifies important failures, recommends and writes fixes, and helps prevent those issues from coming back. We’re building agents that can understand complex software systems and continuously improve the quality of other AI agents.

About the Role:

We’re looking for an experienced research engineer to help make the Engine agent more capable and more efficient.

You’ll study real agent failures, build benchmarks that capture what matters, run experiments to improve performance, and turn successful ideas into production. This may include prompting and agent-harness improvements, model selection, fine-tuning and post-training custom models. The focus is on measurable improvements to the overall agent.

This role also requires a understanding of production engineering and system-level tradeoffs. Engine is a production system, so improving an agent is not just about maximizing benchmark performance-it also means understanding the impact on cost, latency, reliability, and scalability. You’ll work in the same team with production engineers to design, test, and ship improvements that work reliably in real-world environments.

Location: SF and NYC

What You’ll Do:

  • Build and maintain benchmarks and evaluations that measure the quality and efficiency of Engine agents on real-world tasks.

  • Design and run experiments to improve agent performance across models, prompting, context, tools, orchestration, and agent strategies.

  • Explore and implement post-training and fine-tuning techniques when they can meaningfully improve agent capabilities, quality, or cost.

  • Turn successful experiments into production improvements, working closely with engineers and researchers to measure impact and prevent regressions.

  • Help define the ML roadmap and technical direction for improving Engine agents, and mentor other engineers through strong technical leadership.

What You’ll Bring:

  • 4+ years of experience in ML/AI research, or a closely related field.

  • Master’s or Ph.D. in a relevant scientific field.

  • Hands-on experience working with LLMs and AI agents, including analyzing model behavior and improving real-world performance

  • Strong experience designing benchmarks, evaluations, and experiments for AI/ML systems; you know how to tell whether a change actually made an agent better.

  • Strong software engineering skills, with a track record of taking ideas from research prototype to measurable production impact.

  • You have maximum agency and strong research judgment: you can identify high-impact problems, work through ambiguity, move quickly, and communicate your findings clearly.

Nice to Have:

  • Ph.D. in Machine Learning, Computer Science or Physics.

  • Hands on experience with LLM-as-a-judge, automated graders, synthetic data generation, or human evaluation.

  • Hands on experience with reinforcement learning, preference optimization, SFT, RLHF/RLAIF, or other post-training techniques for LLMs.

  • Experience optimizing LLM agents for cost, latency, or task efficiency on productions

  • Experience with model serving, inference optimization, distributed systems, or GPU infrastructure.

Compensation & Benefits

We offer competitive compensation that includes base salary, variable compensation for relevant roles, meaningful equity, benefits, and perks. Benefits include things like medical, dental, and vision coverage, flexible vacation, a 401(k) plan, and life insurance. Actual compensation and offerings will vary based on role, level, and location. Team members in the EU, UK, and APAC receive locally competitive benefits aligned with regional norms and regulations.

Compensation Philosophy:

We offer competitive compensation that includes base salary, variable compensation for relevant roles, meaningful equity, benefits, and perks. Actual compensation and offerings will vary based on role, level, and location. Team members in the EU, UK, and APAC receive locally competitive benefits aligned with regional norms and regulations.

Benefits

Benefits include medical, dental, and vision coverage, flexible vacation, a 401(k) plan, meals on in-office days in the US and more.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,941 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
New York
$47k – $106k per year (Estimated) • In office • Bachelor's Degree • Moscow
Python
SQL
AI/ML
Hadoop
LangChain
LangGraph
LLM
NLP
NumPy
Pandas
Prompt Engineering
PyTorch
Spark
AI Agents
DevOps
Git
SLI/SLO/SLA
Apply
$120k – $250k per year • Equity 0.2–1% • In office • Full-Time • Master's Degree • Seattle
Python
AI/ML
Diffusion Models
Fine-tuning
PyTorch
Self-Supervised Learning
Apply
$120k – $180k per year • Equity 0.2–1.5% • In office • Full-Time • 1+ year exp • San Francisco
Python
AI/ML
AI Agents
ChatGPT
DevOps
Error Budget
Apply
$230k – $300k per year • Equity 0.1–0.2% • In office • Full-Time • 3+ years exp • PhD • New York
TypeScript
Databases
PostgreSQL
AI/ML
AI Agents
Claude
Claude Code
Cursor
Devin
OpenAI Codex
Frontend
Next.js
tRPC
DevOps
Platform Engineering
Vercel
Apply
$17k – $47k per year (Estimated) • Remote/Hybrid • Internship • 5+ years exp • Minsk
Go
JavaScript
Python
AI/ML
AI Agents
Claude
Claude Code
Copilot
Cursor
LLM
Prompt Engineering
RAG
Anthropic
Function Calling
OpenAI
Frontend
React.js
DevOps
CI/CD
Docker
Git
Terraform
GitHub
Apply
In office • Full-Time • 7+ years exp • Singapore
Python
TypeScript
AI/ML
AI Agents
LangChain
LangGraph
LangSmith
LLM
Prompt Engineering
RAG
Edge AI
DevOps
Amazon EKS
AWS
Azure
Azure AKS
CI/CD
Cloudflare
Datadog
GCP
GitOps
Google GKE
Grafana
Helm
Kubernetes
Platform Engineering
Prometheus
Terraform
Vector
Analytics
A/B Testing
Management
Monday.com
Apply
In office • Full-Time • 7+ years exp • Amsterdam
Python
TypeScript
AI/ML
AI Agents
LangChain
LangGraph
LangSmith
LLM
Prompt Engineering
RAG
Edge AI
DevOps
Amazon EKS
AWS
Azure
Azure AKS
CI/CD
Cloudflare
Datadog
GCP
GitOps
Google GKE
Grafana
Helm
Kubernetes
Platform Engineering
Prometheus
Terraform
Vector
Analytics
A/B Testing
Management
Monday.com
Apply
In office • Full-Time • 7+ years exp • London
Python
TypeScript
AI/ML
AI Agents
LangChain
LangGraph
LangSmith
LLM
Prompt Engineering
RAG
Edge AI
DevOps
Amazon EKS
AWS
Azure
Azure AKS
CI/CD
Cloudflare
Datadog
GCP
GitOps
Google GKE
Grafana
Helm
Kubernetes
Platform Engineering
Prometheus
Terraform
Vector
Analytics
A/B Testing
Management
Monday.com
Apply
$122k – $244k per year (Estimated) • In office • Full-Time • 4+ years exp • Atlanta • Austin • Dallas • Las Vegas • Nashville
JavaScript
Python
TypeScript
AI/ML
AI Agents
Axolotl
Fine-tuning
LangChain
LangGraph
LangSmith
RLHF
TRL
Unsloth
Transformers
DPO
Hugging Face
Post-training
SFT
DevOps
Cloudflare
Management
Monday.com
Apply
$140k – $279k per year (Estimated) • In office • Full-Time • 4+ years exp • New York
JavaScript
Python
TypeScript
AI/ML
AI Agents
Axolotl
Fine-tuning
LangChain
LangGraph
LangSmith
RLHF
TRL
Unsloth
Transformers
DPO
Hugging Face
Post-training
SFT
DevOps
Cloudflare
Management
Monday.com
Apply
$197k – $374k per year (Estimated) • In office • Full-Time • 8+ years exp • PhD • New York
AI/ML
AI Agents
Claude
LangChain
OpenAI
Vertex AI
Management
n8n
Zapier
Apply
$350k – $475k per year • In office • Full-Time • 4+ years exp • San Francisco • New York
C++
Python
C++
PyTorch C++
AI/ML
PyTorch
Ray
Reinforcement Learning
RLHF
DPO
InfiniBand
NCCL
Post-training
PPO
TPU
DevOps
Kubernetes
SLURM
SRE
Apply
$350k – $475k per year • In office • Full-Time • San Francisco • New York
AI/ML
Fine-tuning
LoRA
PEFT
DevOps
CI/CD
Kubernetes
SRE
Apply
Senior Data Analyst 4 hours ago
$86k – $171k per year (Estimated) • Equity • Remote • Full-Time • 5+ years exp • Bachelor's Degree • New York
Python
SQL
Databases
Snowflake
AI/ML
Anomaly Detection
Claude
Copilot
Cursor
dbt
Edge AI
DevOps
AWS
Analytics
Tableau
Marketing
Salesforce
Apply
$170k – $230k per year • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • New York • San Francisco
Go
Java
Kotlin
Python
Scala
Databases
Apache Kafka
Databricks
Delta Lake
Snowflake
AI/ML
Dagster
dbt
Flink
Spark
Knowledge Graph
Apply
See all jobs
This is one of many
368,941 more open roles from verified company boards, updated every day.