386,261open jobs
10,153companies
47,934added this week
Browse all
Salary
$159k – $308k per year (Estimated)
Location
In office (San Francisco)
Seniority
Senior
Overview
Company
Impact
Profile match
Research and infrastructure for representational accuracy in human-AI systems. The Behavioral Specification is the interpretive layer above retrieval that supplies the framework facts are read through. Empirical foundation: the Beyond Recall preprint (Gulaya 2026). Open source under Apache 2.0.

ABOUT BASELAYER

Every business in America needs a bank account to exist. The system that decides whether they're real, who's behind them, and whether they're a risk, runs on infrastructure from the 1980s. We're rebuilding that layer from scratch.

Baselayer is the identity layer for institutions across the United States - the most complete business graph in America and every human tied to it. We fuse public records, IRS data, sanctions lists, web signals, and fraud telemetry from 2,200+ financial institutions into a single graph that resolves any business and the humans behind it in milliseconds. The legacy credit bureaus took 50 years to build something that gets 60% match rates. We've built something that gets 98% in under two years.

Today we're trusted by over 20% of financial institutions in America - including FIS, Rho, Socure and leading loan infrastructure providers. But the graph is becoming infrastructure for anyone who needs to know if a business is real and worth trusting: gig platforms, marketplaces, AI companies, and commerce infrastructure at scale.

Trust is the substrate of every financial transaction. We're rebuilding it.

ABOUT THE TEAM

We're solving real-time entity resolution at a scale no one else has cracked - fusing dozens of data sources into a single business identity graph and resolving any entity in milliseconds. It's a graph AI problem, a retrieval problem, and a fraud-modeling problem stacked on top of each other. The technical depth is real.

You'd be joining a small team where the data moat is defensible, the research problems are open, and the infrastructure you build becomes load-bearing for businesses. Ownership is real. Velocity is real. There's no layer of process between an idea and shipping it.

We're at an inflection point - the graph is built, the match rates speak for themselves, and the hardest problems are still ahead: graph embeddings, fraud propagation models across the business network, real-time traversal at sub-100ms latency, and expanding the identity layer beyond finance into every platform that needs to trust a business.

If you want to work on something foundational - the kind of infrastructure that gets built once and everything else runs on top of - this is it.

ABOUT THE ROLE

Baselayer answers questions the loan application didn't ask. For every business that crosses our queues, we need to know things that aren't on the form: what the business actually does, where it actually lives on the web, whether the people it names match the public record, and whether anything across the open web contradicts the story we were told. We answer those questions with LLM-driven agents that crawl, click, search, and extract structured evidence from across the web - and we treat this as a production data pipeline, not a research demo. We're hiring a Senior AI Engineer to own a slice of this enrichment surface end-to-end.

WHAT YOU'LL DO

  • Own industry/category classification of businesses from heterogeneous signals (name, website, directory presence, reviews).
  • Build and maintain discovery and verification systems for a business's real web presence - filtering aggregators, parked domains, brand collisions, and impersonators.
  • Link individuals to businesses via public web evidence (e.g. confirming a named officer or employee genuinely works there).
  • Develop risk/legitimacy scoring derived from web-presence signals, fed back into downstream underwriting.
  • Build and evolve the shared agent infrastructure: provider-agnostic base agents, shared toolset registry (browser navigation, search, scraping, structured database lookups, scoring), eval harness, and instrumentation surface for token-and-tool tracing.
  • Own model selection, agent design, prompt and tool engineering, eval methodology, and cost control across your enrichment surface.

MINIMUM REQUIREMENTS

  • Shipped LLM-driven agents to production - not notebooks, not demos. Real users, real cost, real failure modes, real on-call.
  • Strong async Python including structured-data libraries, modern web frameworks, and relational databases.
  • Experience across multiple frontier LLM providers and at least one agent framework, with deep knowledge of failure modes.
  • Built or maintained eval methodology: curated golden datasets, scoring functions, labelling guidelines, regression diagnostics.
  • Browser automation experience: headless browsers, anti-bot evasion, authenticated flows.
  • Holds informed opinions on structured-output reliability - when to use JSON-schema mode vs. function calling vs. extractor-on-top-of-text.

WHAT SETS YOU APART

  • Web scraping at scale: anti-bot evasion, residential proxies, request fingerprinting, authenticated flows, CDN defeats.
  • Eval-framework experience (e.g., LangSmith, Braintrust, Evals, or custom).
  • Entity resolution / record linkage / fuzzy matching at scale.
  • Browser-automation experience at the devtools-protocol level.
  • Built a tool registry or toolset abstraction over multiple LLM providers.
  • Cost/latency optimization: response caching, semantic caching, model routing (cheap-first then escalate), thinking-budget tuning, prompt-cache hit-rate work.

WORK LOCATION

  • Based in SF; hybrid - 4 days per week in office.

COMPENSATION

  • Salary Range: $230,000 - $340,000 + Equity

BENEFITS

  • Time off when you need it: Flexible PTO so you can recharge without red tape.
  • In-person energy: We're based in SF and meet in the office 4 days a week.
  • Competitive compensation: We pay well and back it with equity. We want you to think and act like an owner.
  • Career rocket fuel: You'll help build the foundation of a high-growth startup, working side by side with experienced founders and team members who've done it before.
  • Benefits on us: We cover 100% of your health, dental, and vision premiums. No surprise deductions from your paycheck.
  • 401(k) with company match: We match your contributions so your future self benefits too
  • HSA contributions included: We contribute to your HSA on applicable plans, so your coverage works as hard as you do
  • Stay healthy, stay sharp: A $250 monthly gym stipend to help you bring your best self to work, and everywhere else
  • A seat at the table: We believe in transparency, radical candor, and giving every team member a voice
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
386,261 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
San Francisco
$149k – $260k per year • In office • Full-Time • 10+ years exp • Master's Degree • San Francisco • New York • Washington • Indianapolis
Python
Apex
Apex
Salesforce Data Cloud
Databases
Apache Kafka
Kafka
Neo4j
AI/ML
Agentforce
AI Agents
Claude
Copilot
Cursor
Embeddings
Knowledge Graph
LangChain
LLM
Model Context Protocol
RAG
Semantic Search
Semantic Search
Windsurf
DevOps
AWS
Azure
GCP
GitHub
gRPC
Vector
Marketing
Salesforce
Apply
$65k – $160k per year (Estimated) • In office • Full-Time • 5+ years exp • PhD • London
Apex
Apex
Salesforce Data Cloud
Databases
BigQuery
Databricks
Google BigQuery
Snowflake
AI/ML
Agentforce
AI Agents
Claude
Cursor
LLM
LLM Guardrails
Prompt Engineering
Marketing
Salesforce
Apply
AI Architect 1 day ago
In office • Helsinki
C#
Java
Python
Databases
Databricks
AI/ML
AI Agents
LangChain
LLM
LLM Guardrails
LLMOps
OpenAI
RAG
Semantic Kernel
DevOps
AWS
Azure
CI/CD
GCP
Marketing
LinkedIn
Apply
$25k – $66k per year (Estimated) • In office • Full-Time • 4+ years exp • India
Go
Python
Rust
Databases
Kafka
AI/ML
AI Agents
LLM
Model Context Protocol
DevOps
Amazon Kinesis
AWS
Azure
CI/CD
Docker
GCP
Kubernetes
Apply
$21k – $68k per year (Estimated) • In office • 8+ years exp • Bengaluru
JavaScript
TypeScript
AI/ML
AI Agents
Claude
Gemini
LLM
OpenAI
RAG
Frontend
React.js
Redux
Redux Toolkit
Mobile
State Management
DevOps
CI/CD
Rest API
QA
Cypress
Jest
Apply
Sr. Product Manager 3 days ago
$156k – $281k per year (Estimated) • In office • Full-Time • 5+ years exp • San Francisco
AI/ML
AI Agents
Embeddings
Apply
$156k – $281k per year (Estimated) • In office • San Francisco
Python
SQL
Python
FastAPI
Databases
PostgreSQL
AI/ML
Embeddings
DevOps
GCP
Google Cloud Run
Vector
Apply
$157k – $305k per year (Estimated) • In office • San Francisco
Python
Python
FastAPI
SQLAlchemy
SQLModel
Databases
PostgreSQL
AI/ML
AI Agents
Embeddings
DevOps
Cloudflare
Fastly
Google Cloud Run
GCP
Apply
$76k – $154k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • San Francisco
C++
Python
Ruby
Databases
PostgreSQL
AI/ML
Copilot
LLM
RAG
LLM Guardrails
Design
Figma
Management
Slack
Marketing
Salesforce
Zendesk
Apply
$133k – $184k per year • Equity • Remote/Hybrid • Full-Time • 2+ years exp • San Francisco
Ruby
TypeScript
Ruby
Ruby on Rails
Analytics
A/B Testing
Apply
$152k – $310k per year (Estimated) • Equity • Remote/Hybrid • 6+ years exp • Bachelor's Degree • San Francisco
Go
Kotlin
Python
Ruby
Ruby
Ruby on Rails
DevOps
Ansible
ArgoCD
AWS
CI/CD
CircleCI
Configuration Management
Gerrit
Git
Gitflow
GitHub
GitHub Actions
Helm
Jenkins
Kubernetes
Self-Healing
Spinnaker
Tekton
Terraform
Apply
$152k – $274k per year (Estimated) • Remote/Hybrid • Internship • 5+ years exp • San Francisco
Design
Figma
Management
Asana
Notion
Marketing
Salesforce
Apply
$197k – $314k per year • In office • Full-Time • 12+ years exp • PhD • New York • San Francisco
AI/ML
Agentforce
AI Agents
Marketing
Salesforce
Apply
See all jobs
This is one of many
386,261 more open roles from verified company boards, updated every day.