368,910open jobs
9,449companies
47,822added this week
Browse all
Salary
$250k – $285k per year
Location
Remote (United States)
Employment
Full-Time
Overview
Company
Impact
Profile match
The unified interface for LLMs. Find the best models & prices for your prompts

About OpenRouter

OpenRouter is the leading AI routing and infrastructure layer that developers and enterprises use to access, manage, and optimize the best large language models across providers without lock-in, capacity constraints, or unnecessary cost. We power the most advanced AI teams in the world by giving them the flexibility to move fast, scale confidently, and stay future-proof as models evolve.

As enterprise adoption of AI accelerates, OpenRouter sits at the center of how organizations operationalize LLMs across research, product, and production workloads.

The Role

As a Research Scientist, you will conduct deep, original research that advances how the world understands, evaluates, and routes large language models. You'll work with one of the richest datasets in AI: billions of LLM generations spanning every major model, provider, and use case.

You will own and pursue a research agenda: designing experiments, developing evaluation frameworks, and producing work that shapes how models are compared, selected, and deployed. Your findings will inform OpenRouter's routing intelligence, public rankings, and the broader AI discourse.

Success in this role is measured by the quality and impact of your research, not by shipping production code or building dashboards. You'll collaborate with product and engineering teams, but your primary focus is depth and rigor.

What You'll Do

  • Own and pursue a research agenda focused on LLM evaluation, model quality, routing optimization, and AI usage patterns, contributing original insights that advance the field.

  • Design novel evaluation frameworks and benchmarks that go beyond standard leaderboards, using real-world generation data to capture how models actually perform across tasks and contexts.

  • Conduct large-scale empirical studies on LLM behavior: how models compare across providers, how performance changes over time, and how usage patterns reveal strengths and weaknesses.

  • Develop the statistical and mathematical foundations behind our routing systems, building the models and heuristics that power intelligent provider and model selection.

  • Identify opportunities to apply research findings to feed back into OpenRouter's product and platform.

  • Collaborate with external researchers, model providers, and the open-source community to advance shared understanding of LLM capabilities and limitations.

  • Work with product and engineering teams to translate research findings into improvements to OpenRouter's platform, without being constrained to a shipping cadence.

What You Bring

Experience & Technical Skills

  • MS or PhD in a quantitative field (machine learning, statistics, computer science, mathematics, computational linguistics, or similar).

  • Track record of original research, demonstrated by first-author publications, significant open-source contributions, or equivalent impact in industry research.

  • Deep expertise in statistics, experimental design, and causal inference. You can design rigorous studies and reason carefully about validity, bias, and generalizability.

  • Strong programming skills in Python. You can build data pipelines, run large-scale experiments, and prototype models efficiently.

  • Proficiency in SQL for working with large-scale analytical databases (ClickHouse, BigQuery, or similar).

  • Hands-on experience with modern ML/NLP techniques such as LLM evaluation, fine-tuning, embeddings, classification, or reinforcement learning from human feedback.

  • Familiarity with the current LLM landscape: model architectures, provider ecosystems, benchmark suites, and the strengths and limitations of leading models.

Mindset & Approach

  • Deeply curious and self-directed. You identify the most important open questions and pursue them without waiting for direction.

  • Rigorous but pragmatic. You hold yourself to high scientific standards while operating at startup speed.

  • AI-first in your own workflow. You use LLMs, coding agents, and modern AI tools heavily in your research process and have strong opinions about what works.

  • Strong communicator. You can explain complex findings clearly in papers, blog posts, internal memos, and conversations with non-technical stakeholders.

  • Collaborative. You work well with product and engineering teams and can translate research insights into actionable recommendations.

The base salary for this full-time position in the United States, spanning multiple internal levels depending on qualifications, ranges between $250,000 to $285,000 plus benefits & equity.

If you don't think you meet all of the criteria below but still are interested in the job, please apply. Nobody checks every box, and we're looking for someone who is excited to join the team.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,910 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
Founding Engineer 1 day ago
$93k – $140k per year • In office • Full-Time • 3+ years exp • Munich
Python
Python
FastAPI
AI/ML
Fine-tuning
LLM
VLM
Apply
$70k – $105k per year • In office • Full-Time • 3+ years exp
Python
SQL
TypeScript
AI/ML
LLM
RAG
Function Calling
LLM Guardrails
Cybersecurity
GDPR
Management
n8n
Apply
$93k – $140k per year • In office • Full-Time • 3+ years exp
Node JS
TypeScript
JavaScript
Node JS
Nest.JS
AI/ML
LLM
EU AI Act
Frontend
React.js
Cybersecurity
GDPR
Apply
$150k – $220k per year • In office • Full-Time • 10+ years exp
JavaScript
Node JS
TypeScript
Databases
Google BigQuery
Frontend
React.js
DevOps
CI/CD
GCP
Apply
$118k – $258k per year (Estimated) • In office • Full-Time • 5+ years exp • Munich
Python
AI/ML
AI Agents
LLM
Apply
Data Scientist 18 days ago
$180k – $240k per year • Remote • Full-Time • 4+ years exp
Python
SQL
TypeScript
Databases
ClickHouse
Google BigQuery
AI/ML
AI Agents
dbt
Embeddings
LLM
NLP
OpenRouter
Prompt Engineering
Human-in-the-Loop
LLM Guardrails
Model Context Protocol
Apply
$245k – $280k per year • In office • Full-Time • 6+ years exp
AI/ML
OpenRouter
Edge AI
Cybersecurity
HIPAA
SOC 2
Apply
Applied AI Engineer 2 months ago
$163k – $357k per year (Estimated) • Remote • Full-Time
AI/ML
LLM
OpenRouter
LLM Guardrails
AI Agents
Apply
$220k – $260k per year • Remote • Full-Time • 4+ years exp
Python
TypeScript
AI/ML
AI Agents
LLM
OpenRouter
Prompt Engineering
OpenAI
DevOps
Rest API
Apply
$215k – $285k per year • Remote • Full-Time • 4+ years exp
TypeScript
JavaScript
AI/ML
OpenRouter
Frontend
Next.js
React.js
Apply
See all jobs
This is one of many
368,910 more open roles from verified company boards, updated every day.