368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$119k – $262k per year (Estimated)
Location
Remote/Hybrid (London, United Kingdom, Toronto, Canada)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Cohere is a Canadian AI company founded in 2019 and headquartered in Toronto. It builds secure, enterprise-focused large language models and AI tools for businesses and regulated industries. The company is known for emphasizing privacy, private deployment, and "sovereign AI" rather than consumer chat products

Who are we?

Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems.

We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that.

We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft.

We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us!

Why This Role?

North is Cohere’s AI workspace platform for enterprises: a secure, customizable environment where companies can use AI across their real workflows while maintaining control over sensitive data. North connects AI agents with workplace tools, applications, and business context, helping users delegate complex work, build automations, inspect outputs, and collaborate with AI in production environments.

As North becomes more capable, one of the most important questions is also one of the hardest: how do we know whether the model is actually getting better for the workflows customers care about?

This role is about being the voice of North inside modelling. You will build the evaluation systems, feedback loops, and applied modelling workflows that make sure model progress translates into better product outcomes for North users. You will work closely with North product teams, customer-facing teams, and modelling teams to define what “good” means across the product surface, turn real usage and product direction into high-quality evals, and use those evals to guide model selection, patches, and regular model updates.

This is neither a pure research role nor a conventional product engineering role. It is a rigorous applied MLE role for someone who cares deeply about measurement, model behavior, and real-world product quality. You should be excited by the craft of building careful evals: evals that capture messy agentic workflows, reflect actual customer needs, resist superficial benchmark hacking, and provide useful signal for where the product and models need to go next.

As a Member of Technical Staff, North Modelling (Evals), You Will:

  • Own the eval strategy for North: define what we need to measure across agent workflows, tool use, enterprise knowledge work such as deep research, document creation or editing, and other human-AI interactions.

  • Build high-quality evals from the realities of the product: user feedback, production failures, privacy-preserving usage logs, internal dogfooding, customer needs, and forward-looking product goals.

  • Create systems that continuously turn what North is learning from users, customers, and feature teams into evals, so measurement keeps pace with the product rather than becoming a static benchmark.

  • Be the voice of North inside modelling: extract clear insight from eval results, customer-derived data and product context, then turn model failures, capability gaps, and North-specific needs into actionable recommendations for the central modelling teams.

You May Be a Good Fit If:

  • You have improved LLM-powered, agent-powered, or AI-product systems through evals, feedback loops, data curation, prompting, model adaptation, or model selection.

  • You care deeply about evaluation as a craft: representative tasks, precise rubrics, clean data, failure analysis, regression tracking, and knowing when a metric is giving false confidence.

  • You have strong applied MLE judgment and can reason clearly about model behavior, eval validity, product outcomes, and production tradeoffs.

  • You are comfortable working close to users and product teams, while translating messy qualitative signals into measurement that other modelling teams can act on.

  • You are self-directed, practical, and motivated by open-ended problems where the right answer requires both technical depth and product understanding.

  • You care about making AI systems genuinely useful in real enterprise settings, not just better on abstract benchmarks.

If some of the above does not line up perfectly with your experience, we still encourage you to apply.

Location

This team works closely across Europe and East Coast North America time zones. We are open to candidates who can collaborate effectively within those hours.

Full-Time Employees at Cohere enjoy these Perks:

  • A weekly lunch stipend of $75/£75 or equivalent in your local currency for lunch.

  • Full health and dental benefits, including a separate budget for mental health.

  • RRSP matching, 401K, Pension Scheme.

  • 100% Parental Leave top-up for up to 6 months, for either parent.

  • Annual enrichment benefits:

    Arts & culture, fitness/wellness, quality time, and a workspace improvement credit.

    Education & learning stipend for conferences, courses, and coaching.

  • 6 weeks of paid vacation (30 working days!)

  • Budget for traveling to other offices if you are remote, plus an annual company offsite.

How and Where We Work:

  • Cohere is remote-friendly, but we also have offices in Toronto, London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul with more opening soon.

  • For those in the office: a daily lunch program, plenty of snacks, and regular community and social events.

  • For those not near an office: a co-working benefit so you can work alongside others in your city.

  • Everyone receives a $500 home office stipend to set up your workspace properly.

If any of the above doesn’t line up exactly with your experience, we still encourage you to apply.

We strive to create an inclusive work environment for all; we welcome applicants from all backgrounds and are committed to providing equal opportunities. Should you require any accommodations during the recruitment process, please submit an Accommodations Request Form, and we will work together to meet your needs.

We may use AI-enabled tools to screen and assess applicants against the criteria for this position. This helps our recruiters identify potentially qualified candidates, but it doesn't limit the applications our recruiters may review or consider.

Beware of Scams: Cohere will never ask for payment or third-party services (e.g., CV writing) as part of our hiring process. All legitimate roles are listed on the Cohere careers page and LinkedIn only, with all communications from Cohere employees coming from an @cohere.com or @cw.cohere email alias. If jobs are viewed on other sites then please verify these through our official careers page.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
London
$305k per year • In office • 8+ years exp • Bachelor's Degree • San Francisco
AI/ML
AI Agents
Anthropic
Claude
LLM
Multimodal AI
Apply
$18k – $22k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Bucharest
AI/ML
AI Agents
LLM
DevOps
GCP
Cybersecurity
OWASP Top 10
Apply
$75k per year • In office • Full-Time • Estonia
Python
AI/ML
AI Agents
DevOps
AWS
Apply
$148k – $267k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Sunnyvale • Boston • Austin • Toronto • New York
AI/ML
AI Agents
Claude
Claude Code
Lovable
Replit
Cybersecurity
BeyondTrust
Crowdstrike
CyberArk
Delinea
Microsoft Entra ID
Okta
PCI DSS
SOC 2
Threat Modeling
Apply
$142k – $215k per year • In office • Full-Time • 12+ years exp • Princeton
JavaScript
TypeScript
Java
Java
Gradle
Hibernate
Maven
Spring Boot
Databases
Databricks
Snowflake
AI/ML
AI Agents
Claude
Claude Code
Frontend
Angular
React.js
Vue.js
DevOps
Amazon CloudWatch
Amazon ECS
Amazon EKS
Amazon EventBridge
Amazon S3
API Gateway
AWS
AWS Lambda
AWS Step Functions
Azure
CI/CD
IAM
Platform Engineering
Rest API
Kubernetes
Apply
$124k – $271k per year (Estimated) • Remote/Hybrid • Full-Time • London
AI/ML
AI Agents
Reinforcement Learning
Synthetic Data
Apply
$97k – $216k per year (Estimated) • Remote • Full-Time • Toronto • Calgary • Ottawa • Montreal
Kotlin
Swift
AI/ML
Edge AI
Apply
$66k – $159k per year (Estimated) • Remote • Full-Time • 3+ years exp • Stockholm • London
Python
AI/ML
Cohere SDK
LLM
NLP
Apply
$83k – $186k per year (Estimated) • Remote • Full-Time • London
Python
AI/ML
AI Agents
Cohere SDK
LLM
RAG
Edge AI
Apply
$80k – $181k per year (Estimated) • Remote • Full-Time • London
AI/ML
AI Agents
Edge AI
DevOps
AWS
Azure
CI/CD
GCP
Git
Helm
Kubernetes
Apply
$77k – $148k per year (Estimated) • Equity • In office • Master's Degree • London
JavaScript
Python
Scala
AI/ML
AI Agents
DevOps
GitHub
Apply
$87k – $159k per year (Estimated) • In office • Contractor • 5+ years exp • London • Stockholm
Design
Figma
Marketing
Zendesk
Apply
$105k – $204k per year (Estimated) • In office • Full-Time • London
Python
SQL
Databases
Snowflake
AI/ML
Dagster
dbt
Analytics
A/B Testing
Apply
$27k – $61k per year (Estimated) • In office • Full-Time • 5+ years exp • Pune • London
Analytics
Power BI
Tableau
Marketing
Salesforce
Apply
$34k – $85k per year (Estimated) • In office • Full-Time • Bachelor's Degree • London
AI/ML
AI Agents
Edge AI
QA
Appium
Cucumber
Cypress
Selenium
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.