368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$58k – $164k per year (Estimated)
Location
Remote (Vancouver, Toronto, Calgary, Edmonton, Winnipeg, Canada)
Seniority
Junior · 1+ year exp
Employment
Contractor
Overview
Company
Impact
Profile match
Cohere is a Canadian AI company founded in 2019 and headquartered in Toronto. It builds secure, enterprise-focused large language models and AI tools for businesses and regulated industries. The company is known for emphasizing privacy, private deployment, and "sovereign AI" rather than consumer chat products

Who are we?

Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems.

We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that.

We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft.

We are a global technology company co-headquartered in Toronto and San Francisco, with key offices in London, New York City, Montreal, Seoul, Germany and Paris. Join us!

Why this role?

We are on a mission to build machines that understand the world and make them safely accessible to all. Data quality is foundational to this process. Machines (or Large Language Models, to be exact) learn in similar ways to humans, by way of feedback. By labelling, ranking, auditing, and correcting model output, you will improve Large Language Models' performance for iterations to come, thus having a lasting impact on Cohere's technology. We are hiring Generalist professionals with broad backgrounds that span multiple consumer-facing or personal domains.

This is a judgment-driven role, not passive data entry. You will review, assess, and provide structured feedback across a broad and evolving range of tasks, evaluating, stress-testing, and improving our models on English-language data spanning multiple modalities (text, image, and structured formats such as JSON, CSV/TSV, and Markdown). This is a great opportunity for professionals with strong analytical skills to contribute to high-impact annotation projects.

Please Note: This is a part-time independent contractor position available within Canada only. We seek candidates who can commit to 16 hours per week at a CAD $30/hour contract rate. This role is BYOD (Bring Your Own Device, laptop). This position is remote.

As an Data Annotation Specialist, you will:

  • Evaluate and rank model outputs: Complete preference and comparison tasks, assessing which responses best conform to project guidelines for accuracy, helpfulness, tone, and safety, and writing clear justifications for your judgments.

  • Stress-test and break models: Probe models adversarially to surface failure modes, unsafe behavior, and capability gaps, and document reproducible cases that engineering and research teams can act on.

  • Create datasets: Author high-quality prompts, responses, and exemplars to build training and evaluation datasets, following detailed specifications and editing machine-written or human-written outputs to standard.

  • Build and apply rubrics and taxonomies: Contribute to the design of grading criteria and rubrics, then apply them consistently to produce structured, high-quality annotations across task types.

  • Annotate and correct multimodal data: Label, audit, and rectify inaccuracies across text, image, and structured data, maintaining a high standard of data integrity and accuracy.

  • Calibrate and maintain consistency: Participate in calibration exercises and inter-annotator agreement checks to align on standards, and flag ambiguous or uncovered edge cases rather than guessing, since a single misjudgment replicated at scale degrades a model.

  • Adapt to experimental work: Take on new and evolving task types as project needs shift, applying sound judgment in areas where guidelines are still being developed.

  • Report on model performance: Surface and communicate quality and performance trends in model and agent behavior, giving cross-functional partners clear, well-evidenced feedback on where models succeed, fail, and degrade.

You may be a good fit if you have:

  • 1+ years of experience in AI data annotation, LLM evaluation, content moderation, research, or a related analytical role, with exposure to quality assurance, and/or preference ranking.

  • Experience applying detailed guidelines to complex and often ambiguous content, with strong contextual and sociocultural judgment, sensitivity to nuance, tone, and register, and the ability to reason well in cases where there is no single correct answer.

  • Comfort with ambiguity: a willingness to flag unclear or uncovered edge cases rather than guess, and to work productively on novel, experimental tasks whose definitions are still evolving.

  • A sharp, curious eye for inconsistencies, subtle errors, and model failure modes, including the instinct to probe models adversarially and surface where they break. Familiarity with how large language models behave, such as hallucination, sycophancy, and instruction-following gaps, is a plus.

  • Excellent command of written English and strong reading comprehension, with the ability to clearly justify your evaluations, including why an output is correct or incorrect, high-quality or low-quality, and to write clear prompts and exemplars that explain your reasoning. Bonus points if you are fluent in another language!

  • Strong attention to detail and commitment to accuracy, with the ability to maintain consistency across high-volume and monotonous tasks.

  • Comfort working with annotation platforms and structured formats such as JSON, CSV/TSV, Markdown, XML, and YAML.

  • Strong execution in a remote environment, including good time management, comfort using new tools, and the ability to work independently in a global, asynchronous team.

Candidate Journey

  • Initial Screening: Once you have submitted your application, our Talent Team will review your resume and writing samples.

  • Virtual Annotation Test: This assignment will test your written and technical skills through various language-based tasks, such as a data science take-home assessment, writing sample, and more.

  • Video Screen: If selected to move forward, you will have a short video call with a member of our Operations Team!

  • Offer: Independent Contractor Agreement.

As an independent contractor, you maintain control over how you complete your work and may work with multiple clients simultaneously, although we ask you to declare if any of these are with a direct competitor of Cohere and maintain IP confidentiality of the Cohere project. Independent contractors are not eligible for health benefits or other benefits provided to employees. Compensation for services is provided to contractors by contractors invoicing for services provided pursuant to the terms of our agreement with the contractor.

It is important to understand thatas an independent contractor, continuous work is not guaranteed. The client-contractor relationship is fundamentally project-based,meaning engagements may be temporary, periodic, or intermittent based on our organizational needs and project availability. As an independent contractor, you should anticipate fluctuations in workflow and, therefore, compensation for services when Cohere does not require as many hours of services in a week.

Prospective candidates, please be advised: this role involves working with human-generated and model-generated tasks that may involve exposure to not safe for work (NSFW) text content as part of data annotation tasks, including explicit, offensive, or other inappropriate material.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Vancouver
Remote • Full-Time
JavaScript
Python
TypeScript
AI/ML
AI Agents
Function Calling
LLM
Prompt Engineering
Structured Outputs
Management
n8n
Zapier
Apply
$27k – $110k per year (Estimated) • Remote • Full-Time
JavaScript
Python
TypeScript
AI/ML
AI Agents
Function Calling
LLM
Prompt Engineering
Structured Outputs
Management
n8n
Zapier
Apply
$27k – $69k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
Python
TypeScript
AI/ML
Embeddings
LLM
RAG
LLM Guardrails
AI Agents
Function Calling
Model Context Protocol
Frontend
Angular
DevOps
AWS
Azure
CI/CD
Docker
GCP
Vector
Apply
$28k – $53k per year (Estimated) • In office • Full-Time • Rostov-on-Don
AI/ML
LLM
RAG
Apply
$39k – $87k per year (Estimated) • In office • 9+ years exp • Bengaluru
Java
Python
Java
Spring Boot
Databases
ElasticSearch
Google BigQuery
Neo4j
PostgreSQL
AI/ML
Claude
Claude Code
Copilot
Cursor
LLM
AI Agents
Devin
Model Context Protocol
OpenAI Codex
DevOps
AWS
Datadog
GCP
Grafana
New Relic
Prometheus
Management
Jira
Apply
$124k – $271k per year (Estimated) • Remote/Hybrid • Full-Time • London
AI/ML
AI Agents
Reinforcement Learning
Synthetic Data
Apply
$119k – $262k per year (Estimated) • Remote/Hybrid • Full-Time • London • Toronto
AI/ML
AI Agents
Cohere SDK
LLM
Function Calling
Apply
$97k – $216k per year (Estimated) • Remote • Full-Time • Toronto • Calgary • Ottawa • Montreal
Kotlin
Swift
AI/ML
Edge AI
Apply
$66k – $159k per year (Estimated) • Remote • Full-Time • 3+ years exp • Stockholm • London
Python
AI/ML
Cohere SDK
LLM
NLP
Apply
$83k – $186k per year (Estimated) • Remote • Full-Time • London
Python
AI/ML
AI Agents
Cohere SDK
LLM
RAG
Edge AI
Apply
$79k – $192k per year (Estimated) • Remote/Hybrid • Contractor • 6+ years exp • Bachelor's Degree • Vancouver
C#
JavaScript
TypeScript
Go
C#
.NET
Go
Temporal
Databases
ClickHouse
PostgreSQL
AI/ML
Edge AI
Frontend
Next.js
React.js
shadcn/ui
Storybook
Tailwind CSS
Radix UI
Mobile
PWA
DevOps
CI/CD
GitLab
GitLab CI
Rest API
Cybersecurity
Keycloak
Management
Confluence
Jira
QA
Cypress
Jest
Playwright
Swagger
Vitest
Apply
Associate QA 2 hours ago
$53k – $122k per year (Estimated) • Remote/Hybrid • 3+ years exp • Vancouver
C#
SQL
TypeScript
JavaScript
Go
C#
.NET
Go
Temporal
Databases
ClickHouse
PostgreSQL
AI/ML
Edge AI
Frontend
Lighthouse
Next.js
React.js
DevOps
CI/CD
Docker
GitLab
GitLab CI
Cybersecurity
Keycloak
Management
Confluence
Jira
QA
Cypress
Jest
Playwright
Postman
Selenium
Swagger
Vitest
Apply
$44k – $58k per year • Remote/Hybrid • Full-Time • 2+ years exp • Associate's Degree • Vancouver
Bash
Python
SQL
Databases
MS SQL
MySQL
Oracle
PostgreSQL
DevOps
AWS
Azure
GCP
VMWare
Windows Server
Apply
$86k – $187k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Vancouver
C#
JavaScript
Python
Java
Java
Spring Boot
Mobile
JUnit
DevOps
Azure
Azure DevOps
CI/CD
Git
GitHub Actions
GitLab CI
Jenkins
Rest API
QA
Rest-Assured
Selenium
TestNG
Apply
$79k – $97k per year • Remote • Full-Time • Vancouver
JavaScript
SQL
PHP
PHP
WordPress
Design
Figma
Webflow
Marketing
GA4
Hotjar
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.