368,530open jobs
9,432companies
50,439added this week
Browse all
Salary
$98k – $234k per year (Estimated)
Location
In office
Seniority
Principal
Employment
Full-Time
Overview
Company
Impact
Profile match
Phrase is the AI-powered language intelligence platform for enterprise. Automate workflows, connect any tool or engine, and scale multilingual content globally.

Principal AI Researcher

At Phrase, we help open the door to global business by providing the world's leading Language Intelligence Platform.

The Phrase Platform combines AI, agentic orchestration, and a headless, API-first architecture in one composable system. Beyond translation, it orchestrates and adapts content to culture, audience, channel, brand voice, and intended outcome in any language, for every audience. It applies the context that makes content perform in every market: quality standards, glossaries, prior translations, and cultural nuance. Every team, in every region, can ship content that is on-brand, on-point, and ready for any audience.

Phrase gives enterprises the intelligence to automate workflows, the freedom to connect their own tools and engines, and the control to govern global content at scale.

The AI Research team sits at the core of Phrase’s product differentiation. We build, train, and evaluate proprietary translation models; we design the agentic workflows that orchestrate translation quality end-to-end; and we run the research cadence that feeds product development with grounded, evidence-based decisions.

The team’s current research programmes span: fine-tuning LLMs for domain-specific machine translation, multi-agent evaluation pipelines, automated quality profiling from style guides, and active learning through feedback loops on real customer data. These are not exploratory proofs of concept; they are live systems serving production traffic.

The next phase requires someone who can take ownership of the most technically demanding of these streams, extend them, and help chart the research direction for the team as a whole.

What you’ll be responsible for:

Core Research

  • Lead design, training, and evaluation of LLMs for translation and language quality tasks, including work with fine-tuning techniques such as LoRA and DPO, and instruction-tuned models at various scales.

  • Design and implement robust evaluation frameworks for translation quality, moving beyond automatic metrics (BLEU, TER, ChrF, MQM, COMET), LLM-as-judge approaches, and hybrid evaluation pipelines.

  • Architect and evaluate complex agentic workflows for NLP problems: multi-step reasoning, tool use, structured output generation, and orchestration across multiple model providers.

  • Take technical ownership of the team’s flagship model development programme, from data curation and training pipeline design through to production integration and ongoing evaluation.

  • Design and run experiments using ML pipelines, maintaining reproducibility and clear documentation of results and decisions.

Applied NLP and Product Collaboration

  • Work closely with Product, Engineering, and Solutions to translate research findings into concrete, shippable capabilities.

  • Provide expert input on model provider strategy: evaluation of frontier models (OpenAI, Gemini, Claude, open-source), benchmarking against internal baselines, and recommendations on production use.

  • Contribute to the evolution of quality evaluation systems, including integration of style guides, and a focus on customer-led, outcome-oriented metrics and evaluation methodologies.

  • Support decisions on active learning, feedback loop design, and data pipeline strategy for continuous model improvement.

Team and Leadership

  • Act as a technical lead for the research team: setting direction on open problems, reviewing PRs and research write-ups, running or contributing to internal enablement sessions such as reading groups, deep dives etc.

  • Mentor junior and mid-level researchers; model good research hygiene and collaborative working practices.

  • Contribute to hiring decisions for the team, including take-home assignment design and technical interviews.

  • Represent the AI Research team in cross-functional forums and with external stakeholders where research credibility matters.

What We’re Looking For

Must have

  • Deep, hands-on experience building and training LLMs or large transformer-based models: pre-training, fine-tuning, RLHF/DPO, or equivalent.

  • Strong background in NLP, with demonstrable expertise in at least one of: machine translation, MT evaluation, information extraction, text generation, search optimization, or reinforcement learning.

  • Proven ability to design, implement, and critically evaluate agentic systems: tool-calling agents, multi-agent pipelines, or LLM-orchestrated workflows at scale.

  • Experience designing and running rigorous evaluation frameworks, including both automatic metrics and human-in-the-loop eval design.

  • Strong Python skills and experience with ML/NLP libraries (Hugging Face, PyTorch, vLLM, or similar); familiarity with workflow orchestration tools such as Flyte, Prefect, or Airflow.

  • PhD in Computer Science, Computational Linguistics, or a related field, OR equivalent research experience demonstrated through publications, open-source contributions, or significant applied research output.

Strongly Preferred

  • Industry experience working on production NLP systems, not just academic research: shipping models, handling data pipelines, monitoring quality in live traffic.

  • Familiarity with the localization and machine translation domain: translation quality metrics (MQM, COMET), post-editing workflows, segment-level quality signals, or TMS integrations.

  • Experience with multi-LLM ensemble architectures or model routing strategies.

  • Track record of publishing or presenting at top NLP venues (ACL, EMNLP, NAACL, AAAI, NeurIPS, ICLR).

  • Experience with experiment tracking tools (Weights & Biases), model serving frameworks (vLLM, TGI), and cloud-based ML infrastructure (AWS SageMaker, GCP Vertex AI).

  • Experience with PydanticAI, LangGraph, or similar frameworks for building structured agentic pipelines.

  • Prior experience in a technical lead or staff research role; evidence of mentoring or growing other researchers.

Character and Working Style

  • Product-minded: you think about research in terms of what it enables and who it serves, not just whether the numbers went up.

  • A natural driver: you identify open problems, propose directions, and move work forward without waiting to be pushed.

  • Collaborative and open: you share results early, invite challenge, give credit generously, and build on others’ work.

  • Precise but not precious: you care about rigour and reproducibility without losing sight of practical impact.

  • Growth-oriented: you are motivated by the prospect of building and leading a team, not just doing individual research.

  • Generative: you bring new ideas to the table, not just rigorous execution of known approaches; intellectual curiosity and a willingness to pursue directions that might not work are as important to us as technical depth.

What you’ll get:

  • Work experience in a successful and growing global SaaS company

  • Be part of an international team in Europe, APAC, and the Americas

  • Expert colleagues in their field who are determined to build the best localization platform on the market, creating a world where language never limits opportunity

  • An agile work environment, where it is encouraged to take smart risks

  • Take part in a culture full of trust, support and loyalty, where respectful and open feedback is valued, and diversity is fully embraced

  • A positive, open-minded, and innovative atmosphere

  • Support in your professional development and personal career goals

What's on top:

  • 4 Company holidays additional to your regular holidays (1 day per quarter where the entire company is off to celebrate our achievements).

  • In addition, the company also provides employees with a Christmas break to allow you to spend time with family and friends without use of your vacation allocation.

  • Your birthday is off because it is important to celebrate you as well.

  • 2 Giveback days where you can support the local community, volunteer, and/or participate in charity events and activities.

  • Professional and extensive onboarding.

  • Learning & development through Phrase learning

  • Additional local benefits depending on the entity you're hired at, just ask your Talent Acquisition Partner

At Phrase we believe in the critical importance of diversity in all its forms and intersectionalities, and are committed to ensuring that the people we interview reflect this diversity. We therefore strongly encourage people of any identity to apply for our exciting career opportunities. We have taken an active decision to make our work environment inclusive every day, starting at the very beginning of your Phrase experience.

We value and welcome different perspectives, experiences and backgrounds as we believe that these differences make our team even stronger on our mission of opening the door to global business by giving everybody access to the content they need in the language they speak.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$44k – $58k per year • Remote/Hybrid • Full-Time • 2+ years exp • Associate's Degree • Vancouver
Bash
Python
SQL
Databases
MS SQL
MySQL
Oracle
PostgreSQL
DevOps
AWS
Azure
GCP
VMWare
Windows Server
Apply
Lead Data Engineer 4 hours ago
$140k – $231k per year • Remote/Hybrid • Full-Time • Bachelor's Degree • O'Fallon
Java
Python
SQL
Python
pySpark
Databases
Apache Kafka
Databricks
Delta Lake
AI/ML
Airflow
Hadoop
Spark
DevOps
AWS
Azure
CI/CD
GCP
Git
GitHub
Analytics
ETL/ELT
Apply
$73k – $186k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Toronto
Java
Python
SQL
Databases
Databricks
AI/ML
Spark
DevOps
AWS
Azure
Bitbucket
GCP
Git
Analytics
ETL/ELT
Apply
$88k – $224k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Waterloo
Java
Python
SQL
Databases
Databricks
AI/ML
Spark
DevOps
AWS
Azure
Bitbucket
GCP
Git
Analytics
ETL/ELT
Apply
Remote/Hybrid • Internship • Bachelor's Degree • Toronto
Java
Python
SQL
Databases
Databricks
AI/ML
Spark
DevOps
AWS
Azure
Bitbucket
GCP
Git
Analytics
ETL/ELT
Apply
Product Manager 1 month ago
$93k – $192k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree
AI/ML
Claude
AI Agents
Apply
Platform Engineer 1 month ago
In office • Full-Time
AI/ML
LLM
AI Agents
LLM Guardrails
DevOps
AWS
CI/CD
Git
Kubernetes
Platform Engineering
Cybersecurity
ISO 27001
SOC 2
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.