377,351open jobs
9,828companies
48,306added this week
Browse all
Location
In office (Prague)
Seniority
Staff · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Automate complex transactional workflows with Rossum’s AI document processing solution. Reduce manual tasks, increase accuracy, drive efficiency.

About Coupa

Coupa is the platform companies run their spending on - sourcing, procurement, invoices, payments, suppliers, contracts. It is the system of record for how large organisations decide what to buy, from whom, and at what price.

That makes it one of the most unusual datasets in enterprise software: a single network of 10M+ buyers and suppliers, and $10 trillion of transacted spend to date - quotes, bids, awards, orders, invoices, contracts, and the documents behind every one of them. Multimodal, longitudinal, and tied to outcomes measured in real money.

About The Team

Rossum joined Coupa earlier this year. We brought the document understanding layer - our proprietary T-LLM (transactional LLM) architectures, which we design and train from scratch, and which read the world's messiest business documents in production, millions of them every week. Now we are pointing the same in-house research capability at a much bigger problem: not just reading the documents, but acting on them.

About the Role

Sourcing is where the money is actually decided.

We are expanding our Data Capture Research team in Prague with a Senior Data Scientist to work on Sourcing.

Which suppliers get invited. How the event is structured. How bids that differ in price, lead time, quality, risk and carbon get compared at all. What a fair price even is. When to award, to whom, and how to split the award across suppliers.

For a researcher this is unusually open ground - not one model family, but several, on the same data:

  • Recommendation and retrieval. Supplier discovery: matching demand to the right suppliers across a 10M-node network.

  • Forecasting and should-cost modelling. What this category, in this region, at this volume, should cost right now.

  • Game theory and mechanism design. Auction formats, bidding behaviour, incentives, and competitive dynamics between real counterparties.

  • Combinatorial optimisation. Award allocation under volume, capacity and multi-sourcing constraints.

  • Multimodal document understanding. RFPs, specs, quotes and contracts carry the actual requirements - and our T-LLM foundations already give us a head start there.

Part of the work is choosing the right instrument for each - neural networks, gradient boosting, optimisation, bandits, mechanism design - instead of forcing one.

Very little of this has been built with modern ML yet. That is the point of the role: real greenfield problems, a dataset nobody else has, and a product that ships to companies whose margins depend on getting these decisions right.

You will work in a small, senior team of researchers and engineers - the group that built Rossum's production models from scratch - with direct access to Product and the AI Platform team. Ideas that work do not sit on paper; they roll into systems used at scale.

What You'll Do

  • Own sourcing research initiatives end to end - from framing the problem on real spend data, through experiments, to a model running in production.

  • Build what does not exist yet. Most of these problems have no baseline to beat and no off-the-shelf answer - you decide what the first version looks like.

  • Turn the network into a training set. Define the datasets, labels and benchmarks that make sourcing problems learnable at all: deriving supervision from historical events and their outcomes, and building evaluation the team can trust.

  • Design and run bold experiments with a hacker mindset - fast prototypes, honest baselines and offline evaluation that actually predicts online behaviour.

  • Build on our T-LLM foundations wherever documents carry the signal - RFPs, specs, quotes, contracts - reusing in-house architectures we train ourselves.

  • Ship with deployment in mind: inference cost, latency, robustness, and what happens when a recommendation is wrong in front of a buyer.

  • Work across the company - Coupa's Sourcing product and data teams and our AI Platform team. Document decisions and make the people around you better.

Who You Are

Sourcing needs several kinds of modelling, so we are deliberately open about which one you bring.

  • 5+ years in applied ML, data science, ML engineering or quantitative research, with models or decision systems you took into production.

  • Strong Python and real comfort with messy, large-scale data - including SQL and the unglamorous work of making a dataset trustworthy.

  • Depth in at least one modelling discipline, curiosity about the rest. Deep learning, recommendation and ranking, forecasting and time series, optimisation and operations research, causal inference and econometrics, RL and bandits, market and mechanism design, or LLM-based systems. Sourcing touches most of these; nobody arrives holding all of them.

  • Scientific rigour. Strong experiment design, healthy scepticism about your own metrics, and real care about leakage, baselines, and evaluation that survives contact with production.

  • Ownership and curiosity. You are comfortable in a greenfield, ambiguous problem space, and you will talk to product people and procurement experts to find where the value actually is.

  • Interest in procurement, supply chains or market design is welcome but not required - we will teach the domain.

Why Join Us

  • Greenfield problems in a mature product: Modern ML has barely been applied to sourcing, inside a platform that already has the users, the workflows and the data.

  • Data nobody else has: $10 trillion of transacted spend, 10M+ buyers and suppliers, and the documents behind all of it.

  • We train our own models: Proprietary T-LLM architectures, designed and trained in-house - not a wrapper around someone else's API.

  • Real ownership, short path to customers: You frame the problem, choose the method, and see it working in front of buyers - no research-to-product handoff.

  • Global impact: Technology used every day by companies around the world.

  • Experiment-driven culture: Pragmatic delivery, and quarterly recognition for standout research contributions.

  • Compute and tools: Frontier LLMs on tap for your own work, and our high-end GPU and large-memory clusters to train on.

  • Conferences: A budget to attend the ones that matter in your field.

  • 33 days off: PTO, personal days, your birthday and two company wellness days. Parental leave on top.

  • Prague, Karlín: Inspiring workspace and full tech setup, including a 200 m² terrace with views of Prague Castle.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
377,351 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Prague
$75k – $110k per year • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Boston
Python
R
SQL
R
ggplot2
Databases
Snowflake
AI/ML
Pandas
Analytics
Tableau
Apply
$111k – $173k per year • Equity • In office • Full-Time • PhD • Waltham
Python
SQL
Databases
Databricks
Snowflake
AI/ML
AI Agents
Scikit-learn
Statsmodels
Management
Slack
Apply
Lead Data Engineer 1 day ago
$110k – $204k per year • Equity • Remote/Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Frisco • Toronto
Python
SQL
Databases
Snowflake
AI/ML
Claude
Copilot
Analytics
ETL/ELT
Power BI
Tableau
Apply
$84k – $157k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Arlington
PowerShell
Python
AI/ML
AI Agents
LLM
Model Context Protocol
DevOps
API Gateway
Azure
Azure DevOps
CI/CD
Configuration Management
Docker
IAM
Kubernetes
Platform Engineering
Terraform
Cybersecurity
Least Privilege
Cryptography
Vault
Apply
$83k – $177k per year (Estimated) • In office • Full-Time • 10+ years exp • Wellington
PowerShell
Python
SQL
DevOps
Ansible
AWS
Azure
CI/CD
Platform Engineering
Terraform
Windows Server
Apply
In office • Full-Time • 5+ years exp • Prague
Python
SQL
AI/ML
LLM
Multimodal AI
Time Series Forecasting
Apply
Lead AI Engineer 1 day ago
Remote/Hybrid • Full-Time • 5+ years exp • Prague
Python
SQL
AI/ML
LLM
Multimodal AI
Time Series Forecasting
Apply
In office • Full-Time • Prague
AI/ML
Edge AI
Apply
In office • Full-Time • Master's Degree • Prague
Python
Python
Django
Flask
Databases
PostgreSQL
AI/ML
Edge AI
DevOps
Kubernetes
Rest API
Apply
Remote/Hybrid • Full-Time • 4+ years exp • Prague
AI/ML
Edge AI
DevOps
AWS
CI/CD
GCP
Kubernetes
Platform Engineering
Apply
Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Prague
Analytics
Power BI
Apply
Remote/Hybrid • 3+ years exp • Prague
C#
PowerShell
C#
.NET
DevOps
AWS
Azure
Azure DevOps
GCP
Git
GitLab
Jenkins
TeamCity
QA
Selenium
TestRail
Apply
In office • Prague
Java
Apply
In office • Prague
Java
Kotlin
DevOps
AWS
Azure
CI/CD
GCP
Kubernetes
VMWare
Apply
Remote/Hybrid • Full-Time • Prague
Python
Python
pySpark
Databases
Databricks
DynamoDB
AI/ML
Spark
DevOps
AWS
AWS Lambda
CI/CD
GCP
Terraform
Amazon S3
AWS Step Functions
IAM
Apply
See all jobs
This is one of many
377,351 more open roles from verified company boards, updated every day.