582,867open jobs
25,570companies
81,102added this week
Browse all
Salary
$138k – $297k per year (Estimated)
Location
Remote (United States, France, United Kingdom, Canada, Poland)
Employment
Full-Time
Overview
Company
Impact
Profile match
Frontier AI lab building architectures and models that autonomously reason, learn, and evolve.

About Pathway

Pathway builds the first post-transformer frontier model that solves AI's fundamental memory problem. While transformers wake up in the same state every time-like Groundhog Day-our architecture enables true continuous learning, infinite context reasoning, and real-time adaptation. We're not optimizing yesterday's technology; we're building what comes after transformers.

Our breakthrough architecture outperforms Transformer and provides the enterprise with full visibility into how the model works. Combining the foundational model with the fastest data processing engine on the market, Pathway enables enterprises to move beyond incremental optimization and toward truly contextualized, experience-driven intelligence. We are trusted by organizations such as NATO, La Poste, and Formula 1 racing teams.

Pathway is led by co-founder & CEO Zuzanna Stamirowska, a complexity scientist who created a team consisting of AI pioneers, including CTO Jan Chorowski who was the first person to apply Attention to speech and worked with Nobel laureate Geoff Hinton at Google Brain, as well as CSO Adrian Kosowski, a leading computer scientist and quantum physicist who obtained his PhD at the age of 20.

The company is backed by leading investors and advisors, including Lukasz Kaiser, co-author of the Transformer (“the T” in ChatGPT) and a key researcher behind OpenAI’s reasoning models. Pathway is headquartered in Palo Alto, California.

The Opportunity

You will design and execute rigorous benchmarks and define dataset standards. Collaborating closely with our R&D team, you will build the evaluation infrastructure that guides the evolution of Pathway’s post-transformer models.

You Will

  • Proactively identify, prioritize, and curate relevant public and client-driven benchmarks across our target use cases and markets.
  • Evaluate candidate benchmarks for clarity, data quality, evaluation methodology, and fit with our model roadmap.
  • Run benchmarks with baseline models to validate setup, uncover edge cases, and de-risk R&D runs.
  • Hand off “benchmark-ready” packages to R&D (specs, data, evaluation scripts, expected metrics, constraints)
  • Maintain a shared vocabulary and documentation around benchmarks, datasets, and evaluation formats that GTM and R&D can both use.
  • Track and organize benchmark results, model leaderboards, and “what good looks like” for different customers and scenarios.
  • Contribute to demos and public-facing proof points based on benchmark outcomes.

You will play a key role in defining and driving the benchmarking process for AI model evaluation. Your work will directly influence what we build, how we talk about it, and how customers and the market experience BDH.

Requirements

Cover letter

It's always a pleasure to say hi! If you could leave us 2-3 lines, we'd really appreciate that.

You are expected to meet at least one of the following criteria:

  • You have published at least one paper at NeurIPS, ICLR, or ICML - where you were the lead author or made significant conceptual & code contributions.
  • You have significantly contributed to an LLM training effort which became newsworthy (topped a Hugging Face benchmark, best in class model, etc.), preferably using multiple GPU's.
  • You have spent at least 6 months working in a leading Machine Learning research center (e.g. at: Google Brain / Deepmind, Apple, Meta, Anthropic, Nvidia, MILA).
  • You were an ICPC World Finalist, or an IOI, IMO, or IPhO medalist in High School.

You

  • Have experience with ML/LLM evaluation, data science, or technical product roles, ideally around benchmarks or experimentation.
  • Are comfortable reading papers, leaderboards, and Github repos, and turning them into clear, repeatable benchmark specs.
  • Can talk comfortably with both engineers and customers, and translate between technical detail and business value.
  • Care about high-quality data, reproducible experiments, and crisp documentation
  • Arerespectful of others
  • Arefluent in English

Bonus Points

  • Published or open-sourced work on LLM evaluation, benchmarking or data quality.
  • Experience designing custom benchmarks or evaluation protocols for novel model capabilities.

Why You Should Apply

  • Join an intellectually stimulating work environment.
  • Be a pioneer: you get to work with a new type of "Live AI" challenges around long sequences and changing data.
  • Be part of one of an early-stage AI startup that believes in impactful research and foundational changes.

Benefits

  • Type of contract: Full-time, permanent
  • Preferable joining date: Immediate. The positions are open until filled - please apply immediately.
  • Compensation: based on profile and location.
  • Location: Remote work. Possibility to work or meet with other team members in one of our offices: Palo Alto, CA; Paris, France or Wroclaw, Poland. Candidates based anywhere in the EU, UK, United States, and Canada will be considered.

If you meet our broad requirements but are missing some experience, don’t hesitate to reach out to us.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
582,867 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Palo Alto
In office • Contractor • 1+ year exp • Bachelor's Degree • Hanoi
Python
AI/ML
Copilot
Cursor
Claude Code
Model Context Protocol
vLLM
Multimodal AI
Function Calling
AI Agents
Gemini
LLM
RAG
Reranking
Triton
OpenAI
Anthropic
OCR
Structured Outputs
LLM Guardrails
Speculative Decoding
KV Cache
Prompt Caching
Tool Use
DevOps
Azure
CI/CD
Git
Docker
Kubernetes
Azure AKS
Apply
$26k – $62k per year (Estimated) • In office • Full-Time • 12+ years exp • Taguig
AI/ML
ChatGPT
AI Agents
LLM
Perplexity
LLM Guardrails
Design
Sketch
Apply
Sr. AI Engineer 1 day ago
Remote/Hybrid • 10+ years exp
Python
C#
C#
.NET
Databases
Weaviate
Milvus
Pinecone
FAISS
AI/ML
LangGraph
LangChain
LlamaIndex
LoRA
Vertex AI
Fine-tuning
Embeddings
RLHF
Reinforcement Learning
Prompt Engineering
Function Calling
AI Agents
PEFT
QLoRA
RAG
OpenAI
Anthropic
Amazon SageMaker
DPO
SFT
PPO
GRPO
LLM Guardrails
Multi-Agent Systems
Tool Use
DevOps
GCP
Azure
AWS
Kubernetes
Vector
Apply
AI Engineer 1 day ago
$22k – $59k per year (Estimated) • In office • 3+ years exp • Bengaluru
Python
Python
FastAPI
Databases
PostgreSQL
Weaviate
Chroma
Milvus
pgvector
Pinecone
AI/ML
LangGraph
AutoGen
LangChain
Claude
LlamaIndex
vLLM
Fine-tuning
Prompt Engineering
Multimodal AI
AI Agents
TensorRT
TensorRT-LLM
TGI
Llama
Mistral
TensorFlow
PyTorch
CrewAI
Gemini
LLM
RAG
Hallucination
Agentic Workflows
Multi-Agent Systems
DevOps
Rest API
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Vector
Apply
In office • 5+ years exp • Bachelor's Degree
AI/ML
LLM
Marketing
LinkedIn
Apply
$87k – $253k per year (Estimated) • Remote/Hybrid • Internship • Bachelor's Degree • Palo Alto
AI/ML
ChatGPT
Transformers
OpenAI
Apply
$139k – $299k per year (Estimated) • Equity • Remote • Full-Time • PhD • Palo Alto
AI/ML
ChatGPT
Transformers
TensorFlow
PyTorch
LLM
Triton
OpenAI
Anthropic
Hugging Face
DevOps
CI/CD
Apply
$108k – $241k per year (Estimated) • Remote • Full-Time • Bachelor's Degree • Palo Alto
Python
AI/ML
ChatGPT
MLFlow
Vertex AI
Kubeflow
Transformers
TensorFlow
PyTorch
LLM
OpenAI
Amazon SageMaker
Metaflow
DevOps
Terraform
GCP
GitHub Actions
Loki
CloudFormation
Prometheus
GitLab CI
SLURM
Azure
CI/CD
Jenkins
AWS
Docker
Kubernetes
Grafana
Amazon CloudWatch
Apply
$165k – $260k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Houston • Palo Alto
Python
Go
JavaScript
Rust
TypeScript
PowerShell
C#
C++
Node JS
Rust
Tauri
C++
TensorFlow C++
PyTorch C++
AI/ML
Embeddings
Quantization
Multimodal AI
Computer Vision
AI Agents
TensorFlow
PyTorch
LLM
RAG
Edge AI
ONNX Runtime
Agentic Workflows
Frontend
Angular
React.js
DevOps
WebRTC
Azure
CI/CD
Git
Apply
$147k – $231k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Houston • Palo Alto
Python
Go
JavaScript
Rust
TypeScript
PowerShell
C#
C++
Node JS
Rust
Tauri
C++
TensorFlow C++
PyTorch C++
AI/ML
Fine-tuning
Quantization
Multimodal AI
Knowledge Distillation
Function Calling
Computer Vision
AI Agents
Speech Recognition
TensorRT
OpenVINO
Transformers
TensorFlow
PyTorch
LLM
RAG
Hugging Face
OCR
Text-to-Speech
Human-in-the-Loop
LLM Evaluation
LLM Guardrails
Edge AI
LiteRT
ONNX Runtime
Multi-Agent Systems
Tool Use
Model Distillation
Frontend
Angular
React.js
DevOps
gRPC
GCP
WebRTC
WebSockets
Azure
CI/CD
Git
AWS
Robotics
Localization
Apply
$62k – $70k per year • In office • Full-Time • Palo Alto
Apply
$39k – $82k per year • In office • 2+ years exp • Palo Alto
QA
Rest-Assured
Apply
$148k – $300k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Palo Alto
Python
JavaScript
TypeScript
Node JS
AI/ML
Mistral
LLM
Edge AI
Frontend
Next.js
React.js
Apply
See all jobs
This is one of many
582,867 more open roles from verified company boards, updated every day.