389,999open jobs
10,334companies
49,652added this week
Browse all
Salary
$78k – $184k per year (Estimated)
Location
Remote/Hybrid (Paris, France, Palo Alto, New York, United States, Zurich, Switzerland, Amsterdam, Netherlands)
Employment
Full-Time
Overview
Company
Impact
Profile match
Mistral AI is a French artificial intelligence company founded in Paris in 2023 by former researchers from Google DeepMind and Meta. It builds efficient open-weight and commercial large language models, including the Mistral, Mixtral, Magistral and Codestral families, and distributes them through its own developer platform as well as the major cloud marketplaces. Positioned as Europe's leading frontier model developer, the company also ships the Le Chat assistant, on-premise deployment options for regulated industries and a sovereign compute offering for customers who must keep data inside the region.

About Mistral

Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems-across high-stakes industries like finance, manufacturing, defense, healthcare, and the public sector-co-creating customized AI systems that they can run on their terms.

We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Our teams are distributed between Europe, North America, Asia and the Middle East. We are creative, low-ego and team-spirited.

Role summary

As a Research Engineer on Forge, you will turn real customer requirements into reliable training and deployment workflows. You’ll work end-to-end across model adaptation and post-training (CPT/SFT/RL/distillation), evaluation, data, and infrastructure. The role bridges research experimentation and production constraints.

This role sits in Applied Science, with direct impact on client outcomes. You’ll collaborate closely with scientists, engineers, product, and customer-facing teams to ensure Forge projects ship, are maintainable, and can be trusted by others.

Interview focus can vary (algorithms, infrastructure, evals, or data). You don’t need to match every bullet below to apply.

What you will do

  • Build and improve post-training and evaluation workflows (CPT/SFT/RL/distillation), turning prototypes into repeatable Forge “recipes”.

  • Develop tools and pipelines for synthetic data generation, data curation, training, evaluation, and deployment.

  • Debug and harden large-scale ML systems: distributed training, scheduling/execution, checkpointing, observability, and reproducibility.

  • Improve the Forge codebase via clear APIs, tests, documentation, and maintainable abstractions.

  • Push the frontier of our RL training stack (e.g., high-throughput async rollout and scalable post-training systems at frontier-model scale)

  • Make sure Forge deployment is seamless and adaptable to a diversity of clients (hardware access, software stack, cloud and on-premises, …)

  • Partner with researchers and infrastructure engineers to translate bottlenecks into concrete system improvements.

About you

  • Strong Python engineering skills and experience working in large codebases (testing, code review, CI, operational ownership).

  • Hands-on experience with PyTorch, JAX, or similar.

  • Strong systems and infrastructure fundamentals.

  • Experience with LLM training or post-training: fine-tuning, RL, distillation, evaluation, and/or data pipelines.

  • Excellent debugging skills in ambiguous systems (distributed jobs, data issues, quality regressions, infra failures).

  • Clear communication with technical and non-technical stakeholders.

  • High agency, low ego, and comfort in fast-moving, under-specified environments.

Nice to have

  • Distributed training experience (FSDP, DeepSpeed, Megatron, etc.).

  • Cluster/orchestration experience (SLURM, Ray, Kubernetes, Kueue, Karpenter, Skypilot, etc.).

  • Experience building reliable ML infrastructure, evaluation systems, or large-scale data processing pipelines.

  • Research experience in LLMs, agents, multimodal models, reasoning, code, or domain adaptation.

  • Open-source contributions, publications, or widely used internal tooling.

  • Experience training multi-billion-parameter models (pre-training or RL). Experience training on petabyte- and exabyte-scale datasets.

  • Ability to identify bottlenecks across the stack and drive improvements from first principles.

What We Offer

We offer a comprehensive benefits package designed to support your well-being, growth, and work-life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks.

For the most up-to-date details on benefits available in your location, please refer to our Benefits page.

Privacy Policy

Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant Privacy Policy.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
389,999 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Paris
$17k – $50k per year (Estimated) • Remote • Full-Time • Moscow
Python
Databases
Apache Kafka
Kafka
Redis
AI/ML
AI Agents
Anthropic
Embeddings
Fine-tuning
Function Calling
Hugging Face
LLM
LoRA
MLFlow
OpenAI
PEFT
RAG
Structured Outputs
Transformers
DevOps
Grafana
Kubernetes
Prometheus
Management
Confluence
SharePoint
Apply
AI Engineer AWS F/H 6 hours ago
$50k – $127k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Bordeaux
Python
SQL
Databases
Amazon Aurora
DynamoDB
OpenSearch
pgvector
Pinecone
PostgreSQL
AI/ML
AWS Bedrock
AWS Bedrock AgentCore
Embeddings
LangChain
LangSmith
LlamaIndex
LLM
LLM Guardrails
LLMOps
Prompt Engineering
RAG
Ragas
Reranking
DevOps
Amazon CloudWatch
Amazon S3
API Gateway
AWS
AWS Lambda
AWS Step Functions
CI/CD
CloudFormation
IAM
Terraform
Vector
Apply
$224k – $478k per year (Estimated) • In office • Contractor • 8+ years exp • Bachelor's Degree • San Francisco
AI/ML
AI Agents
Anthropic
Claude
Claude Code
Function Calling
Instructor
LLM
Model Context Protocol
Multimodal AI
Vertex AI
Apply
$25k – $70k per year (Estimated) • In office • Full-Time • 3+ years exp • Bengaluru
Python
SQL
Databases
OpenSearch
pgvector
Pinecone
Snowflake
Weaviate
PostgreSQL
AI/ML
AI Agents
Amazon SageMaker
AutoGen
AWS Bedrock
AWS Bedrock AgentCore
AWS Strands Agents
CrewAI
Embeddings
KServe
Kubeflow
LangChain
Langfuse
LangGraph
LangSmith
LLM
LLM Guardrails
MLFlow
Model Context Protocol
Pydantic AI
RAG
Ray
Ray Serve
Semantic Kernel
TGI
vLLM
DevOps
Amazon EC2
Amazon EKS
Amazon S3
AWS
AWS Lambda
CI/CD
Datadog
GitHub
GitHub Actions
Grafana
IAM
Jenkins
KEDA
Kubernetes
Loki
Prometheus
Terraform
Vector
Apply
Principal AI Engineer 7 hours ago
$32k – $75k per year (Estimated) • In office • Full-Time • 9+ years exp • Bachelor's Degree • Hyderabad
Python
Python
Asyncio
FastAPI
Databases
PostgreSQL
Redis
AI/ML
A2A
AI Agents
Anthropic
AutoGen
AWS Bedrock
Claude
CrewAI
Function Calling
GPT-4
LangChain
Langfuse
LangGraph
LangSmith
LLM
LLMOps
Model Context Protocol
OpenAI
Prompt Engineering
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
WebSockets
Apply
$151k – $331k per year (Estimated) • Remote/Hybrid • Full-Time • Master's Degree • San Francisco • Palo Alto
Python
AI/ML
Mistral SDK
PyTorch
Apply
$95k – $200k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Palo Alto
Go
Python
AI/ML
CUDA
CUDA Toolkit
Fine-tuning
Mistral SDK
NCCL
PyTorch
DevOps
Karpenter
Kubernetes
Platform Engineering
Cybersecurity
Kyverno
Apply
$58k – $154k per year (Estimated) • In office • Full-Time • Paris
C#
Go
Python
TypeScript
Python
Django
FastAPI
Flask
Apply
$77k – $184k per year (Estimated) • In office • Full-Time • Master's Degree • Paris • Amsterdam • Linz • Munich • London
Python
AI/ML
JAX
DevOps
HPC
Apply
$76k – $180k per year (Estimated) • Remote/Hybrid • Full-Time • Paris • Amsterdam • Linz • Munich • London
Python
AI/ML
Fine-tuning
LLM
Mistral
RLHF
Post-training
SFT
AI Agents
Function Calling
DevOps
HPC
Design
SolidWorks
Apply
$58k – $111k per year (Estimated) • In office • Full-Time • Paris
Apply
$70k – $152k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Paris • Amsterdam
DevOps
AWS
Azure
Marketing
LinkedIn
Apply
$43k – $106k per year (Estimated) • In office • Contractor • Paris
PHP
PHP
Drupal
Symfony
DevOps
AWS
CI/CD
Cloudflare
Datadog
GitLab
GitLab CI
Kubernetes
Management
Confluence
Jira
Miro
Apply
$69k – $130k per year (Estimated) • In office • Full-Time • 5+ years exp • Paris
DevOps
AWS
Azure
Apply
Release Manager H/F 8 hours ago
$66k – $158k per year (Estimated) • In office • Contractor • 10+ years exp • Paris
AI/ML
Supervision
DevOps
ArgoCD
AWS
CI/CD
Datadog
Docker
Git
GitHub
GitHub Actions
GitLab
GitLab CI
Grafana
Jenkins
Kubernetes
Prometheus
Terraform
Cybersecurity
SonarQube
Management
Confluence
Jira
Miro
ServiceNow
Apply
See all jobs
This is one of many
389,999 more open roles from verified company boards, updated every day.