695,350open jobs
40,785companies
98,938added this week
Browse all
Salary
$50k – $115k per year (Estimated)
Location
In office (Paris)
Seniority
Senior · 5+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Mistral AI is a French artificial intelligence company founded in Paris in 2023 by former researchers from Google DeepMind and Meta. It builds efficient open-weight and commercial large language models, including the Mistral, Mixtral, Magistral and Codestral families, and distributes them through its own developer platform as well as the major cloud marketplaces. Positioned as Europe's leading frontier model developer, the company also ships the Le Chat assistant, on-premise deployment options for regulated industries and a sovereign compute offering for customers who must keep data inside the region.

About Mistral

Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems across high-stakes industries like finance, manufacturing, defense, healthcare, and the public sector, co-creating customized AI systems that they can run on their terms.

We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Our teams are distributed between Europe, North America, Asia and the Middle East. We are creative, low-ego and team-spirited.

Mistral AI is building a new Compute Support team to ensure the reliability, performance, and scalability of our GPU clusters in a Kubernetes environment. As one of the founding members of this team, you will play a pivotal role in shaping its processes, standards, and culture.

In this hybrid L1/L2 support role, you will be the first point of contact for customers and internal teams, providing technical guidance, troubleshooting, and issue resolution for compute-related inquiries. You’ll also serve as the escalation point for complex issues, leveraging your deep systems knowledge, Kubernetes expertise, and operational debugging skills to ensure our AI workloads run smoothly at scale.

This is a support-focused role that blends system administration, customer-facing communication, and compute infrastructure expertise. While coding is not the primary focus, familiarity with Go for debugging and automation is a plus.

Key Responsibilities

First Customer’s Point of Contact

  • Act as the first interlocutor for customers and internal teams, providing timely and effective responses to compute-related inquiries.

  • Triage and prioritize incoming requests, ensuring SLAs are met for acknowledgment, first response, and resolution.

  • Gather and analyze initial issue details (e.g., logs, error messages, system metrics) to diagnose problems efficiently.

  • Provide clear, actionable guidance to customers and internal users, including temporary workarounds where applicable.

Technical Support

  • Serve as the L2 escalation point for Linux, Kubernetes, and compute infrastructure issues, providing deep technical troubleshooting for:

    • GPU/TPU workloads (e.g., CUDA errors, memory leaks, job failures).

    • Kubernetes clusters (e.g., pod crashes, node failures, networking misconfigurations).

    • Bare metal and cloud environments (e.g., AWS EC2, GCP VMs, HPC clusters).

  • Diagnose and resolve performance bottlenecks, hardware failures, and resource contention in distributed systems.

  • Analyze system metrics, logs, and traces (e.g., dmesg, journalctl, nvidia-smi, Prometheus, Grafana) to identify root causes of issues.

  • Participate in on-call rotations to provide 24/7 support for critical compute systems

  • Nice to have: Optimize system configurations (e.g., kernel parameters, filesystem tuning, network settings) for high-performance computing (HPC) and AI workloads.

Kubernetes & Containerization

  • Debug Kubernetes clusters with a focus on:

    • Pod and node issues (e.g., CrashLoopBackOff, OOMKilled, ImagePullBackOff).

    • Networking and storage (e.g., CNI plugins, PersistentVolumes, StorageClasses).

    • Resource management (e.g., Requests/Limits, QOS classes, node affinity).

  • Troubleshoot container runtime issues (Docker, containerd) such as image pull failures, OCI compliance, or runtime errors.

  • Collaborate with SRE teams to understand and improve existing IaC (Terraform, Ansible) and Go-based tooling.

Documentation & Process Improvement

  • Create and maintain runbooks, playbooks, and internal documentation for common compute issues (e.g., GPU debugging, Kubernetes troubleshooting).

  • Contribute to post-mortems with actionable follow-ups to prevent recurring incidents.

  • Train internal teams on best practices for compute infrastructure and debugging.

  • Help define and refine support processes as a founding member of the new Compute Support team.

What We’re Looking For

Required Skills & Experience

  • 5+ years of experience in system administration, technical support, or infrastructure operations, with a strong focus on compute-heavy environments (bare metal, cloud, HPC, or virtualization).

  • Deep Linux/Unix expertise:

    • Debugging kernel-level issues (e.g., OOM killer, I/O bottlenecks, CPU throttling).

    • Performance tuning (e.g., sysctl, ulimit, filesystem optimizations).

    • Networking troubleshooting (e.g., iptables, tcpdump, DNS, NFS, SSH).

    • Storage management (e.g., LVM, RAID, NVMe, GPU-local storage).

  • Hands-on Kubernetes experience is mandatory:

    • Debugging pods, nodes, and clusters (e.g., kubectl describe, kubectl logs, crictl).

    • Networking and storage (e.g., CNI plugins, PersistentVolumes).

    • Resource management (e.g., Requests/Limits, QOS classes).

  • Familiarity with containerization (Docker, containerd) and basic IaC awareness Proficiency with monitoring tools (Prometheus, Grafana, ELK, OpenTelemetry).

  • Basic scripting/automation skills (preferably Go, but Bash or Python are acceptable for ad-hoc tasks).

  • Experience with bare metal, virtualization, or HPC environments (e.g., Fluidstack, Coreweave, Vast) is a strong plus.

  • Excellent problem-solving and communication skills, with the ability to explain technical issues clearly and patiently to both technical and non-technical stakeholders.

  • Customer-focused mindset with a passion for resolving issues efficiently and improving system reliability.

  • Ability to work in on-call rotations and handle high-pressure situations with a structured, calm approach.

  • Commitment to meeting tight SLAs for issue acknowledgment, triage, and first response.

What We Offer

We offer a comprehensive benefits package designed to support your well-being, growth, and work-life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks.

For the most up-to-date details on benefits available in your location, please refer to our Benefits page.

Privacy Policy

Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant Privacy Policy.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
695,350 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Paris
Senior DevOps Engineer 10 hours ago
$76k – $97k per year (gross) • Remote/Hybrid • Full-Time • Vilnius
Go
Lua
Databases
Redis
Apache Kafka
DevOps
Terraform
Ansible
Zabbix
Prometheus
HAProxy
CI/CD
ArgoCD
Kubernetes
Grafana
Linux
Management
Agile
Scrum
Apply
$43k per year (net) • Remote • Full-Time • Moscow
Python
JavaScript
TypeScript
Node JS
Node JS
Nest.JS
Databases
PostgreSQL
Redis
ElasticSearch
Apache Kafka
DevOps
Terraform
Helm
GitHub Actions
cert-manager
Prometheus
CI/CD
Kubernetes
Alertmanager
SLI/SLO/SLA
Linux
DNS
Cybersecurity
Tcpdump
Apply
$71k – $144k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Getafe
Python
PowerShell
Bash
Python
Flask
FastAPI
DevOps
Rest API
Terraform
Ansible
CI/CD
Jenkins
Git
Docker
Kubernetes
SOAP
Apply
In office • Internship • Paris
DevOps
Kubernetes
Linux
Apply
$17k – $44k per year (Estimated) • Remote • Kazan
Python
Bash
1C
Databases
Redis
RabbitMQ
Apache Kafka
AI/ML
Langfuse
DevOps
Terraform
Ansible
Helm
Azure DevOps
Istio
Loki
containerd
Prometheus
Yandex Cloud
GitLab CI
Azure
CI/CD
Git
Docker
Kubernetes
Grafana
Service Mesh
SLI/SLO/SLA
GitLab
IAM
Linux
DNS
Cybersecurity
Keycloak
HashiCorp Vault
PKI
Apply
$102k – $247k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Singapore
Go
AI/ML
Mistral
LLM
Apply
$107k – $222k per year (Estimated) • In office • Full-Time • 10+ years exp • Paris • London
Go
Apply
$71k – $128k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Paris • London
AI/ML
Model Context Protocol
AI Agents
Context Engineering
Edge AI
Apply
$47k – $77k per year (Estimated) • In office • Full-Time • 3+ years exp • Paris
AI/ML
Mistral SDK
Apply
$78k – $165k per year (Estimated) • In office • Full-Time • 5+ years exp • Singapore
Management
Slack
Notion
Google Workspace
Apply
GTM intern 6 hours ago
$19k per year • In office • Internship • 18+ years exp • Paris
SQL
AI/ML
dbt
AI Agents
DevOps
GitHub
Apply
up to $71k per year • Equity • In office • Paris
SQL
Apply
$44k – $90k per year (Estimated) • In office • Full-Time • Paris
Apply
DevTool Growth Lead 8 hours ago
$46k – $92k per year • Remote/Hybrid • Full-Time • Paris
DevOps
GitHub
Management
n8n
Discord
Marketing
Reddit
Apply
$69k – $103k per year • Remote/Hybrid • Full-Time • 3+ years exp • Paris
Python
Go
Java
Rust
TypeScript
SQL
C#
C++
Databases
PostgreSQL
DevOps
Docker
Apply
See all jobs
This is one of many
695,350 more open roles from verified company boards, updated every day.