368,941open jobs
9,452companies
47,951added this week
Browse all
Salary
$91k – $189k per year (Estimated)
Location
In office (Leeds)
Seniority
Senior · 5+ years exp
Employment
Contractor
Overview
Company
Impact
Profile match
CreateFuture is an AI transformation and software engineering partner that builds digital products, platforms, and AI-native solutions for enterprise clients. Headquartered in Edinburgh, Scotland, the company assists organizations across sectors like fintech, energy, and government in modernizing legacy technology and scaling intelligent workflows. Its services span business strategy, cloud infrastructure, customer experience design, and data engineering to drive digital growth.

Working at CreateFuture

CreateFuture is Version 1 company and an AI-native consulting partner where people do work that matters and are supported to do it well. We work alongside organisations such as PayPal, adidas, NatWest, FanDuel and Money Saving Expert, building digital products and services that make a difference  while always putting people first.

We’re a team of creators. We write code, shape delivery, build go-to-market strategies, develop AI solutions and create the practices that support our people. We work side by side with our clients, challenging what’s not working and helping them to build the future. Our commitment to craft, quality, and culture has helped us scale to over 600 people in just a few years.

Our UK Benefits 

  • 35 days leave (including bank holidays). 
  • Private medical insurance.
  • Enhanced parental and adoption leave. 
  • Financial coaching + 5% pension match.
  • 40 hours of paid learning and development.

 View our full list of UK benefits. 

CreateFuture is aGreat Place to Work-Certified™ company and has won Best Workplaces UK multiple years in a row.

Join us on our journey. Let’s create tomorrow, together, today.

About the role and team:

AI platforms are still mostly run like prototypes. There's usually a Kubernetes cluster somebody set up in a hurry, no SLOs, no runbooks, and a cost line nobody can explain. We're engaging a Senior Cloud Engineer to bring proper SRE discipline to a major AI platform programme with a client in a high-traffic, heavily regulated consumer sector.

You'll own how the platform runs. That means the Kubernetes infrastructure behind model serving, agent orchestration and batch inference. It means the CI/CD pipelines that ship models, agents, tools and prompt changes. It means the SLOs and on-call practice that make reliability a stated commitment rather than a hope, and the observability that makes AI-specific failure modes

visible, including drift, silent quality regression, cost blowouts and agent loops.

You'll be the operational conscience of the programme. When a new capability is about to ship without limits, monitoring or a runbook, you're the person who says so. This is a role for someone who finds that work satisfying rather than thankless, and who wants to do it on a platform where the SRE patterns are still being written.

What you'll be doing:

Technical Delivery & Implementation

  • Kubernetes for AI workloads: Design, build and operate the Kubernetes infrastructure behind model serving, agent orchestration and batch inference, defined in Terraform and

deployed through GitOps.

  • CI/CD for non-deterministic systems: Build deployment pipelines suited to AI workloads, covering model rollouts, agent and tool updates, and prompt and configuration changes, with

automated eval and regression checks before promotion.

  • Reliability engineering: Define and track SLOs and SLAs for platform services, run incident response and root cause analysis, and write post-mortems people will read.
  • On-call and runbooks: Take part in the programme's on-call rotation, and build runbooks clear enough that someone else can use them at 3am.
  • AI-specific observability: Instrument latency, token usage, cost per request, model and agent error rates, retrieval quality and drift, with dashboards and alerting that surface

problems before users do.

  • FinOps: Embed cost estimation, anomaly detection and resource optimisation into the pipeline as a default rather than a monthly review.
  • Security posture: Apply zero-trust networking, least-privilege access and secrets discipline to a platform handling sensitive data across multiple third-party model providers.

Client Delivery & Stakeholder Management

  • Operability from day one: Work with the inference control plane, evaluation and knowledge platform workstreams so new capabilities arrive with monitoring, limits and a runbook,

instead of acquiring them after the first incident.

  • Technical advisory: Guide the client's architecture and AWS service choices on this platform, and build the working relationships that let that guidance land.
  • Cost accountability: Understand the programme's budget context, build the cost-benefit case for infrastructure decisions, and be able to justify spend to a finance audience.
  • Risk and delivery: Spot emerging operational bottlenecks in the client's estate, raise them early, and plan and estimate your own stream accurately.
  • Stakeholder communication: Explain reliability and cost trade-offs to technical and non- technical audiences, and pivot your framing to whoever is in the room.
  • Fitting in fast: Work within the client's existing platform standards and tooling where they're sound, and make the case for change where they aren't.

Documentation & Handover

  • Runbooks and decision records: Leave operational documentation, architecture decision records and Terraform the client's own engineers can maintain confidently.
  • Knowledge transfer: Bring the client's engineers along in SRE practice as you go, so the reliability discipline holds after the engagement closes.
  • Exit readiness: Treat a clean handover as part of the definition of done from the first sprint, not something arranged in the final fortnight.

Skills & Experience

Core Technical Capabilities

We're looking for a mix of platform and SRE (50%), software engineering (30%), AI systems

operations (20%).

  • Experience: 5+ years in DevOps, SRE or platform engineering, including real production ownership of Kubernetes rather than consuming a cluster someone else runs.
  • Infrastructure as code: Strong Terraform, and production experience on AWS. GCP or Azure alongside it is welcome.
  • Python: Solid enough for real automation and operational tooling, with sound engineering practice around it.
  • Pipelines: Built and maintained CI/CD with GitHub Actions, GitLab CI, Argo CD, Jenkins or similar.
  • Observability: Hands-on with Datadog, Prometheus, Grafana or OpenTelemetry, and opinionated about what's worth alerting on.
  • Incident management: Calm, methodical instincts under pressure, and a habit of fixing the class of problem rather than the instance.
  • AI and ML operations: Experience operating model serving, LLM inference or MLOps pipelines in production is a strong plus. Familiarity with model APIs, vector databases, MCP

or orchestration frameworks helps.

  • Networking and security: VPC design, service mesh and zero-trust patterns.

Domain & Sector Experience

  • Regulated industries: Experience running platforms under strict compliance and data privacy controls, in iGaming, financial services or banking, is highly advantageous.
  • Scale: Comfort with high-traffic, latency-sensitive systems where peak load is a sporting fixture rather than a gradual curve.
  • Contract and consulting delivery: A track record of taking on operational ownership of an unfamiliar estate quickly, and being trusted with it.

Useful credentials

Track record matters more than certification here, but the following are useful evidence of depth:

  • Certified Kubernetes Administrator (CKA).
  • AWS Solutions Architect Professional or DevOps Engineer Professional.
  • FinOps Certified Practitioner (FOCP), which is a strong differentiator for this role.
  • Terraform Associate.

What we’ll offer you:

We trust people to do their best work. That means flexibility over rigid rules, impact over activity, and real investment in your growth both professionally and personally. You’ll be part of a supportive, and friendly culture, surrounded by smart, curious people who care deeply about what they do.

We offer flexible working, including hybrid and remote options. Our office hubs are located in Edinburgh, Leeds, Manchester, London and Bulgaria, with occasional travel to client sites or CreateFuture offices when needed.

We trust you to manage your time balancing collaboration with client time and focused work. What matters is the impact you have, not how busy you look.

Our hiring process

We try to keep our hiring process clear, fair and respectful of your time. We aim to get back to everyone who applies and we will be upfront about where you are in the process.

It usually looks like this:

  • Call with our Talent Acquisition Team 
  • Role specific capability interview 

Depending on the role, we might also ask you to do a short presentation, a practical or technical task or have a values focused conversation. We will explain what is involved before anything happens.

Inclusion at CreateFuture 

We believe diverse teams build better workplaces and better products. We want CreateFuture to be a place where people feel able to be themselves and do their best work.

If you need any adjustments or support during the application process, just. We will do what we can to help.

We look forward to your application!

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,941 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Leeds
$87k – $130k per year • In office • Full-Time • 6+ years exp • Murray
Python
TypeScript
JavaScript
Python
Alembic
FastAPI
Pydantic
SQLAlchemy
Databases
PostgreSQL
Redis
AI/ML
Embeddings
LLM
LLM Guardrails
Ollama
RAG
Frontend
React Query
React Router
React.js
Vite
DevOps
AWS
CI/CD
Docker
Docker Compose
IAM
Terraform
Cybersecurity
FedRAMP
NIST 800-53
QA
Pytest
Apply
$122k – $200k per year • In office • Full-Time • 10+ years exp • Charlotte • New York
AI/ML
AI Agents
Context Engineering
Copilot
Hallucination
Human-in-the-Loop
Knowledge Graph
LLM
LLM Guardrails
LLMOps
Prompt Engineering
RAG
DevOps
Azure
CI/CD
GitHub
Kubernetes
Vector
Analytics
A/B Testing
Apply
$145k – $193k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Chicago • Washington • Denver • Jersey City • Boston
Python
AI/ML
AI Agents
Anomaly Detection
Embeddings
Fine-tuning
Function Calling
Hugging Face
LangChain
LLM
LLM Evaluation
LLM Guardrails
PyTorch
RAG
Red Teaming
Scikit-learn
Vertex AI
DevOps
Azure
GCP
Cybersecurity
Threat Modeling
Apply
HLS specialist 5 hours ago
In office • Full-Time • Israel
DevOps
AWS
Azure
Apply
$169k – $253k per year • Equity • In office • Full-Time • 8+ years exp • Bachelor's Degree • Boulder
JavaScript
TypeScript
AI/ML
AI Agents
Frontend
GraphQL
React.js
DevOps
AWS
Azure
GCP
Google GKE
Helm
Kubernetes
Apply
In office
C#
JavaScript
Kotlin
Node JS
Python
Scala
TypeScript
C#
.NET
Python
Django
Mobile
Dependency Injection
DevOps
AWS
CI/CD
Apply
$92k – $185k per year (Estimated) • In office • Manchester
JavaScript
C#
C#
.NET
Frontend
Next.js
React.js
DevOps
AWS
CI/CD
Terraform
Apply
$92k – $190k per year (Estimated) • In office • PhD • Leeds
Python
AI/ML
AWS Bedrock
LLM
Anthropic
OpenAI
Model Context Protocol
DevOps
AWS
Kong
Platform Engineering
API Gateway
Apply
Senior AI Engineer 5 days ago
$91k – $189k per year (Estimated) • In office • Contractor • PhD • Leeds
Python
AI/ML
Braintrust
Hallucination
LangSmith
LLM
Context Engineering
Human-in-the-Loop
Arize Phoenix
DevOps
AWS
CI/CD
GCP
GitHub Actions
GitLab CI
GitHub
GitLab
Apply
In office
DevOps
Azure
CI/CD
Apply
Remote/Hybrid • Full-Time • Glasgow • Bristol • London • Manchester • Cambridge
Design
AutoCAD
SpaceTech
QGIS
Apply
$87k – $171k per year (Estimated) • In office • Contractor • Leeds
Python
SQL
Databases
Amazon Redshift
Databricks
DynamoDB
Snowflake
DevOps
AWS
AWS Lambda
Azure
GCP
Amazon Kinesis
Apply
$91k – $204k per year (Estimated) • In office • Full-Time • Leeds
C#
Java
Python
TypeScript
DevOps
AWS
CI/CD
Platform Engineering
Apply
$58k – $150k per year (Estimated) • In office • Full-Time • Leeds
C#
JavaScript
Python
TypeScript
AI/ML
Claude
Copilot
LangChain
LangGraph
Model Context Protocol
DevOps
Platform Engineering
GitHub
QA
Playwright
Apply
$92k – $190k per year (Estimated) • In office • PhD • Leeds
Python
AI/ML
AWS Bedrock
LLM
Anthropic
OpenAI
Model Context Protocol
DevOps
AWS
Kong
Platform Engineering
API Gateway
Apply
See all jobs
This is one of many
368,941 more open roles from verified company boards, updated every day.