368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$46k – $137k per year (Estimated)
Location
In office (London)
Seniority
Middle · 3+ years exp
Employment
Internship
Overview
Company
Impact
Profile match
Fractile is a London semiconductor company founded in 2022 that designs chips for large language model inference. Its in-memory computing architecture aims to remove the memory bottleneck that limits how fast transformer models can run. The company is backed by prominent artificial intelligence investors including a NVIDIA venture arm.

Infrastructure Engineer (Linux)

About Fractile

Fractile was founded in 2022 on the bet that, eventually, the world’s most capable AI systems would be limited in their impact by the time taken to produce useful outputs. We bet everything on the logical conclusion: that the only way to truly unlock this latent value, to make speed viable at scale, was to radically re-invent the hardware that we run our frontier AI models on. Ever since, we have been building chips and systems that tackle this problem: how to efficiently generate output at thousands of tokens per second, while handling the complexity and capacity challenges of operating large models at very long contexts.

The workloads that push to the limits of the current frontier are already transformational; it is the technical and economic limits on inference speed that are constraining progress. The defining work of the 21st century will be marked by the engine of inference delivering immense and diffuse chains of intellectual inquiry, in drug discovery, in software engineering, in materials discovery, in any field where progress is driven by deep reasoning and intelligence to resolve complex problems.

We are seeking an Infrastructure Engineer (Linux) to help build and maintain the compute, storage, and networking foundations that support our silicon development workloads, supporting an on-premise compute environment used for computationally-intensive engineering work, while growing your skills across Linux systems administration, HPC-style cluster management, networking, storage, and performance tuning. This is a great opportunity for someone with 2-3+ years of solid Linux experience who picks up new technical areas quickly, works methodically, and is comfortable operating independently, and who wants to develop deeper expertise in infrastructure, high-performance computing, and large-scale systems engineering.

Key Responsibilities:

  • Help maintain and troubleshoot a fleet of on-premise Linux (Rocky/RHEL-family) servers.
  • Support the deployment and upkeep of on-premise compute infrastructure using infrastructure-as-code tooling (e.g. Ansible).
  • Assist with monitoring and observability - setting up and maintaining tooling for resource utilisation, service health, and machine failures (e.g. Prometheus, node_exporter, Grafana, Zabbix).
  • Help diagnose and resolve networking issues (DNS, VLANs, bonding, routing) and storage issues (network filesystems, capacity management, performance troubleshooting).
  • Support day-to-day operation of a cluster compute/job scheduling environment (e.g. Slurm).
  • Help with user and identity management tasks (e.g. FreeIPA/LDAP), including onboarding and access provisioning.
  • Document infrastructure changes and contribute to runbooks and playbooks.
  • Work with engineers across the organisation to understand and help resolve infrastructure-related bottlenecks.

Requirements:

  • 2-3+ years of solid, hands-on production Linux system administration experience.
  • Good working knowledge of monitoring solutions (e.g. Prometheus, Grafana, Zabbix, or similar).
  • Solid networking fundamentals - TCP/IP, DNS, VLANs, routing, basic troubleshooting.
  • Good understanding of storage concepts - network filesystems, parallel/distributed filesystems, disk/volume management, basic performance troubleshooting.
  • Comfortable working from the command line and with scripting (Bash, Python, or similar) to automate routine tasks.
  • A methodical, curious approach to troubleshooting, with the ability to pick up new tools and technical areas quickly.
  • Comfortable working independently and taking ownership of problems through to resolution.

Desirable:

  • Exposure to or knowledge of HPC (High Performance Computing) environments or cluster compute concepts.
  • Familiarity with infrastructure-as-code tools (e.g. Ansible, Terraform).
  • Linux certifications (e.g. RHCSA, LFCS, or similar).
  • Experience working in a data centre environment.
  • Knowledge of server hardware (e.g. installation, diagnostics, component replacement).
  • Silicon industry background.

How we work

  • Ownership and execution: you will have full agency to drive your work forward
  • Rapid iteration: we all work directly with top leadership to move from idea to hardware on ambitious timelines
  • Full-stack engagement: hardware, software, silicon, and modelling teams all work closely together to create a product with generational impact
  • Optimistic and pragmatic: we possess the will to win, and to do the hard work to get us there
  • Team player mentality: the mission is bigger than any of us, and we have the curiosity and technical focus to see the best idea shipped, no matter who’s it is

What we offer:

  • Competitive salary: A competitive salary reflective of your experience and the specialist nature of the role.
  • Equity & Ownership: meaningful equity so everyone shares in the value creation.
  • Benefits: Private Medical, Dental and Vision, Contributory Pension, 25 Days holiday plus bank holidays and Life/Critical Illness Insurance.
  • Diverse & fun office: we believe the hardest problems get solved by the broadest range of minds. We are committed to Equal Employment Opportunity through attracting and retaining a diverse team and building an inclusive environment.

    Fractile is seeking to increase the clock speed of global progress, one chip at a time. We’ve recently raised $220M from investors including Founders Fund and Accel and our most important work lies ahead. Join us!

    Export controls

    Our work involves technologies subject to UK, US and other international export control regulations. Certain roles may require additional eligibility checks to ensure compliance with applicable law. We'll be transparent about this throughout the hiring process.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
London
$34k – $83k per year (Estimated) • Remote/Hybrid • Full-Time • Bengaluru
Node JS
JavaScript
Databases
Apache Kafka
DevOps
ArgoCD
Azure
Azure AKS
Azure DevOps
CI/CD
Datadog
FinOps
GitHub Actions
Grafana
Istio
Kubernetes
Platform Engineering
Prometheus
SLI/SLO/SLA
Terraform
GitHub
IAM
Cybersecurity
GDPR
Microsoft Defender
Microsoft Defender for Cloud
Okta
PCI DSS
Management
ServiceNow
Apply
$60k – $137k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Singapore
Python
DevOps
Git
Apply
NLP / LLM Engineer 1 day ago
$18k – $24k per year (net) • In office • Full-Time • 3+ years exp • Tashkent
Python
Python
FastAPI
Databases
ElasticSearch
Milvus
Pinecone
Qdrant
Weaviate
AI/ML
AI Agents
ChatGPT
DeepEval
Embeddings
Gemini
Hybrid Search
LangChain
Langfuse
LangGraph
LlamaIndex
LLM
LoRA
NLP
PEFT
Prompt Engineering
PyTorch
QLoRA
RAG
Reranking
Semantic Search
Synthetic Data
Tokenization
Triton
vLLM
Transformers
Anthropic
DPO
GraphRAG
Hugging Face
OCR
OpenAI
Semantic Search
SFT
Structured Outputs
Function Calling
TGI
DevOps
CI/CD
Docker
Git
GitHub
Analytics
A/B Testing
Apply
$20k – $50k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Moscow
Python
DevOps
CI/CD
Git
Apply
Data Engineer 3 1 day ago
$21k – $52k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Bengaluru
Java
Python
Scala
SQL
Java
Maven
Python
pySpark
Databases
Apache Kafka
Databricks
Snowflake
AI/ML
Hadoop
Spark
DevOps
Azure
Analytics
Tableau
ETL/ELT
Apply
$89k – $184k per year (Estimated) • In office • Internship • 5+ years exp
C++
Python
Rust
SystemVerilog
DevOps
Bazel
CI/CD
Datadog
Grafana
Incident Management
Prometheus
SLI/SLO/SLA
Apply
$89k – $185k per year (Estimated) • In office • Internship • 5+ years exp • London
C++
Python
Rust
SystemVerilog
DevOps
Bazel
CI/CD
Datadog
Grafana
Incident Management
Prometheus
SLI/SLO/SLA
Apply
$76k – $203k per year (Estimated) • In office • Internship • 8+ years exp • Master's Degree • London
Perl
Python
DevOps
HPC
Chips/EDA
Cadence Innovus
Formal Verification
Siemens Calibre
Synopsys Fusion Compiler
Apply
$75k – $202k per year (Estimated) • In office • Internship • 8+ years exp • Master's Degree
Perl
Python
DevOps
HPC
Chips/EDA
Cadence Innovus
Formal Verification
Siemens Calibre
Synopsys Fusion Compiler
Apply
ML Runtime Engineer 12 days ago
$103k – $215k per year (Estimated) • In office • Bachelor's Degree
Rust
AI/ML
SGLang
vLLM
Apply
$65k – $155k per year (Estimated) • In office • Internship • Bachelor's Degree • London
Go
JavaScript
Ruby
Scala
Apply
In office • Internship • 1+ year exp • Bachelor's Degree • London
Go
JavaScript
Ruby
Scala
Apply
$221k – $370k per year • In office • Full-Time • 5+ years exp • London
AI/ML
OpenAI
Apply
$72k – $136k per year (Estimated) • Equity • Remote/Hybrid • 5+ years exp • London
Python
SQL
Databases
Snowflake
AI/ML
AI Agents
DevOps
Kibana
Marketing
Salesforce
Apply
$77k – $148k per year (Estimated) • Equity • In office • Master's Degree • London
JavaScript
Python
Scala
AI/ML
AI Agents
DevOps
GitHub
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.