368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$118k – $255k per year (Estimated)
Location
In office (Singapore)
Seniority
Architect
Employment
Full-Time
Overview
Company
Impact
Profile match
Firmus Technologies builds immersion-cooled artificial intelligence factories that run large GPU fleets on renewable power. Founded in 2021 in Singapore, it develops both the data centre design and the cloud service on top. Its Project Southgate campuses in Australia are among the region's largest planned artificial intelligence sites.

Firmus Technology

Firmus Technologies is a global leader pioneering the development and operation of efficient AI infrastructure across Asia Pacific. 

Founded in Australia in 2019, our mission is to create the most efficient AI infrastructure by combining cutting-edge technology with a steadfast commitment to sustainability.

At Firmus, we are unique in our approach. We design, build, and operate a new class of digital infrastructure - the AI Factory. Through our model-to-grid technology approach, we have pushed the boundaries of multi-generational liquid cooling systems, energy management, AI software orchestration, and construction. For our customers, this approach allows us to make every watt count and deliver low-cost AI tokens globally.

Firmus AI Cloud

Our large-scale GPU cloud platform, Firmus AI Cloud, is purpose-built to deliver energy-efficient AI compute at scale to customers.

It empowers developers, enterprises, educational institutions, and government users to train and deploy AI models with unmatched efficiency and cost savings. With an ever-growing suite of services and applications, we are committed to delivering a cloud experience that is market-leading, proprietary, and built to scale.

Role Summary

We are seeking a Principal Solution Architect, AI Data Infrastructure  to serve as Firmus’s subject-matter expert on everywhere data lives and moves across the AI fabric-storage, memory, and the caching layers that keep GPUs fully utilised. You will own the architecture, evaluation, and optimisation of the systems that feed our AI Factory and Firmus AI Cloud, from high-performance parallel and software-defined storage, through NVMe and next-generation memory fabrics, to the KV-cache and data-movement paths that determine real training and inference performance. In a business built on making every watt count, your job is to make every byte count with it.

You will operate as the trusted technical authority across both internal engineering teams and external customers - translating complex data-path decisions into measurable performance, capacity, and TCO outcomes. Whether validating a customer’s workload requirements or shaping our own platform roadmap, you’ll be the person who understands the principles deeply enough that specific vendors and brands are secondary to the architecture.

Key Responsibilities

You’ll own the reference architecture for the AI data path end to end-parallel and software-defined storage, NVMe and NVMe-oF, object and file systems, memory disaggregation and pooling, and the KV-cache and data-movement strategies that eliminate bottlenecks between storage, memory, and GPUs. You’ll evaluate, benchmark, and select technologies against real AI workloads, turning results into clear decisions on performance, capacity, power efficiency, and total cost of ownership.

Depending on the engagement, you’ll work externally as the trusted advisor to customers, sizing solutions, validating requirements, and resolving performance issues in production - or internally, shaping the platform roadmap and integrating storage and memory into our Kubernetes-based and bare-metal environments alongside compute and networking teams. Either way, you’ll bring the low-latency fabric expertise (RDMA, RoCEv2, InfiniBand) to connect it all, and establish the operational best practices for data protection, resilience, and lifecycle management that a large-scale cloud demands.

Skills & Experience

You’ll bring deep, vendor-agnostic expertise across the AI data path, with hands-on experience of one or more leading high-performance storage and memory platforms-systems in the class of WEKA, VAST Data, DDN, or Dell (PowerScale/PowerFlex)-understanding the underlying principles well enough that the specific brand is secondary. You’re fluent in NVMe and NVMe-oF, parallel and software-defined storage, object and file systems, and the low-latency network fabrics (RDMA, RoCEv2, InfiniBand) that link data to GPUs, and you track next-generation memory directions such as CXL, memory disaggregation, and KV-cache optimisation for LLM workloads.

Just as important is benchmarking discipline - you can design and run workload-driven evaluations across bare-metal, virtualised, and containerised environments and translate the numbers into architecture and business cases. You’ve done this at scale, integrating storage and memory into GPU clusters (DGX/HGX and similar) and reasoning about the CAPEX, server-count, and power trade-offs that come with it. Typically this comes with 10+ years in storage, memory, HPC, or AI infrastructure roles.

Communication rounds it out. You can advise a C-level customer, author a clear reference architecture or whitepaper, and mentor engineers with equal ease, and you’re comfortable in high-ambiguity environments where the right architecture has to be created rather than copied.

Key Competencies

You bring deep command of the full AI data path - storage, memory, and cache - and back it with rigorous, workload-driven benchmarking. You think in systems, reasoning fluently across compute, network, and data, and you stay pragmatic and vendor-agnostic, optimising for outcomes rather than badges. You operate as a single-threaded owner, driving to the outcome, and you communicate with equal ease across engineering teams and executive audiences.

Location & Reporting

  • Singapore
  • Reporting to the Director of Solution Architecture & Delivery

Employment Basis

Full-time

Diversity

At Firmus, we are committed to building a diverse and inclusive workplace. We encourage applications from candidates of all backgrounds who are passionate about creating a more sustainable future through innovative engineering solutions.

Join us in our mission to revolutionize the AI industry through sustainable practices and cutting-edge engineering. Apply now to be part of shaping the future of sustainable AI infrastructure.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Singapore
$217k – $304k per year • Equity • Remote • Full-Time • 8+ years exp
Go
Databases
Apache Kafka
ClickHouse
Google BigQuery
AI/ML
Flink
Recommender Systems
DevOps
Incident Management
Kubernetes
Apply
$152k – $239k per year • Remote • Full-Time • 8+ years exp
SQL
Databases
Snowflake
AI/ML
LLM
Model Context Protocol
Analytics
A/B Testing
Apply
$105k – $252k per year • Remote • Full-Time • 18+ years exp • Bachelor's Degree
Python
Java
Java
Gradle
DevOps
Ansible
AWS
CI/CD
CloudFormation
Configuration Management
Docker
GitHub Actions
GitLab CI
Helm
Jenkins
Kubernetes
Platform Engineering
Terraform
GitHub
GitLab
Cybersecurity
Sonatype Nexus IQ
Management
Confluence
Jira
Apply
$69k – $171k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Bachelor's Degree • Bogotá
SQL
Databases
Databricks
AI/ML
dbt
LLM
NLP
Context Engineering
Analytics
Power BI
Tableau
Apply
$54k – $175k per year (Estimated) • Remote • Full-Time
JavaScript
Node JS
TypeScript
Frontend
React.js
Sass
DevOps
AWS
CI/CD
Docker
Kubernetes
Rest API
Apply
$60k – $148k per year (Estimated) • In office • Full-Time • Launceston
DevOps
HPC
IoT
OPC UA
Apply
$83k – $193k per year (Estimated) • In office • Full-Time • 12+ years exp • Bachelor's Degree • Sydney
Apply
$149k – $269k per year (Estimated) • In office • Full-Time • 5+ years exp • San Francisco
AI/ML
LLM
Management
Jira
Apply
$181k – $331k per year (Estimated) • In office • Full-Time • 8+ years exp • Bachelor's Degree • San Francisco
C++
Python
C++
PyTorch C++
TensorFlow C++
AI/ML
PyTorch
TensorFlow
InfiniBand
Apply
$166k – $356k per year (Estimated) • In office • Full-Time • 5+ years exp • San Francisco
AI/ML
Fine-tuning
Apply
$117k – $251k per year (Estimated) • Remote/Hybrid • Full-Time • Singapore
Apply
$74k – $126k per year (Estimated) • In office • Full-Time • Singapore
Python
Apply
$88k – $191k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Singapore
C++
Java
Kotlin
Python
Mobile
JUnit
DevOps
Git
gRPC
Jenkins
JFrog Artifactory
Shift-Left
Cybersecurity
Shift-Left Security
QA
Pytest
Robot Framework
TestNG
Apply
Senior AI Architect 4 hours ago
$138k – $304k per year (Estimated) • In office • Full-Time • 6+ years exp • Bachelor's Degree • Singapore
Python
SQL
Databases
Databricks
AI/ML
AI Agents
LangGraph
OpenAI
RAG
Spark
LangChain
DevOps
Azure
Apply
$64k – $189k per year (Estimated) • Remote/Hybrid • Full-Time • 1+ year exp • Bachelor's Degree • Singapore
Python
SQL
Databases
Apache Kafka
AI/ML
Amazon SageMaker
Kubeflow
MLFlow
Spark
Vertex AI
DevOps
AWS
Azure
Azure DevOps
CI/CD
Docker
GCP
GitLab
GitLab CI
Jenkins
Kubernetes
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.