368,746open jobs
9,444companies
47,506added this week
Browse all
Salary
$74k – $207k per year (Estimated)
Location
In office (Melbourne)
Seniority
Middle · 3+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Base Compute is an AI inference lab. We build the runtimes and infrastructure that make powerful AI run on-device. Fast, private, and at near-zero marginal cost.

About Us

Base Compute is an AI inference lab. Our mission is to bring AGI on device. We believe in a world where everyone has access to intelligence: fast, private and always available on your device.

We’re building the infrastructure for the next generation of on-device AI, from silicon-level optimizations to distributed inference systems.

We’re working on hard problems at the intersection of inference efficiency, model intelligence and autonomous research.

The Role

We’re looking for a Founding ML Engineer to work at the frontier of on-device AI. This role is for someone who lives at the intersection of systems engineering and machine learning, turning state-of-the-art research into hyper-optimized, production-ready infrastructure.

You’ll have significant ownership over our entire inference stack and direct influence on the technical bets the company makes.

What You’ll Work On

  • Inference engine development: Building and scaling our custom inference engine, handling everything from weight loading and KV-cache management to efficient request scheduling

  • Cross-platform silicon optimization: Writing and tuning custom kernels and leveraging hardware-specific instructions to squeeze maximum performance out of diverse architectures, including Apple Silicon, NVIDIA, AMD, Snapdragon, and other edge platforms

  • Systems architecture: Developing robust, low-latency serving runtimes in C++ to manage model routing, continuous batching, and novel decoding strategies under strict thermal and memory constraints

  • Performance profiling: Identifying and eliminating bottlenecks across the entire stack, from memory bandwidth ceilings to kernel interleaving

What We’re Looking For

  • 3+ years of experience in ML engineering or systems programming (Rust, C/C++), with a strong track record of building performance-critical software

  • Expertise in GPU programming and hardware optimization across various platforms (CUDA, ROCm, Metal, Triton, or similar)

  • Solid understanding of modern LLM architectures, including parsing formats and implementing optimization techniques (quantization, speculative decoding, etc.)

  • A strong sense of ownership and autonomy: the ability to take ambiguous architectural challenges and drive them from research translation directly into production-ready infrastructure

  • Good communication: the ability to explain complex architectural decisions simply, give honest feedback and document systems cleanly

  • Nice-to-haves:

    • Familiarity with ML compilers (torch.compile, custom operators)

    • Experience with low-precision inference (INT8/FP8/FP4)

    • Knowledge of Edge LLMOps

What We Offer

  • Founding team equity and strong base salary

  • Direct influence on technical direction: your ideas will shape the roadmap

  • Work on genuinely hard problems that haven't been solved yet

  • Small team, fast iteration, low bureaucracy

Location

The team is based in Melbourne and Berlin and works in-person from the office most days. We require strong written and spoken English, since the team collaborates across time zones.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,746 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Melbourne
$42k – $107k per year (Estimated) • Remote/Hybrid • Full-Time • 8+ years exp • Mexico City
C#
Go
Java
Rust
AI/ML
AI Agents
LLM
Model Context Protocol
Cybersecurity
Delinea
Zero Trust
Apply
$72k – $168k per year (Estimated) • Equity • Remote • Full-Time
C++
Go
Java
Python
Rust
Scala
Apply
UI/UX Designer 1 day ago
$12k – $55k per year (Estimated) • Remote • Full-Time
AI/ML
Claude
LLM
LLM Guardrails
Lovable
Design
Figma
Apply
$20k – $49k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
Ruby
SQL
Databases
Amazon Neptune
Neo4j
AI/ML
Hallucination
LangChain
LangGraph
LLM
Model Context Protocol
Spark
AI Agents
LLM Guardrails
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
GitHub Actions
GitLab CI
Jenkins
Kubernetes
Rest API
Terraform
GitHub
GitLab
QA
Playwright
Postman
Selenium
Swagger
Apply
$70k – $105k per year • In office • Full-Time • 3+ years exp
Python
SQL
TypeScript
AI/ML
LLM
RAG
Function Calling
LLM Guardrails
Cybersecurity
GDPR
Management
n8n
Apply
$85k – $202k per year (Estimated) • In office • Full-Time • PhD • Berlin
AI/ML
CUDA Toolkit
Knowledge Distillation
LLM
Quantization
Reinforcement Learning
Edge AI
ROCm
Apply
Founding ML Researcher 2 months ago
In office • Full-Time • PhD • Melbourne
AI/ML
CUDA
CUDA Toolkit
Knowledge Distillation
LLM
Quantization
Reinforcement Learning
Triton
Edge AI
ROCm
Apply
$96k – $227k per year (Estimated) • In office • Full-Time • Melbourne • Sydney
PowerShell
Python
DevOps
AWS
Azure
CI/CD
GCP
Incident Management
Kubernetes
Platform Engineering
Service Mesh
Terraform
Apply
$80k – $147k per year (Estimated) • Remote/Hybrid • Full-Time • Melbourne
DevOps
CI/CD
Apply
$43k – $108k per year • Equity • Remote/Hybrid • Full-Time • Melbourne • Sydney
AI/ML
Claude
Claude Code
Design
Figma
Management
Confluence
Jira
Slack
Marketing
Mixpanel
Apply
$24k – $99k per year (Estimated) • In office • Full-Time • 2+ years exp • Melbourne
JavaScript
Python
DevOps
Azure
Azure DevOps
Management
Jira
Marketing
Salesforce
QA
Selenium
TestRail
Apply
$170k – $195k per year • Remote/Hybrid • Melbourne
Management
Confluence
Jira
Apply
See all jobs
This is one of many
368,746 more open roles from verified company boards, updated every day.