703,247open jobs
41,500companies
99,451added this week
Browse all
Salary
$86k – $109k per year
Location
Remote/Hybrid (Auckland, New Zealand)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Our ambition is to help growth companies to use brand marketing to drive short and long term success. Tracksuit is an online dashboard that enables any company to track the strength of their brand, compare against their competitor's brands, and get valuable advice on what the numbers mean and what to do next. All for a tenth of the price of traditional brand tracking.

Tracksuit exists to help marketers prove their brand building is working. We give teams the data they need to make smarter decisions, convince stakeholders, defend budgets, and track their progress. The brand tracking industry is dominated by 100-page reports, static data and big price-tags. We're doing things differently by being built for the modern marketer: always-on, accessible, and approachable.

We're now tracking more than 1,000 brands across 25 countries globally. With offices in Auckland, Sydney, London and New York City, we're scaling fast with a brilliant team of collaborative and ambitious humans. Our culture is defined by "high care, high performance". We strive to be the best and look after each other while we do it.

Are you our next Senior Site Reliability Engineer?

We’re on the lookout for a Senior Site Reliability Engineer to join Tracksuit, based in our Auckland office.

You’ll set how reliability, observability and security work at Tracksuit, from the AWS infrastructure underneath the platform through to the agents, MCP servers and model-backed features running on top. Teams build and run their own services, and the SRE team makes it straightforward for them to do that well, through the golden paths, guardrails, tooling and defaults that make the reliable way the easy way, plus the support to lean on when something does break.

Why this role is special

  • Platform-wide impact: You set how reliability, observability and operational readiness work across the whole platform, and every team shipping on it feels the difference.
  • Agentic systems in production: You’ll set the guardrails, identity and blast-radius controls for agents operating against real systems and data, and work out what could go wrong before it does.
  • Golden paths: You’ll make the platform legible to agents as well as people, through golden paths, MCP servers, skills and documentation that both can actually use.
  • Startup pace with scaleup reach: Five years in, we’re working with over 1,000 brands across AU, NZ, USA and the UK, with plenty of scaling still ahead.

What you'll do

  • Build and maintain the cloud infrastructure and paved paths teams ship on, so resilience, security and scalability come as defaults rather than as decisions each team makes on its own
  • Give teams observability they can self-serve: monitoring, alerting, logging and tracing, plus automation for provisioning, deployments and the operational work nobody should be doing by hand
  • Make incidents easier to handle: the tooling, runbooks and practice that let whoever is closest to the problem debug it quickly, and post-incident reviews that turn into actual changes
  • Set the identity and least-privilege model for non-human actors, including agents, CI and MCP servers, so teams can give agents real access without real risk
  • Document infrastructure designs and operational procedures in a form both people and agents can act on, including machine-readable runbooks
  • Set the standard for what is safe to ship, and make it easy to meet through guardrails and checks in the pipeline rather than through gatekeeping
  • Give teams visibility and controls over cloud and inference spend, so the cost of what they run is something they can see and act on
  • Coach engineers in reliability practices, and work with Engineering and Product to balance feature delivery against platform needs

That's the role, so who are you?

  • You’ve run incident command on real production incidents, and you’ve owned a platform through a meaningful scaling step
  • Deep infrastructure skills: Infrastructure as Code (Terraform, Terragrunt, CDK), containers (ECS), cloud platforms (AWS preferred, or GCP/Azure), CI/CD, and scripting in Python, TypeScript or Bash.
  • Strong on observability: Datadog or similar, distributed tracing and structured logging, and a solid grasp of incident management methodology.
  • Security-minded: Networking and cloud architecture fundamentals, plus an understanding of the security model for agentic systems, including credential handling, least privilege and prompt injection.
  • Works with agents: You use coding agents in your own work (Claude Code or equivalent) and have a point of view on where they help and where they don’t.
  • Product-oriented: You navigate technical complexity with business outcomes in mind, and can clearly communicate system health, risks and trade-offs to technical and non-technical people.
  • Kind: This is a collaborative role, so we ultimately want a great, supportive team player.
  • Bonus points if you’re passionate about marketing, design, and building exceptional user experiences.
  • Some of the tools we use: AWS, ECS, Terraform and Terragrunt, GitHub Actions, Datadog, Claude Code, Linear, Notion, Postgres, DynamoDB and Snowflake.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
703,247 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Auckland
$54k – $144k per year (Estimated) • Remote • Full-Time • Bachelor's Degree
Python
Databases
PostgreSQL
pgvector
Pinecone
FAISS
Qdrant
AI/ML
Copilot
Cursor
LangGraph
LangChain
Claude Code
LoRA
Model Context Protocol
Fine-tuning
Multimodal AI
Function Calling
Langfuse
LangSmith
PEFT
QLoRA
Ragas
Transformers
CrewAI
LLM
RAG
Reranking
Hybrid Search
OpenAI
Anthropic
OpenAI Codex
Human-in-the-Loop
Structured Outputs
LLM Evaluation
LLM Guardrails
Agentic Workflows
Tool Use
DevOps
Terraform
GitHub Actions
Azure
CI/CD
Git
AWS
Docker
Bicep
Cybersecurity
SOC 2
HIPAA
Apply
$65k – $188k per year (Estimated) • Remote • Full-Time • Montevideo
Python
Bash
Databases
PostgreSQL
DynamoDB
AI/ML
Machine Learning
DevOps
Terraform
GCP
Red Hat
Helm
GitHub Actions
Istio
Datadog
Debian
Kustomize
Azure
CI/CD
AWS
Docker
Kubernetes
Ubuntu
Amazon EKS
Google GKE
Azure AKS
GitLab
Amazon CloudWatch
Linux
Apply
$209k – $239k per year • In office • Full-Time • 9+ years exp • Bachelor's Degree • Riverwoods
Python
JavaScript
Rust
TypeScript
C#
Node JS
Scala
DevOps
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Bitbucket
GitHub
Management
Agile
Apply
$197k – $225k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • McLean
Python
JavaScript
Rust
TypeScript
C#
Node JS
Scala
Frontend
Vue.js
DevOps
GCP
Azure
CI/CD
Git
AWS
Docker
Kubernetes
Bitbucket
GitHub
Management
Agile
Apply
IT Audit Manager 4 hours ago
$98k – $183k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Houston • Atlanta
Python
SQL
Databases
Snowflake
DevOps
GCP
Azure
AWS
Windows
Cybersecurity
OWASP
Analytics
Tableau
Power BI
Alteryx
Management
Agile
Microsoft Office
Apply
$68k – $89k per year • Remote/Hybrid • Full-Time • 5+ years exp • Sydney
AI/ML
Claude
Management
Notion
Marketing
HubSpot
Apply
Head of Marketing US 13 days ago
$195k – $265k per year • Remote/Hybrid • Full-Time • 15+ years exp • New York
Apply
$83k – $107k per year • Remote/Hybrid • Sydney
Apply
$170k – $185k per year • Remote/Hybrid • Full-Time • New York
Apply
$111k – $136k per year • Remote/Hybrid • Sydney
Apply
Remote/Hybrid • Full-Time • Auckland
Analytics
Microsoft Excel
Management
Agile
Apply
$47k – $115k per year (Estimated) • In office • Bachelor's Degree • Auckland
Apply
In office • 3+ years exp • Bachelor's Degree • Auckland
Python
C++
C++
CMake
DevOps
Git
Bazel
Linux
Management
Jira
Apply
In office • Full-Time • Auckland
Apply
In office • Full-Time • Auckland
Apply
See all jobs
This is one of many
703,247 more open roles from verified company boards, updated every day.