Head of Platform
partly.com · Christchurch
Job description
About the role
Partly is building the AI infrastructure layer for the global repair industry. As Head of Platform you will own the internal platform that enables product teams to ship and operate software quickly and safely, including specialised ML/AI compute and agentic development support.
Key responsibilities
- Build and lead the Platform function covering SRE, developer experience, infrastructure, and security enablement.
- Define roadmap, hiring plan and prioritisation trade‑offs for the platform team.
- Create default developer pathways – service templates, CI/CD pipelines, environment provisioning and feature‑flag systems.
- Own reliability foundations: observability, alerting, incident response, SLOs, error‑budget practices and on‑call toil reduction.
- Maintain scalable, secure cloud and Kubernetes foundations using IaC (Terraform) and GitOps (ArgoCD).
- Develop the infrastructure for foundational ML models – GPU/TPU provisioning, training orchestration, model serving and inference cost optimisation.
- Design self‑service, machine‑consumable golden paths for AI agents and forward‑deployed engineers.
- Partner with security/compliance to embed secure‑by‑default patterns and policy‑as‑code.
- Implement FinOps visibility and guardrails to optimise cost without sacrificing reliability.
Required profile
- Proven experience leading platform, infrastructure or SRE teams in fast‑moving environments.
- Strong grounding in SRE practices (SLOs, incident management, observability, capacity planning).
- Deep experience running production workloads on GCP and Kubernetes.
- Hands‑on expertise with Infrastructure‑as‑Code and GitOps workflows.
- Developer‑experience mindset: define golden paths, write clear docs and measure impact.
- Practical experience embedding security into platforms and SDLC.
- Solid computer‑science fundamentals and distributed‑systems knowledge.
Required skills
- Google Cloud Platform (GCP)
- Kubernetes
- Terraform
- ArgoCD
- Python
- Bash
- GPU / TPU compute
- CI/CD pipelines
- GitOps
- Infrastructure as Code
- Observability (metrics, logs, traces)
- Site Reliability Engineering (SRE)
- FinOps
- MLOps (model registry, experiment tracking)
- Incident Management
- Service Level Objectives (SLO) and error budgets
What we offer
- Travel and accommodation covered for quarterly “Season Openers”.
- Relocation assistance for candidates moving to join the team.
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in New Zealand.
Salaries by job title
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
A question about this job?
Ask it here: you will get the full job summary by e-mail, right away.
Published 1 week ago
Expires 1 month from now
19 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
partly.com
Christchurch