Responsibilities
- Ensure the stability, scalability, security, uptime, and performance of cloud infrastructure, applications, and Kubernetes-powered clusters.
- Build and manage repeatable infrastructure and workload deployments using Infrastructure-as-Code, Terraform, GitOps, ArgoCD, Python, and Bash.
- Monitor and optimise cloud and on-prem infrastructure costs and recommend resource allocation or architecture changes.
- Partner with developers, data engineers, and leadership on infrastructure planning, improvements, tooling, training, and software delivery.
- Drive improvements to software production readiness and resolve problems affecting critical systems.
- Participate in incident resolution and help debug applications, networks, databases, and compute systems.
- Maintain awareness of system capacity and performance and support scalable networks, platforms, and engineering tools.
Requirements
- Demonstrated ability to develop and maintain large, established software systems beyond simple scripts and utilities.
- Strong computer science fundamentals, including data structures, concurrency, architecture, APIs, testing, and design patterns.
- Knowledge of system engineering techniques such as profiling, diagnosing lock contention, and identifying network issues.
- Hands-on experience with SRE practices, containerisation, Kubernetes, Infrastructure-as-Code, GitOps, scalable infrastructure, and CI/CD systems.
- Deep familiarity with at least one major cloud platform and Linux, including server tuning, database and storage management, and Kubernetes cluster operations.
- A track record of ownership over critical systems and successful delivery of complex projects.
- Leadership experience through mentoring junior engineers or leading small teams or projects, even without formal management responsibility.
- Strong written and verbal communication, collaboration, initiative, adaptability, and willingness to work across teams.
- Preferred experience in high-growth startups, security compliance and certifications, GCP, ArgoCD, GitLab CI, Kafka, Apache Cassandra, Postgres, or Rust.
Benefits
- Catered lunches, snacks, and drinks on workdays at Auckland, Christchurch, London, and San Francisco offices.
- A $1,500 annual wellness allowance or local equivalent.
- Three months of fully paid parental leave for primary caregivers plus a flexible paid return-to-work arrangement.
- Paid 24/7 car parking or a commute allowance for commuting to a Partly office or co-working space.
- Office-first work arrangement in Christchurch, Auckland, London, and San Francisco, with flexibility to manage schedules in a high-trust environment.
- Company happy hours, monthly lunches, quarterly Season Openers, and an annual global offsite.
- Travel and accommodation covered for onboarding and quarterly Season Openers; relocation assistance is available for moves to Partly HQ.
Tech Stack
Categories
About Partly
Partly's mission is to connect the world's parts. We're a team of automotive veterans and world-class engineers, with backgrounds at Rocket Lab, Microsoft, F1, Apple, and Intel. As a result of our team's collective and unusually diverse experience, we have the perspective to understand what needs fixing and the expertise to actually fix it. We're building the first foundational AI model for the auto parts industry. Our technology delivers efficiencies across the entire collision parts supply chain - from procurement to order management. Partly is independent from any automotive body and backed by industry-leading investors including Octopus Ventures, Blackbird, Squarepeg, I2BF, Ten13, Shasta Ventures, Icehouse Ventures, Peter Beck (Rocket Lab), Randy Reddig (Square), Dylan Field (Figma), and Akshay Kothari (Notion). Partly is UK-based, expanding into the US, with our core engineering team based in Christchurch.
