Partly

Senior Site Reliability Engineer - ANZ

Partly

Auckland, New Zealand · Full Time

Be the first to apply

Experience
5+ yrs
Salary
Openings
1
Posted
1 week ago
Work mode
In office
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About Partly

Partly is a pioneering company developing AI infrastructure for the global repair industry, initially focusing on the $2 trillion automotive market. Their flagship AI model, Interpreter, is designed to analyze vehicle damage and identify necessary replacement parts, already trusted by thousands of businesses worldwide. Founded by former Rocket Lab engineers, Partly is headquartered in Austin, Texas, with offices in London, Christchurch, and Auckland. The company has experienced rapid growth, tripling its workforce in recent years and securing a $50 million Series B funding led by DST Global.

Partly is committed to building a workplace where employees can excel. They prioritize strong company culture and values, providing support for onboarding including travel to nearest offices and quarterly team events, plus relocation assistance if applicable.

Role Overview

The Senior Site Reliability Engineer (SRE) role involves blending software and systems engineering expertise to maintain large-scale distributed systems. The objective is to ensure high reliability, uptime, and smooth performance of both internal and customer-facing services. This position requires strategic leadership and independent initiative, focusing on infrastructure that supports global parts connectivity.

Key Responsibilities

  • Maintain and optimize the stability, scalability, and security of cloud infrastructure and third-party applications within Kubernetes-based environments.
  • Utilize Infrastructure-as-Code tools such as Terraform for Google Cloud Platform and implement GitOps workflows using ArgoCD alongside automation scripts in Python and Bash.
  • Manage and improve cost efficiency across cloud and on-premises resources while recommending architecture and resource allocation changes that enhance value without compromising reliability.
  • Collaborate across teams including developers, data engineers, and leadership to forecast infrastructure needs and provide SRE guidance, tooling, and training to ensure smooth software delivery.
  • Ensure software production readiness with an eye for improvements and proactively address issues.
  • Contribute to incident management and troubleshooting, assisting developers with debugging involving applications, networks, databases, and computing systems.

Required Skills and Experience

  • Strong software engineering skills with experience maintaining extensive, complex systems and writing durable, maintainable code.
  • Deep understanding of computer science essentials such as data structures, concurrency, architecture, APIs, testing, and design patterns.
  • Proficient in system engineering concepts, including profiling performance, identifying lock contention, and diagnosing network problems.
  • Hands-on expertise in current SRE methodologies and tools: containerization (Docker, Kubernetes), Terraform infrastructure automation, and GitOps tools like ArgoCD.
  • Comprehensive knowledge of at least one major cloud platform (preferably GCP) and Linux system administration, including server tuning, database and storage management, and Kubernetes cluster operations.
  • Demonstrated leadership and ownership of critical systems, as well as mentoring or leading engineering teams or projects.
  • Excellent communication abilities enabling collaboration with technical and non-technical stakeholders, and a team-oriented mindset.
  • Adaptable to diverse responsibilities in a fast-paced startup setting and motivated by continuous learning and challenge.

Additional Assets

  • Startup experience in fast-growth environments.
  • Knowledge of security compliance and certification maintenance.
  • Experience with tools like GCP, ArgoCD, GitLab CI, Kafka.
  • Operational familiarity with Apache Cassandra or Postgres databases in production.
  • Proficiency in Rust programming and ability to coach others in Rust best practices.

Benefits

  • Daily nutritious catered lunches at offices in Auckland, Christchurch, London, and San Francisco, including snacks and beverages.
  • An annual wellness allowance equivalent to $1,500 to support health-related expenses such as gym memberships, physiotherapy, or medical treatments.
  • Three months of fully paid parental leave for primary caregivers with flexible return-to-work options.
  • Commuting support including paid parking or allowances.
  • Modern, thoughtfully designed office spaces promoting collaboration and comfort.
  • Office-first culture complemented by flexibility and trust to balance work and personal life.
  • Regular social events including weekly happy hours, monthly lunches, quarterly team gatherings, and yearly global offsites.

Relocation Support

For candidates relocating domestically or internationally to the Partly headquarters, a generous relocation package is available to assist with moving costs.

Level

Senior

Industry

Automotive

Tools & software

How they work

Communication Teamwork & Collaboration Adaptability Leadership

Leave it if you'd like a reply — we won't use it for anything else.

Click to browse, drag & drop, or paste a screenshot

PNG, JPG, GIF, MP4, WebM, MOV · Max 20MB each · Up to 5 files

🤖
Online · instant AI help
Broxer