I

Cloud Site Reliability Engineer

Intrinzic

Sydney, New South Wales, Australia · Full Time

Be the first to apply

Experience
Any
Salary
AUD 130,000 – AUD 130,000 / year
Openings
1
Posted
1 week ago
Work mode
In office
Eligibility
Candidates must be Australian permanent residents or citizens to apply for this opportunity.
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

Overview

Join a leading Australian FinTech which operates critical real-time payments infrastructure underpinning the country's card payments, bill payments, and instant money transfers. This permanent role offers a $130,000 base salary plus superannuation and demands excellence in reliability engineering to maintain seamless operations that are essential to the nation.

Key Responsibilities

  • Manage and enhance cloud and platform infrastructure focusing on availability, resilience, security, and optimal performance.
  • Respond as the first line of defense during platform incidents, tackling complex reliability and capacity challenges, aiming to reduce mean time to detect (MTTD) and mean time to recover (MTTR).
  • Engineer and optimize AWS and Kubernetes environments, including EKS, EC2, EFS, VPC, and IAM components.
  • Develop and refine CI/CD pipelines and automate recurring operational tasks to minimize manual interventions.
  • Improve observability systems to proactively identify issues before affecting customers.
  • Implement robust resilience strategies, including disaster recovery and controlled change and configuration management.

Required Qualifications & Skills

  • Practical experience as a Site Reliability Engineer, cloud engineer, or platform engineer in complex, high-availability systems.
  • Deep expertise in AWS, particularly with EKS, EC2, EFS, VPC, IAM, and strong underlying cloud networking knowledge.
  • Extensive hands-on experience with Kubernetes cluster management, scaling, upgrades, and performance tuning.
  • Proficiency in infrastructure as code tools such as Terraform, CloudFormation, CDK, or Ansible, coupled with scripting skills in Python, TypeScript, or similar languages.
  • Proven ability to design, build, and maintain CI/CD pipelines using GitHub, GitLab, or Bitbucket, with a solid understanding of observability best practices.
  • Candidates must be located in Sydney and able to work onsite a few days each week.

Preferred Experience

  • Background in cloud security, incident response, and disaster recovery within regulated or mission-critical environments.
  • Industry experience in payments, banking, or financial services.
  • Demonstrated success in reducing noisy, alert-heavy platforms to more manageable and quiet operations.

Additional Information

  • The role is based onsite in Sydney, requiring presence in the office several days per week.
  • Australian permanent residents or citizens only are eligible to apply.

Tools & software

Amazon Web Services AWS required Amazon Elastic Compute Cloud EC2 required Terraform · 2 to 5 years required

How they work

Teamwork & Collaboration Problem Solving Attention to Detail Work Ethic

Leave it if you'd like a reply — we won't use it for anything else.

Click to browse, drag & drop, or paste a screenshot

PNG, JPG, GIF, MP4, WebM, MOV · Max 20MB each · Up to 5 files

🤖
Online · instant AI help
Broxer