I
Cloud Site Reliability Engineer
Sydney, New South Wales, Australia · Full Time
Be the first to apply
- Experience
- Any
- Salary
- AUD 130,000 – AUD 130,000 / year
- Openings
- 1
- Posted
- 1 week ago
- Work mode
- In office
- Eligibility
- Candidates must be Australian permanent residents or citizens to apply for this opportunity.
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
Overview
Join a leading Australian FinTech which operates critical real-time payments infrastructure underpinning the country's card payments, bill payments, and instant money transfers. This permanent role offers a $130,000 base salary plus superannuation and demands excellence in reliability engineering to maintain seamless operations that are essential to the nation.
Key Responsibilities
- Manage and enhance cloud and platform infrastructure focusing on availability, resilience, security, and optimal performance.
- Respond as the first line of defense during platform incidents, tackling complex reliability and capacity challenges, aiming to reduce mean time to detect (MTTD) and mean time to recover (MTTR).
- Engineer and optimize AWS and Kubernetes environments, including EKS, EC2, EFS, VPC, and IAM components.
- Develop and refine CI/CD pipelines and automate recurring operational tasks to minimize manual interventions.
- Improve observability systems to proactively identify issues before affecting customers.
- Implement robust resilience strategies, including disaster recovery and controlled change and configuration management.
Required Qualifications & Skills
- Practical experience as a Site Reliability Engineer, cloud engineer, or platform engineer in complex, high-availability systems.
- Deep expertise in AWS, particularly with EKS, EC2, EFS, VPC, IAM, and strong underlying cloud networking knowledge.
- Extensive hands-on experience with Kubernetes cluster management, scaling, upgrades, and performance tuning.
- Proficiency in infrastructure as code tools such as Terraform, CloudFormation, CDK, or Ansible, coupled with scripting skills in Python, TypeScript, or similar languages.
- Proven ability to design, build, and maintain CI/CD pipelines using GitHub, GitLab, or Bitbucket, with a solid understanding of observability best practices.
- Candidates must be located in Sydney and able to work onsite a few days each week.
Preferred Experience
- Background in cloud security, incident response, and disaster recovery within regulated or mission-critical environments.
- Industry experience in payments, banking, or financial services.
- Demonstrated success in reducing noisy, alert-heavy platforms to more manageable and quiet operations.
Additional Information
- The role is based onsite in Sydney, requiring presence in the office several days per week.
- Australian permanent residents or citizens only are eligible to apply.
Skills
Tools & software
Amazon Web Services AWS
required
Amazon Elastic Compute Cloud EC2
required
Terraform
· 2 to 5 years required
How they work
Teamwork & Collaboration
Problem Solving
Attention to Detail
Work Ethic