Life @ AP+:
We are one connected team in pursuit of one inspiring purpose – to unite people and technology to power better experiences. Each of us has a part to play in making that happen. You’ll be encouraged to bring your big ideas forward and make a difference through your work. Taking steps forward in your career whilst still having room for fun, friendships, and flexibility in your daily life.
We’re driven by our core values: lead with heart, learn for tomorrow and live our legacy. A purpose like ours takes the inspired impact of an incredible team. Ready to change the game? We’re ready to help you do it.
The Purpose:
Reliability matters when you’re supporting technology that Australians depend on every day.
As a Cloud Site Reliability Engineer, you’ll help keep AP+’s cloud and platform services reliable, available, secure and resilient. This is a hands-on engineering role where you’ll combine cloud infrastructure, automation, observability and SRE practices to improve the performance and reliability of critical technology services.
Working primarily across AWS and Kubernetes environments, you’ll engineer for resilience, respond to complex incidents, reduce operational toil through automation and continually improve how our platforms perform at scale.
Key Responsibilities the Role Owns:
- Operate, monitor and continually improve cloud and platform infrastructure, engineering for high availability, resilience, performance and security.
- Act as a first responder to cloud and platform incidents, troubleshooting complex reliability, capacity and performance issues while improving MTTD and MTTR through lessons learned.
- Engineer and optimise AWS and Kubernetes environments, including EKS, EC2, EFS, VPC and IAM, supporting scalable and resilient workloads.
- Build and improve CI/CD pipelines and automate operational tasks to reduce manual effort, minimise risk and enable safe, frequent and reliable delivery.
- Strengthen observability across infrastructure, applications and cloud services, using meaningful metrics to identify issues and drive improvements in availability and performance.
- Build resilience into our platforms through disaster recovery, controlled change, configuration management and close collaboration with Engineering, Security and Service Management teams.
You’ll likely be a strong fit for this role if:
- You bring hands-on experience in Site Reliability, Cloud or Platform Engineering, supporting complex and highly available cloud environments.
- You have strong AWS expertise, ideally across EKS, EC2, EFS, VPC and IAM, underpinned by a solid understanding of cloud networking.
- You’re experienced with Kubernetes and containerised workloads, including cluster management, scaling, upgrades and optimisation.
- Infrastructure-as-code and automation are central to how you work, with experience using tools such as Terraform, CloudFormation, CDK or Ansible, along with Python, TypeScript or similar scripting languages.
- You’ve designed and operated CI/CD pipelines using platforms such as GitHub, GitLab or Bitbucket and understand the importance of strong observability and operational monitoring.
- You bring an engineering mindset focused on reliability and resilience, with experience across cloud security, incident response, disaster recovery and controlled change, ideally within a regulated or high-availability environment.
What happens next:
At AP+, we believe in the power of passion, pride and purpose. Our team is driven by a shared mission to make a difference in the world of payments, and we're proud to work together towards this common goal.
If you’re an SRE or Cloud Engineer who gets excited about automation, observability and engineering platforms that need to be there when it matters, we’d love to hear from you.
If you’re ready to be a game changer, please submit your application. The Talent Acquisition team will endeavour to review your application and notify you of the outcome within the next two weeks.
We want to remove all barriers to inclusion, so if you need advice or support with your application, we’re here to help. Please reach out to [email protected]. We also encourage you to let us know your pronouns at any point during the recruitment process.
AP+ are not partnering with Recruitment agencies for this role
Skills Required
- Hands-on experience in Site Reliability, Cloud, or Platform Engineering supporting complex, highly available cloud environments
- Strong AWS expertise and understanding of cloud networking
- Experience with Kubernetes and containerized workloads, including cluster management, scaling, upgrades, and optimization
- Experience with infrastructure as code and automation using Terraform, CloudFormation, CDK, Ansible, or similar tools
- Experience with Python, TypeScript, or similar scripting languages
- Experience designing and operating CI/CD pipelines using GitHub, GitLab, Bitbucket, or similar platforms
- Understanding of observability and operational monitoring
- Experience with cloud security, incident response, disaster recovery, and controlled change
- Experience with AWS EKS, EC2, EFS, VPC, and IAM
- Experience in regulated or high-availability environments
What We Do
Introducing Australian Payments Plus Australian Payments Plus (AP+) brings together Australia’s three domestic payment providers, BPAY Group, eftpos and NPP Australia, into one integrated entity. Bringing these businesses together enables AP+ to create a more competitive and coordinated Australian payments organisation that is strategically placed to respond to the impacts of regulatory and technological change today, and into the future. Operating in the public interest, AP+ is a member-owned organisation, with a diverse range of members including Australia’s domestic banks, international banks operating in Australia, some of the country’s largest merchants, payment service providers and payment processors, together with a range of challenger and disruptor brands focused on specific markets and products.







