Site Reliability Engineer/Lead

Posted 3 Days Ago
Be an Early Applicant
Melbourne, Victoria, AUS
In-Office
Senior level
Fintech • Payments • Financial Services
The Role
Build and lead an SRE practice and enterprise observability platform across Azure hybrid environments. Design observability, automation, IaC and CI/CD integrations, incident management, governance, and reliability engineering practices. Mentor teams, influence architecture, embed SRE across the SDLC, and adopt AI-assisted operations to improve platform robustness and compliance.
Summary Generated by Built In
Making a difference isn’t just our purpose, it’s what motivates our people every day. For us, work is not about ticking a box. It’s about knowing that it matters. We make a meaningful difference every day by making it easier for our customers to focus on what matters in their lives.

This is a great opportunity to establish and lead Site Reliability Engineering at Macmillan Shakespeare.

Here’s how you will make a difference in this role… 

You'll start as a hands-on technical leader, designing and building our enterprise observability platform across Azure, cloud and hybrid environments. As the capability grows, you'll transition into leading and building our SRE practice, setting engineering standards, mentoring a team and influencing technology strategy across the business.

If you're excited by creating platforms, automating operations, improving reliability and using modern technologies—including the adoption of AI-assisted operations—this role offers the autonomy and support to make a lasting impact.

At MMS clear expectations help our people Be the Difference.

Your key responsibilities in this role will include: 

  • Build our SRE capability from the ground up.
  • Shape the enterprise observability platform strategy and technology roadmap
  • Influence architecture and engineering practices across the organisation
  • Work with modern Azure cloud technologies, automation and AI-enabled operations
  • Ensure Platform Robustness, Compliance & Cyber Security
  • Grow into a strategic leadership role while staying close to technology
  • Partnering with Product Owners, Engineering Managers and Platform teams to embed reliability engineering practices throughout the software development lifecycle.
  • Join a collaborative team that values innovation, continuous improvement and engineering excellence

To be considered for this role you will have: 

  • Demonstrated experience designing and operating enterprise monitoring, observability and operational platforms in complex hybrid/cloud environments.
  • Experience with Infrastructure as Code (IaC), CI/CD pipelines, platform automation, and Microsoft Azure integrations, including logging, monitoring and alerting.
  • Experience implementing monitoring, logging and distributed tracing using modern observability practices.
  • Experience designing automated operational workflows using APIs and event-driven architectures.
  • Strong understanding of Site Reliability Engineering (SRE), operational automation and continuous improvement.
  • Experience leading major incident management, root cause analysis and post-incident reviews.
  • Ability to simplify complex technical concepts and influence engineering teams through technical leadership.
  • Strong stakeholder engagement skills across engineering, product, operations, cyber security and executive leadership.
  • Experience developing engineering standards, governance frameworks and operational practices.

Desirable

  • Experience establishing or leading SRE, Platform Engineering or DevOps functions.
  • Experience with enterprise IT Service Management (ITSM) platforms and automated incident management.
  • Experience integrating Jira, Microsoft Teams and collaboration platforms into operational workflows.
  • Experience with Azure AI Foundry, Azure OpenAI or AI-assisted operational tooling.
  • Experience in regulated industries with strong governance and compliance requirements.
  • Knowledge of modern cloud architecture, resilience engineering and distributed systems.

Essential

  • Bachelor's degree in Information Technology, Computer Science, Software Engineering or a related discipline, or equivalent industry experience.

Desirable But Not Essential

  • ITIL® 4 Foundation or Managing Professional certification.
  • Relevant Microsoft Cloud certifications.
  • Certified Site Reliability Engineer (SRE) qualification or equivalent industry training.
  • Agile and/or Scrum certifications.

What we can offer you:

·         Strong culture and the support of a values-based recognition program

·         Novated leasing benefits and discounts

·         12 weeks paid parental leave 

·         Comprehensive learning and development opportunities to support your career growth

·         Sonder digital wellbeing platform, providing personalised support 24/7, plus annual flu vaccinations

·         Default Income Protection Insurance reimbursed for members of the MMS Default Super Fund

·         Exempt Employee Share Plan

·         Volunteer leave & Career break

·         MMS Rewards program

 

Embracing our value of Everyone Matters we hold a collective commitment to foster an environment where all differences are valued and respected. 

·         We encourage individuals from all backgrounds including Aboriginal and Torres Strait Islander peoples, those caring for someone or living with a disability, LGBTQIA+ and culturally diverse applicants to apply.

·         We value the skills and attributes veterans can bring to our organisation.

·         We embrace hybrid working and welcome conversations about flexibility.

Please note all successful candidates will undergo background checks (including criminal history and ASIC checks) and an NDIS Workers Screening Check if appropriate.  All information provided will be treated confidentially.

We acknowledge Aboriginal and Torres Strait Islander people as the Traditional Custodians of the lands where we live, learn and work.

If you identify as a person living with disability and require adjustments to our recruitment process, please contact us at [email protected]

Skills Required

  • Designing and operating enterprise monitoring, observability and operational platforms in complex hybrid/cloud environments.
  • Experience with Infrastructure as Code (IaC), CI/CD pipelines, platform automation, and Microsoft Azure integrations.
  • Implementing monitoring, logging and distributed tracing using modern observability practices.
  • Designing automated operational workflows using APIs and event-driven architectures.
  • Strong understanding of Site Reliability Engineering (SRE), operational automation and continuous improvement.
  • Leading major incident management, root cause analysis and post-incident reviews.
  • Ability to simplify complex technical concepts and influence engineering teams through technical leadership.
  • Strong stakeholder engagement skills across engineering, product, operations, cyber security and executive leadership.
  • Developing engineering standards, governance frameworks and operational practices.
  • Bachelor's degree in Information Technology, Computer Science, Software Engineering or equivalent industry experience.
  • Establishing or leading SRE, Platform Engineering or DevOps functions.
  • Experience with enterprise IT Service Management (ITSM) platforms and automated incident management.
  • Integrating Jira, Microsoft Teams and collaboration platforms into operational workflows.
  • Experience with Azure AI Foundry, Azure OpenAI or AI-assisted operational tooling.
  • Experience in regulated industries with strong governance and compliance requirements.
  • Knowledge of modern cloud architecture, resilience engineering and distributed systems.
  • ITIL 4 Foundation, Microsoft Cloud certifications, Certified SRE or Agile/Scrum certifications.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Melbourne, Victoria
1,199 Employees
Year Founded: 1988

What We Do

With eight brands across the employee benefits, fleet management and disability support industries, MMS employs around 1300 people in Australia and New Zealand. Established in 1989, MMS blazed the trail for salary packaging in Australia, and we have grown from a small family business to the ASX-listed house of brands we are today. At MMS, we're proud of our history, our heart, and our commitment to making a difference to people's lives. ​We care because people matter. ​We collaborate because the greatest achievements are made together, and we continuously create because some of the best innovations have yet to be imagined.​ Our vision – to be a trusted partner that provides solutions in making complex matters simple – reflects our strong nationwide presence and the many long-term clients we partner with including Federal and State governments and some of the largest public and private sector, health and charitable organisations

Similar Jobs

Airwallex Logo Airwallex

Security Engineer

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
In-Office
2 Locations
2300 Employees
150K-220K Annually
In-Office
2 Locations
2653 Employees

Block Logo Block

Account Executive

Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
In-Office or Remote
Melbourne, Victoria, AUS
12000 Employees

Block Logo Block

Project Manager

Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
In-Office
Melbourne, Victoria, AUS
12000 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account