Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
Office, Lilongwe, Central Region, MWI
In-Office
Senior level
Blockchain • Fintech • Payments • Software • Financial Services • Cryptocurrency
The Role
Operate and improve reliability of market-critical systems by automating toil, building CI/CD and IaC, improving observability, leading incident reviews, performing capacity and resilience testing, and providing 24/7 on-call support for production services.
Summary Generated by Built In
ASX: Powering Australia's financial marketsWhy join the ASX?

When you join ASX, you’re joining a company with a strong purpose – to power a stronger economic future by enabling a fair and dynamic marketplace for all.

In your new role, you’ll be part of a leading global securities exchange with a strong brand. We are known for being a trusted market operator and an exciting data hub. 

Want to know why we are a great place to work, click on the link to learn more.

www.asx.com.au/about/careers/a-great-place-to-work

We are more than a securities exchange!

The ASX team brings together talented people from a diverse range of disciplines. 

We run critical market infrastructure, with 1 in 3 people employed within technology.  Yet we have a unique complexity of roles across a range of disciplines such as operations, program delivery, financial products, investor engagement, risk and compliance.

We’re proud to foster a workplace where diversity is celebrated and inclusion is part of our everyday culture. Our employee-led networks champion LGBTIQ+ inclusion, promote gender equality, accessibility and wellbeing, inspire giving and volunteering, and celebrate cultural and religious events, creating a sense of belonging for all. As an AWEI Bronze employer and member of the Champions of Change Coalition for gender equality, we’re committed to a fair and inclusive workplace where everyone can thrive.

Your Team

Provides first and second level technical support on all systems within the Markets Line of Business, this includes trading, Reference Data Management, Market Surveillance, Market Announcements, Regulatory Feeds, Clearing, Risk and Margining systems.

Your Responsibilities

  • The role requires a strong SRE mindset including; reducing operational toil through automation; improving observability across logs, metrics and traces; supporting production readiness and non-functional testing; and driving resilience, capacity, incident learning and continuous reliability improvement across market-critical services.

  • Bring a reliability-first mindset to market-critical systems where availability, integrity, recoverability and controlled change are essential to maintaining confidence in ASX services.

  • Design and Implement: CI/CD pipeline, Infrastructure as Code (IaC) development and maintenance and automate operational tasks, deployments, and service recovery processes.

  • Contribute to production readiness assessments by reviewing observability, supportability, resilience, capacity, recovery procedures and operational documentation before release.

  • Support capacity planning and performance engineering by analysing utilisation trends, saturation signals, workload patterns and scalability risks across critical services.

  • Strengthen service resilience through failover validation, recovery testing, fault-tolerance reviews and participation in disaster recovery or operational resilience exercises.

  • Lead or contribute to post-incident reviews, identifying systemic causes, corrective actions and reliability improvements that reduce recurrence and improve recovery outcomes.

  • Design observability practices across logs, metrics and traces, ensuring alerts are actionable, dashboards reflect service health, and monitoring aligns to SLOs and operational risk.

  • The role will also include providing 24 x 7 on-call support (rostered) and weekend/after-hour installations and upgrades.

  • The scope of the role also includes incident, problem, release and change management, and undertaking BAU work within the team.

Must Have

  • 5+ years of experience in a similar SRE role with previous markets experience is highly regarded.

  • Strong understanding of incident response, post-incident review and problem management.

  • Experience with production readiness, release readiness and operational acceptance.

  • Experience with capacity, performance and resilience testing.

  • Strong knowledge of observability design, not only monitoring tools.

  • Experience supporting high-availability, distributed, business-critical platforms.

  • Strong scripting and automation skills using Python, PowerShell, shell scripting and scheduled automation tooling.

  • Demonstrated experience supporting critical applications using AWS Cloud Architecture covering EC2, S3, Lambda, RDS with strong understanding of microservices, containerisation (Docker), Kubernetes administration, and CI/CD pipelines.

  • Experience setting up CloudWatch alerts, supporting observability tools such as Grafana, Prometheus, OpenTelemetry.

  • Strong database operational experience across Oracle and/or Microsoft SQL Server, including SQL scripting, performance tuning, indexing, query plans and production troubleshooting.

  • Strong Linux/Unix administration and troubleshooting experience across cloud, container and distributed application environments.

  • Excellent knowledge and technical skills on MS Windows Servers (2019-2022).

  • Excellent troubleshooting, problem solving and root cause analysis skills.

  • Ability to communicate effectively with both technical and non-technical stakeholders.

  • Willingness to learn new technologies, self-motivation with the ability to multi-task and prioritise work with minimal supervision.

  • AWS certification at Associate level or above.

  • Technical experience in distributed transactions, high availability, performance-critical systems.

Nice to Have

  • Exposure to the standards, practices and procedures of the financial services industry.

  • Securities industry experience such as understanding of exchange-traded futures and options, derivatives markets, and the business processes related to clearing of member trades.

  • Strong networking troubleshooting across DNS, TLS, load balancers, firewalls and TCP/IP.

We make hiring decisions based on your skills, capabilities and experience, and how you’ll help us to live our values. We encourage you to apply even if you don’t meet all the criteria of this role.

If you need any adjustments during the application or interview process to help you present your best self, please let us know at [email protected].

At ASX Group, our diverse workforce is essential to build and maintain a fair and dynamic marketplace. We support flexible working and offer hybrid working options. Even if our roles are advertised as full-time, we encourage you to apply if you are interested in part-time or other flexible working arrangements.

We will arrange for successful candidates to have background checks, including reference and police checks, completed as part of the on-boarding process.

To be considered for this position, candidates must be legally authorised to work in Australia on a permanent basis without any restrictions.

Skills Required

  • 5+ years of experience in a similar SRE role (markets experience highly regarded)
  • Strong understanding of incident response, post-incident review and problem management
  • Experience with production readiness, release readiness and operational acceptance
  • Experience with capacity, performance and resilience testing
  • Strong knowledge of observability design across logs, metrics and traces
  • Experience supporting high-availability, distributed, business-critical platforms
  • Strong scripting and automation skills using Python, PowerShell and shell scripting
  • Demonstrated experience in AWS Cloud Architecture covering EC2, S3, Lambda, RDS
  • Strong understanding of microservices, containerisation (Docker) and Kubernetes administration
  • Experience with CI/CD pipelines and automating deployments/service recovery
  • Experience setting up CloudWatch alerts and supporting Grafana, Prometheus, OpenTelemetry
  • Strong database operational experience across Oracle and/or Microsoft SQL Server, including SQL scripting and performance tuning
  • Strong Linux/Unix administration and troubleshooting experience across cloud, container and distributed environments
  • Excellent knowledge and technical skills on MS Windows Servers (2019-2022)
  • AWS certification at Associate level or above
  • Excellent troubleshooting, problem solving and root cause analysis skills
  • Ability to communicate effectively with technical and non-technical stakeholders
  • Willingness to learn new technologies and ability to multi-task with minimal supervision
  • Technical experience in distributed transactions, high availability, performance-critical systems
  • Exposure to financial services standards, practices and procedures
  • Securities industry experience (exchange-traded futures, options, derivatives, clearing processes)
  • Strong networking troubleshooting across DNS, TLS, load balancers, firewalls and TCP/IP
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Sydney, NSW
1,848 Employees
Year Founded: 1987

What We Do

ASX is one of the world’s top ten exchanges. As a full-service exchange, we offer trading, clearing, settlement, market insights, connectivity, and depository services across all major asset classes including equities, derivatives, ETFs, options, and managed funds. With a total market capitalisation of around $1.5 trillion, ASX is home to some of the world’s leading resource, finance, and technology companies. Our $47 trillion interest rate derivatives market is the largest in Asia and among the biggest in the world. ASX’s network and data centre (The Australian Liquidity Centre) provides a world class financial infrastructure and access to Australia’s largest pools of liquidity.

Similar Jobs

F5 Logo F5

Site Reliability Engineer

Cloud • Information Technology • Security • Software
In-Office
Office, Lilongwe, Central Region, MWI
5847 Employees

Signal AI Logo Signal AI

Site Reliability Engineer

Artificial Intelligence
Hybrid
Office, Lilongwe, Central Region, MWI
226 Employees
50K-50K Annually
Hybrid
Office, Lilongwe, Central Region, MWI
35 Employees

F5 Logo F5

Senior Site Reliability Engineer

Cloud • Information Technology • Security • Software
In-Office
Office, Lilongwe, Central Region, MWI
5847 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account