Senior Site Reliability Engineer – AUS Region

Posted 24 Days Ago
Hiring Remotely in Melbourne, Victoria, AUS
In-Office or Remote
Senior level
Security
The Role
Maintain secure, highly available, and performant production systems across cloud and Kubernetes environments. Build and improve infrastructure, CI/CD pipelines, Infrastructure as Code, observability, monitoring, automation, and incident response processes. Use AI-assisted tools for troubleshooting, anomaly detection, root-cause analysis, and reducing operational toil. Provide technical leadership, mentor engineers, and promote reliability and continuous improvement. Applicants must reside in APAC and hold citizenship in their current country of residence.
Summary Generated by Built In

Senior Site Reliability Engineer – AUS Region 

 Are you looking for more in life than just building another web app? Does upending cyber security resonate with you? We're a growth stage cyber security startup that is paving the way forward for how vulnerability management is run in large enterprise organizations. For our customers, vulnerability management has always been a game of catch up, with limited asset coverage and manual processes. Nucleus’ core goal is to build a fast and scalable platform that solves these problems and many more so that vulnerability   management isn't just possible, it's easy. We're looking for a passionate Senior Site Reliability Engineer to join our growing team of engineers AUS region. 

What You Will Do  

  • Maintain Reliable, Secure, and AI-Assisted Production Operations 
    Keep production systems highly available, secure, patched, and performant. Use AI-assisted tooling to accelerate troubleshooting, identify risks, analyze incidents, and improve operational response. 

  • Build and Maintain Kubernetes, Cloud, and DevOps Infrastructure 
    Own and improve Kubernetes clusters, containerized workloads, Infrastructure as Code, CI/CD pipelines, and cloud infrastructure. Leverage AI-assisted development and automation tools to improve delivery speed, configuration quality, and operational consistency. 

  • Build Observability and Automation That Reduces Toil 
    Improve monitoring, alerting, logging, dashboards, and automated remediation to identify issues earlier and reduce repetitive operational work. Apply AI and intelligent automation to correlate signals, surface anomalies, assist with root-cause analysis, and automate common SRE workflows. 

Expectations of Your Experience  

Required Qualifications: 

  • 8+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, Infrastructure Engineering, or related field. 

  • Strong hands-on experience with cloud platforms, including AWS, GCP, Azure, and/or OpenShift (OCP). 

  • Deep experience with Kubernetes, containers, and production container orchestration. 

  • Experience building and maintaining highly available, scalable, and secure production infrastructure. 

  • Strong experience with Infrastructure as Code, preferably Terraform, and configuration/automation tools such as Ansible. 

  • Strong scripting and automation skills using Python, Bash, or similar languages. 

  • Experience building and maintaining CI/CD pipelines using GitHub, GitLab, Bitbucket, or similar platforms. 

  • Strong experience with observability and monitoring platforms such as Prometheus, Grafana, Loki, CloudWatch, or equivalent tools. 

  • Experience with incident response, root-cause analysis, production troubleshooting, and reliability engineering practices. 

  • Experience using AI-assisted engineering tools to improve infrastructure automation, troubleshooting, documentation, code generation, or operational workflows. 

  • Ability to identify opportunities where AI and automation can reduce operational toil, improve signal detection, and accelerate incident investigation. 

  • Strong understanding of Linux, networking, security, cloud architecture, and distributed systems. 

  • Ability to provide technical leadership, mentor engineers, and help drive a culture of automation, reliability, and continuous improvement. 

  • Must be located in the AUS region and have citizenship in the country you currently reside. 

Minimal Requirements 

  • Minimum 8 years of experience in SRE, DevOps, Cloud Engineering, Infrastructure Engineering, or a related field. 

  • Strong hands-on experience with cloud providers like AWS, GCP, Azure, and OpenShift (OCP). 

  • Strong experience with Kubernetes, Infrastructure as Code (terraform, tofu, CloudFormation), and automation (Python, bash, PowerShell). 

  • Proven experience supporting highly available production systems, including observability, incident response, troubleshooting, and reliability improvements. 

Preferred Qualifications: 

  • Experience integrating LLMs or AI-enabled tools into engineering or operational workflows. 

  • Familiarity with AI-assisted log analysis, anomaly detection, incident summarization, or root-cause investigation. 

  • Experience building internal automation or tooling that combines APIs, scripting, infrastructure data, and AI models. 

  • Understanding of how to use AI safely in production engineering environments, including data security, access controls, validation, and human review. 

Additional Information 

At Nucleus we are committed to achieving excellence in our field by combining diversity, collaboration, teamwork, and pride in our work. All qualified applicants will receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or disability. 

Skills Required

  • 8+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, Infrastructure Engineering, or a related field
  • Hands-on experience with AWS, GCP, Azure, and/or OpenShift
  • Deep experience with Kubernetes, containers, and production container orchestration
  • Experience building and maintaining highly available, scalable, and secure production infrastructure
  • Experience with Infrastructure as Code, preferably Terraform, and configuration or automation tools such as Ansible
  • Strong scripting and automation skills using Python, Bash, or similar languages
  • Experience building and maintaining CI/CD pipelines using GitHub, GitLab, Bitbucket, or similar platforms
  • Experience with observability and monitoring platforms such as Prometheus, Grafana, Loki, CloudWatch, or equivalent tools
  • Experience with incident response, root-cause analysis, production troubleshooting, and reliability engineering practices
  • Experience using AI-assisted engineering tools to improve infrastructure automation, troubleshooting, documentation, code generation, or operational workflows
  • Ability to identify opportunities where AI and automation can reduce operational toil and accelerate incident investigation
  • Strong understanding of Linux, networking, security, cloud architecture, and distributed systems
  • Ability to provide technical leadership, mentor engineers, and promote automation, reliability, and continuous improvement
  • Must be located in the APAC region and hold citizenship in the country of current residence
  • Experience integrating LLMs or AI-enabled tools into engineering or operational workflows
  • Familiarity with AI-assisted log analysis, anomaly detection, incident summarization, or root-cause investigation
  • Experience building internal automation or tooling that combines APIs, scripting, infrastructure data, and AI models
  • Understanding of safe AI use in production engineering environments, including data security, access controls, validation, and human review
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Sarasota, FL
69 Employees
Year Founded: 2018

What We Do

Nucleus is a platform that automates vulnerability analysis, prioritization, and response, to help your organization make better risk decisions and mitigate vulnerabilities much faster than they can today.

Similar Jobs

Xero Logo Xero

Senior Engineer

Cloud • Fintech • Information Technology • Machine Learning • Software
Remote or Hybrid
2 Locations
4500 Employees

Block Logo Block

Staff Software Engineer

Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
In-Office or Remote
Melbourne, Victoria, AUS
12000 Employees

Cash App Logo Cash App

Staff Software Engineer

Blockchain • Fintech • Mobile • Payments • Software • Financial Services
Remote or Hybrid
Melbourne, Victoria, AUS
3500 Employees

Dragos Logo Dragos

Technical Account Manager

Security • Cybersecurity
Remote
Australia
295 Employees
176K-176K Annually

Similar Companies Hiring

Closinglock Thumbnail
Fintech • Real Estate • Security • Software • Financial Services • Cybersecurity • PropTech
Austin, TX
110 Employees
Credal.ai Thumbnail
Software • Security • Productivity • Machine Learning • Artificial Intelligence
Brooklyn, NY
Milestone Systems Thumbnail
Artificial Intelligence • Security • Software • Analytics • Big Data Analytics
Lake Oswego, OR
1500 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account