Senior Site Reliability Engineer – APAC Region

Posted 2 Days Ago
Be an Early Applicant
Kountríon, Trifylia, GRC
In-Office
Senior level
Security
The Role
Maintain secure, highly available, and performant production systems across cloud and Kubernetes environments. Build and improve infrastructure, CI/CD pipelines, Infrastructure as Code, observability, monitoring, automation, and incident response processes. Use AI-assisted tools for troubleshooting, anomaly detection, root-cause analysis, and reducing operational toil. Provide technical leadership, mentor engineers, and promote reliability and continuous improvement. Applicants must reside in APAC and hold citizenship in their current country of residence.
Summary Generated by Built In

Senior Site Reliability Engineer – APAC Region 

 Are you looking for more in life than just building another web app? Does upending cyber security resonate with you? We're a growth stage cyber security startup that is paving the way forward for how vulnerability management is run in large enterprise organizations. For our customers, vulnerability management has always been a game of catch up, with limited asset coverage and manual processes. Nucleus’ core goal is to build a fast and scalable platform that solves these problems and many more so that vulnerability   management isn't just possible, it's easy. We're looking for a passionate Senior Site Reliability Engineer to join our growing team of engineers APAC region. 

What You Will Do  

  • Maintain Reliable, Secure, and AI-Assisted Production Operations 
    Keep production systems highly available, secure, patched, and performant. Use AI-assisted tooling to accelerate troubleshooting, identify risks, analyze incidents, and improve operational response. 

  • Build and Maintain Kubernetes, Cloud, and DevOps Infrastructure 
    Own and improve Kubernetes clusters, containerized workloads, Infrastructure as Code, CI/CD pipelines, and cloud infrastructure. Leverage AI-assisted development and automation tools to improve delivery speed, configuration quality, and operational consistency. 

  • Build Observability and Automation That Reduces Toil 
    Improve monitoring, alerting, logging, dashboards, and automated remediation to identify issues earlier and reduce repetitive operational work. Apply AI and intelligent automation to correlate signals, surface anomalies, assist with root-cause analysis, and automate common SRE workflows. 

Expectations of Your Experience  

Required Qualifications: 

  • 8+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, Infrastructure Engineering, or related field. 

  • Strong hands-on experience with cloud platforms, including AWS, GCP, Azure, and/or OpenShift (OCP). 

  • Deep experience with Kubernetes, containers, and production container orchestration. 

  • Experience building and maintaining highly available, scalable, and secure production infrastructure. 

  • Strong experience with Infrastructure as Code, preferably Terraform, and configuration/automation tools such as Ansible. 

  • Strong scripting and automation skills using Python, Bash, or similar languages. 

  • Experience building and maintaining CI/CD pipelines using GitHub, GitLab, Bitbucket, or similar platforms. 

  • Strong experience with observability and monitoring platforms such as Prometheus, Grafana, Loki, CloudWatch, or equivalent tools. 

  • Experience with incident response, root-cause analysis, production troubleshooting, and reliability engineering practices. 

  • Experience using AI-assisted engineering tools to improve infrastructure automation, troubleshooting, documentation, code generation, or operational workflows. 

  • Ability to identify opportunities where AI and automation can reduce operational toil, improve signal detection, and accelerate incident investigation. 

  • Strong understanding of Linux, networking, security, cloud architecture, and distributed systems. 

  • Ability to provide technical leadership, mentor engineers, and help drive a culture of automation, reliability, and continuous improvement. 

  • Must be located in the APAC region and have citizenship in the country you currently reside. 

Minimal Requirements 

  • Minimum 8 years of experience in SRE, DevOps, Cloud Engineering, Infrastructure Engineering, or a related field. 

  • Strong hands-on experience with cloud providers like AWS, GCP, Azure, and OpenShift (OCP). 

  • Strong experience with Kubernetes, Infrastructure as Code (terraform, tofu, CloudFormation), and automation (Python, bash, PowerShell). 

  • Proven experience supporting highly available production systems, including observability, incident response, troubleshooting, and reliability improvements. 

Preferred Qualifications: 

  • Experience integrating LLMs or AI-enabled tools into engineering or operational workflows. 

  • Familiarity with AI-assisted log analysis, anomaly detection, incident summarization, or root-cause investigation. 

  • Experience building internal automation or tooling that combines APIs, scripting, infrastructure data, and AI models. 

  • Understanding of how to use AI safely in production engineering environments, including data security, access controls, validation, and human review. 

Why You Should Be Excited 

  • 100% company-paid health, dental, vision, life, and short-term disability insurance options  

  • Generous 401k contribution (not a match) 

  • Flexible PTO + 10 company holidays 

  • Equity in a high-growth, VC-backed startup 

  • Fantastic company culture 

  • Work on a truly unique, market- defining product 

Additional Information 

At Nucleus we are committed to achieving excellence in our field by combining diversity, collaboration, teamwork, and pride in our work. All qualified applicants will receive consideration for employment without regard to race, sex, color, religion, sexual orientation, gender identity, national origin, protected veteran status, or disability. 

Skills Required

  • 8+ years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, Infrastructure Engineering, or a related field
  • Hands-on experience with AWS, GCP, Azure, and/or OpenShift
  • Deep experience with Kubernetes, containers, and production container orchestration
  • Experience building and maintaining highly available, scalable, and secure production infrastructure
  • Experience with Infrastructure as Code, preferably Terraform, and configuration or automation tools such as Ansible
  • Strong scripting and automation skills using Python, Bash, or similar languages
  • Experience building and maintaining CI/CD pipelines using GitHub, GitLab, Bitbucket, or similar platforms
  • Experience with observability and monitoring platforms such as Prometheus, Grafana, Loki, CloudWatch, or equivalent tools
  • Experience with incident response, root-cause analysis, production troubleshooting, and reliability engineering practices
  • Experience using AI-assisted engineering tools to improve infrastructure automation, troubleshooting, documentation, code generation, or operational workflows
  • Ability to identify opportunities where AI and automation can reduce operational toil and accelerate incident investigation
  • Strong understanding of Linux, networking, security, cloud architecture, and distributed systems
  • Ability to provide technical leadership, mentor engineers, and promote automation, reliability, and continuous improvement
  • Must be located in the APAC region and hold citizenship in the country of current residence
  • Experience integrating LLMs or AI-enabled tools into engineering or operational workflows
  • Familiarity with AI-assisted log analysis, anomaly detection, incident summarization, or root-cause investigation
  • Experience building internal automation or tooling that combines APIs, scripting, infrastructure data, and AI models
  • Understanding of safe AI use in production engineering environments, including data security, access controls, validation, and human review
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Sarasota, FL
69 Employees
Year Founded: 2018

What We Do

Nucleus is a platform that automates vulnerability analysis, prioritization, and response, to help your organization make better risk decisions and mitigate vulnerabilities much faster than they can today.

Similar Jobs

Carbon Robotics Logo Carbon Robotics

Performance Quality Technician

Artificial Intelligence • Computer Vision • Hardware • Machine Learning • Robotics • Software • Agriculture
Easy Apply
Remote or Hybrid
26 Locations
350 Employees
75K-85K Annually

Pfizer Logo Pfizer

Digital Operations Agentic Lead - Senior Manager

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
29 Locations
121990 Employees

Pfizer Logo Pfizer

Product Specialist

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Remote or Hybrid
29 Locations
121990 Employees

Deepgram Logo Deepgram

Sales Development Representative

Artificial Intelligence • Machine Learning • Natural Language Processing • Software • Conversational AI
In-Office or Remote
28 Locations
150 Employees

Similar Companies Hiring

Closinglock Thumbnail
Fintech • Real Estate • Security • Software • Financial Services • Cybersecurity • PropTech
Austin, TX
110 Employees
Credal.ai Thumbnail
Software • Security • Productivity • Machine Learning • Artificial Intelligence
Brooklyn, NY
Milestone Systems Thumbnail
Artificial Intelligence • Security • Software • Analytics • Big Data Analytics
Lake Oswego, OR
1500 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account