Cloud Operations Engineer

Posted 3 Days Ago
Hiring Remotely in United States
Remote or Hybrid
109K-178K Annually
Senior level
Healthtech
The Role
Operate and improve production AWS workloads through automation, monitoring, patching, backups, certificate renewal, disaster recovery, and restore testing. Define operational readiness standards, manage SLOs and error budgets, respond to incidents, participate in on-call rotation, troubleshoot Windows and Linux compute, and create self-healing workflows. Partner with infrastructure teams to transition migrated workloads into steady-state support while maintaining runbooks and operational documentation.
Summary Generated by Built In

 

We're moving workloads to AWS, and this is the person who keeps them running afterward: patched, backed up, monitored, and recoverable. The Cloud Operations Engineer owns that ongoing operational work and builds automation so it doesn't keep growing headcount as the estate grows.

This role is a standard business-hours on eastern standard time zone engineering role with a shared on-call rotation, not a NOC or shift job. The focus is building automation and operational standards, not watching dashboards.

What You Will Be Doing

 

  • Build and maintain monitoring and alerting for migrated workloads. Alert on conditions that require action.

  • Automate patching, backup validation, certificate renewal, and other recurring operational tasks.

  • Define and enforce operational readiness criteria a workload must meet before it goes live in cloud.

  • Build and test disaster recovery runbooks. Validate recovery time and recovery point targets through failover exercises.

  • Respond to incidents, participate in the on-call rotation, and drive root cause to closure.

  • Convert manual procedures into executable automation and self-healing where the task is repeatable.

  • Track availability and operational health against defined targets. Report on recurring issues.

  • Manage backup, retention, and restore testing for cloud workloads.

  • Troubleshoot and maintain AWS compute, including Windows Server and Linux instances.

  • Partner with infrastructure operations to transition migrated workloads into steady-state support.

  • Maintain runbooks and operational documentation a new on-call engineer can use unaided.

What We Require

  • 6+ years related work experience in infrastructure operations or systems engineering with 4+ years experience operating cloud workloads 
  • 1+ years direct supervisory/management experience 
  • Related Bachelor's degree required 
  • Strong hands-on AWS experience running production workloads. 
  • Proficiency in a programming or scripting language such as Python or Go. 
  • Hands-on Terraform and CI/CD pipeline experience. 
  • Experience defining and operating against SLOs and error budgets. 
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent. 
  • Incident management and on-call experience in a production environment. 
  • Strong hands-on AWS experience running production workloads. 
  • Proficiency in a programming or scripting language such as Python or Go. 
  • Hands-on Terraform and CI/CD pipeline experience. 
  • Experience defining and operating against SLOs and error budgets. 
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent. 
  • Incident management and on-call experience in a production environment. 
  • Proficiency in a programming or scripting language such as Python or Go. 
  • Hands-on Terraform and CI/CD pipeline experience. 
  • Experience defining and operating against SLOs and error budgets. 
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent. 
  • Incident management and on-call experience in a production environment. 
  • Kubernetes or EKS in production.

 

What We Prefer 

 

  • Multi-region or disaster recovery design, including failover testing. 

  • Healthcare, financial services, or other regulated industry experience. 

  • AWS GovCloud experience. Chaos engineering or resilience testing. 

  • AWS certification such as DevOps Engineer, SysOps Administrator, or Solutions Architect. 
     

General Physical Demands

 

  • Exerting up to 10 pounds of force occasionally to move objects. 

  • Jobs are sedentary if traversing activities are required only occasionally. 

  • This position offers a hybrid work arrangement which combines flexibility with in office engagement. Team schedules are determined based on role requirements and business needs. Travel to and from our corporate offices may be required.
     

What We Offer
As a Florida Blue employee, you will be at the heart of GuideWell’s vision – to lead the nation in transforming health through compassionate, connected, and technology-enabled care that delivers personalized value and empowered living. 
To support your wellbeing, comprehensive benefits are offered. As an employee, you will have access to: 
 

  • Medical, dental, vision, life and global travel health insurance
  • Income protection benefits: life insurance, short- and long-term disability programs
  • Leave programs to support personal circumstances
  • Retirement Savings Plan including employer match
  • Paid time off, volunteer time off, 10 holidays and 2 well-being days
  • Additional voluntary benefits available; and a comprehensive wellness program

Employee benefits are designed to align with federal and state employment laws. Benefits may vary based on the state in which work is performed. Benefits for intern, part-time and seasonal employees may differ.
To support your financial wellbeing, we offer competitive pay as well as opportunities for incentive or commission compensation. We also conduct regular annual reviews with pay for performance considerations for base pay increases. 

Typical Annualized Offer/Hiring Range: $109,300 - $136,600
Annualized Salary Range: $109,300 - $177,600
Final pay will be determined with consideration of market competitiveness, internal equity, and the job-related knowledge, skills, training, and experience you bring.
We are an Equal Employment Opportunity employer committed to cultivating a work experience where everyone feels like they belong and can perform at their best in pursuit of our mission. All qualified applicants will receive consideration for employment.

Skills Required

  • 6+ years of related experience in infrastructure operations or systems engineering
  • 4+ years of experience operating cloud workloads
  • 1+ years of direct supervisory or management experience
  • Related bachelor's degree
  • Strong hands-on AWS experience running production workloads
  • Proficiency in a programming or scripting language such as Python or Go
  • Hands-on Terraform experience
  • Hands-on CI/CD pipeline experience
  • Experience defining and operating against SLOs and error budgets
  • Observability tooling experience with CloudWatch, Prometheus, Grafana, Datadog, or equivalent
  • Incident management and production on-call experience
  • Kubernetes or EKS production experience
  • Multi-region or disaster recovery design, including failover testing
  • Healthcare, financial services, or other regulated industry experience
  • AWS GovCloud experience
  • Chaos engineering or resilience testing experience
  • AWS certification such as DevOps Engineer, SysOps Administrator, or Solutions Architect
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Jacksonville, FL
200 Employees
Year Founded: 2014

What We Do

GuideWell Mutual Holding Corporation is a not-for-profit mutual holding company that is the parent to a family of forward-thinking companies focused on transforming health care. We’re at the forefront, forging ahead by innovating, collaborating and advocating for better health. We help people make sense of this new world, forming an integrated ecosystem of products and services and ensuring they get the best experience. We’re relentlessly building and refining to drive higher efficiency and exceptional care. GuideWell – Built for the future of health.

Similar Jobs

Optum Logo Optum

Cloud Engineer

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
In-Office or Remote
Minnetonka, MN, USA
160000 Employees
113K-193K Annually

Optum Logo Optum

Senior I O Engineer - Azure Cloud Ops and Linux-Remote (Nationwide)

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
In-Office or Remote
Eden Prairie, MN, USA
160000 Employees
92K-164K Annually
In-Office or Remote
Redmond, WA, USA
2900 Employees
50-50 Hourly
Remote or Hybrid
United States
200 Employees
109K-178K Annually

Similar Companies Hiring

Sailor Health Thumbnail
Healthtech • Social Impact • Telehealth
New York City, NY
20 Employees
Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
OneImaging Thumbnail
Healthtech
Miami, FL
62 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account