Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
Atlanta, GA, USA
In-Office
98K-202K Annually
Senior level
Insurance
The Role
Design, build, and maintain observability, monitoring, and alerting for cloud and on-prem systems. Develop reliability dashboards and automation, troubleshoot distributed systems, support CI/CD pipeline reliability, participate in incident response and root-cause analysis, and collaborate with engineering and run teams to improve service reliability and operational processes.
Summary Generated by Built In
Job Posting End Date: August 04

Our Fortune 500 company is driving a digital transformation and looking for forward-thinking innovators to disrupt how our industry thinks about and uses technology. As one of the world's leading employee benefits providers, we help millions of people gain affordable access to benefits that help them protect their families, their finances and their futures.

Are you an asker of questions, a solver of problems, and a challenger of the status quo? Our mission is to provide a differentiated customer experience and exceed the expectations people have of technology at any company — not just insurers. 

We are seeking individuals to join our team of talented IT professionals who share never-ending passion and an unwavering focus on our customer experience. Team members comfortable working in an agile, fast-paced, and delivery-focused environment thrive in our environment where we value an entrepreneurial spirit and those who challenge the status-quo.

Unum is changing, and we’re excited about what’s next. Join us.

General Summary:

Unum Group seeks Site Reliability Engineers in Atlanta, GA.

Applicants who are interested in this position may apply at www.jobpostingtoday.com (Ref #66753) for consideration.

  • Design, build, and maintain observability, monitoring, and alerting capabilities across consumer and client-facing digital platforms
  • Develop and maintain dashboards that measure availability, latency, error rate, throughput, capacity, MTTR/MBTI, and other reliability metrics
  • Diagnose and troubleshoot distributed systems issues across cloud-based and on-prem services
  • Partner with engineering teams to improve service reliability, reduce operational toil, and mature incident response practices
  • Implement automation for service health checks, performance monitoring, and remediation
  • Manage CI/CD pipeline reliability and deployment quality controls
  • Conduct root-cause analysis and drive long-term corrective actions
  • Collaborate with Run teams to transition monitoring, dashboards, and operational insights into production support processes
  • Provide guidance on service re-platforming, performance improvements, and architectural decisions based on reliability data

Requires a Bachelor’s degree in Computer Science, Engineering, or related field plus 5 years of experience. Requires 5 years of experience with the following: Observability and monitoring platforms used to monitor application performance and system health using Dynatrace, AWS CloudWatch, Datadog, Grafana, or Amplitude; Working with containerized and cloud-native architectures, including deployment, configuration, and operational support in cloud environments, using AWS; Supporting incident response processes, including participation in on-call rotations, post-incident reviews, and implementation of service-level objectives (SLOs), service-level indicators (SLIs), or service-level agreements (SLAs); Developing scripts or automation to improve system reliability or operational efficiency using Python, Bash, or PowerShell; Troubleshooting distributed systems and analyzing performance bottlenecks across multi-tier or microservices-based architectures; Collaborating with cross-functional engineering teams, including software engineering, platform, infrastructure, or operations teams, within a DevOps or reliability-focused environment; working with version control systems and collaborative development workflows using GitHub, GitLab, or Bitbucket. Requires 4 years of experience with the following: Designing, implementing, or maintaining logging, metrics, and distributed tracing pipelines for enterprise or cloud-based systems; Hands-on experience with continuous integration and continuous deployment (CI/CD) tools and pipelines, using GitHub Actions, Jenkins, or Azure DevOps. Requires 3 years of experience with using infrastructure-as-code or configuration management tools to provision, manage, or maintain environments, including Terraform, AWS CloudFormation, or Ansible. Telecommuting w/i worksite. Up to 5% domestic travel.

40 hours/week; $152,131 - $162,131 per year. This wage range supersedes the base salary range listed below, due to the salary range below reflecting a national range.

#LI-TS1

~IN1

Our company is built on helping individuals and families, and this starts with our employees. We want employees to maintain a positive balance, which is why we provide access to the benefits and resources they need to invest in themselves. From our onsite fitness facilities and generous paid time off to employee professional development programs, we are committed to helping employees live and work their best – both inside and outside the office.

Unum is an equal opportunity employer, considering all qualified applicants and employees for hiring, placement, and advancement, without regard to a person's race, color, religion, national origin, age, genetic information, military status, gender, sexual orientation, gender identity or expression, disability, or protected veteran status.

The base salary range for applicants for this position is listed below. Unless actual salary is indicated above in the job description, actual pay will be based on skill, geographical location and experience.

$98,340.00-$201,900.00

Additionally, Unum offers a portfolio of benefits and rewards that are competitive and comprehensive including healthcare benefits (health, vision, dental), insurance benefits (short & long-term disability), performance-based incentive plans, paid time off, and a 401(k) retirement plan with an employer match up to 5% and an additional 4.5% contribution whether you contribute to the plan or not.  All benefits are subject to the terms and conditions of individual Plans.

Company:

Unum

Skills Required

  • Bachelor's degree in Computer Science, Engineering, or related field
  • 5 years of overall relevant experience
  • 5 years experience with observability/monitoring platforms (Dynatrace, AWS CloudWatch, Datadog, Grafana, or Amplitude)
  • 5 years experience with containerized and cloud-native architectures and operational support in AWS
  • 5 years supporting incident response, on-call rotations, post-incident reviews, and implementing SLOs/SLIs/SLAs
  • 5 years developing automation or scripts to improve reliability using Python, Bash, or PowerShell
  • 5 years troubleshooting distributed systems and analyzing performance bottlenecks in multi-tier or microservices architectures
  • 5 years collaborating with cross-functional engineering, platform, infrastructure, or operations teams in DevOps/reliability environments
  • 5 years working with version control systems and collaborative workflows using GitHub, GitLab, or Bitbucket
  • 4 years designing, implementing, or maintaining logging, metrics, and distributed tracing pipelines for enterprise or cloud systems
  • 4 years hands-on experience with CI/CD tools and pipelines (GitHub Actions, Jenkins, or Azure DevOps)
  • 3 years experience using infrastructure-as-code or configuration management tools (Terraform, AWS CloudFormation, or Ansible)

Unum Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Unum and has not been reviewed or approved by Unum.

  • Healthcare Strength Benefits are described as starting day one and include medical coverage along with employer-provided life, short-term disability, and long-term disability. On-site amenities like a gym and café are also part of the health-and-extras package in some locations.
  • Retirement Support The retirement offering is positioned as a standout through a high employer 401(k) contribution. This retirement support is repeatedly framed as a key draw within the total rewards package.
  • Leave & Time Off Breadth Paid time off is characterized as generous, with additional personal days and company closure days also described. The time-off program is presented as broader than many peers on paper.

Unum Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Chattanooga, TN
9,536 Employees

What We Do

Unum is a company of people serving people. A Fortune 500 company, we help millions of people gain affordable access to disability, life, accident, critical illness, dental and vision benefits through the workplace — benefits that help them protect their families, their finances and their futures. Our products and services meet the needs of a diverse workforce that includes four generations, growing ethnic diversity and changing family dynamics. A strong commitment to social responsibility is one of Unum's core values. In fact, we place particular emphasis on contributing to positive change in the communities in which we live and work. Helping our communities become better is a natural extension of the commitment we make each and every day to our customers — to help employers manage their businesses and employees protect their families and livelihoods. We look forward to sharing industry news, community involvement, recognition and other info we hope you find interesting and informative. We encourage your comments and feedback so we can be a valuable resource for you and get to know you a little better.

Similar Jobs

DraftKings Logo DraftKings

Site Reliability Engineer

Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Remote or Hybrid
United States
6400 Employees
200K-250K Annually
In-Office
30339, Atlanta, GA, USA
3700 Employees

Akamai Technologies Logo Akamai Technologies

Site Reliability Engineer

Cloud • Security • Software • Cybersecurity
In-Office or Remote
2 Locations
10285 Employees
146K-264K Annually

Fabric Health Logo Fabric Health

Site Reliability Engineer

Artificial Intelligence • Healthtech • Software • Telehealth
In-Office or Remote
2 Locations
304 Employees
135K-160K Annually

Similar Companies Hiring

Globe Life Thumbnail
Insurance • Financial Services
McKinney, TX
3000 Employees
MassMutual India Thumbnail
Big Data • Fintech • Information Technology • Insurance • Financial Services
Hyderabad, Telangana
Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account