Reliability Engineer

Posted 2 Days Ago
Be an Early Applicant
4 Locations
In-Office
91K-137K Annually
Mid level
Fintech • Payments • Financial Services
The Role
Lead infrastructure resilience by implementing observability, IaC, automation, and AI-driven detection to ensure cloud and SaaS system stability. Drive tooling, alerts, self-healing, CI/CD optimization, incident triage/restore, and partner across architecture and engineering to improve performance, reliability, and cost efficiency.
Summary Generated by Built In
Reliability Engineer - IE08GE

We’re determined to make a difference and are proud to be an insurance company that goes well beyond coverages and policies. Working here means having every opportunity to achieve your goals – and to help others accomplish theirs, too. Join our team as we help shape the future.   

         

The Hartford’s Corporate / HIMCO IT team is seeking a highly motivated, detail-oriented, and results-driven Reliability Engineer to join our team. This position will play a crucial role to lead infrastructure resilience in ensuring the stability and performance of our systems in cloud and SAAS environments.

Successful candidates will be expected to demonstrate strong technical skills, excellent partnership with stakeholders and partner teams, willingness to understand existing processes and systems, solid technical acumen, experience in delivering quality technical solutions and ensure the systems are stable, performant, and secure.

Responsibilities:

  • Assist in the use of best-in-class software engineering standards and design practices for instrumenting code/application technology stack to enable the generation of relevant metrics on overall technology health - availability, performance, quality, currency and resiliency.
  • Assist the architecture and software engineering teams to influence the technical strategy for the organization, keeping in mind its cross-functional impacts, integration across the organization, and architecture rationalization.
  • Assist on a team as a technical leader for the applications supported, requiring depth and breadth of knowledge in technologies, applications, integration, interfaces and business domain.

DevSecOps Solution Responsibilities:

  • Assist in developing effective tooling, alerts, and response mechanisms to identify and address reliability risks leveraging automation to support problem prevention, detection, mitigation, and resolution.
  • Assist in enhancing the delivery flow by engineering the appropriate solutions to increase delivery speed while adhering to technology standards for sustained reliability.
  • Partner to implement preventative controls and drive increased automation and self-healing capabilities. Continue to improve cost efficiency baselines
  • Promote and implement innovative solutions.

IT Ops Responsibilities:

  • Ensure operational excellence. Collaborate to drive the triaging and service restoration of all high impact incidents in order to minimize the mean time to service restoration and impact to the business. Demonstrate end-to-end ownership.
  • Partner with infrastructure teams to design and implement intelligent incident routing, enhanced monitoring/alerting capabilities and automated service restoration processes. Take proactive measures to prevent high impactful incidents.
  • Achieve and maintain the continuity of Hartford and third-party assets that support a business function. Accountable for keeping the IT application and infrastructure metadata repositories current.

AI-Driven Automation:

  • Research and implement AI-based anomaly detection to predict infrastructure failures and automate preventive measures.
  • Develop AI-powered troubleshooting copilots and LLM-driven operational assistants to accelerate incident resolution and root cause analysis.
  • Implement AI/ML-based runbooks to automate system recovery and optimize operational efficiency.

Qualifications:

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
  • 3+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), or DevOps.
  • Hands-on experience with observability tools: Splunk, Dynatrace, CloudWatch.
  • Deep knowledge of Infrastructure as Code (IaC) with Terraform, CloudFormation.
  • Proven ability to optimize CI/CD pipelines, automate deployments, and enforce DevSecOps best practices.
  • Expertise in cloud platforms (AWS) and Kubernetes-based microservices environments.
  • Strong proficiency in Python, Java for infrastructure automation and tooling development.
  • Experience in AI/ML frameworks for observability, predictive failure detection, and AI-driven troubleshooting desirable
  • Experience with Oracle and SQL Server relational database technologies. Knowledge of open-source database technologies is beneficial.
  • Demonstrated experience working within Agile frameworks and methodologies.
  • Excellent analytical, problem solving and interpersonal skills.

This role will have a Hybrid work schedule, with the expectation of working in an office (Columbus, OH, Chicago, IL, Hartford, CT or Charlotte, NC) 3 days a week (Tuesday through Thursday).

Candidates must be authorized to work in the US without company sponsorship. The company will not support the STEM OPT I-983 Training Plan endorsement for this position.

As a condition of your employment for HIMCO, you will be required to affirm to HIMCO’s Code of Ethics and understand that you will be required to comply with the disclosure of accounts, holdings and pre-clearance of trades for the accounts of you and your household family members as more fully described in the Code of Ethics Key Points.  If you will be deemed to be a “Covered Associate” under HIMCO’s Pay to Play Policy, you will also need to disclose all political contributions that you have given within the past 2 calendar years.

Compensation

The listed annualized base pay range is primarily based on analysis of similar positions in the external market. Actual base pay could vary and may be above or below the listed range based on factors including but not limited to performance, proficiency and demonstration of competencies required for the role. The base pay is just one component of The Hartford’s total compensation package for employees. Other rewards may include short-term or annual bonuses, long-term incentives, and on-the-spot recognition. The annualized base pay range for this role is:

$91,200 - $136,800

Equal Opportunity Employer/Sex/Race/Color/Veterans/Disability/Sexual Orientation/Gender Identity or Expression/Religion/Age

About Us | Our Culture | What It’s Like to Work Here | Perks & Benefits

Skills Required

  • Bachelor's or Master's degree in Computer Science, Engineering, or related field.
  • 3+ years of experience in Infrastructure Engineering, Site Reliability Engineering (SRE), or DevOps.
  • Hands-on experience with observability tools: Splunk, Dynatrace, CloudWatch.
  • Deep knowledge of Infrastructure as Code (IaC) with Terraform and CloudFormation.
  • Proven ability to optimize CI/CD pipelines, automate deployments, and enforce DevSecOps best practices.
  • Expertise in cloud platforms (AWS).
  • Experience with Kubernetes-based microservices environments.
  • Strong proficiency in Python for infrastructure automation and tooling development.
  • Strong proficiency in Java for infrastructure automation and tooling development.
  • Experience with Oracle and SQL Server relational database technologies.
  • Experience with AI/ML frameworks for observability, predictive failure detection, and AI-driven troubleshooting.
  • Knowledge of open-source database technologies.
  • Demonstrated experience working within Agile frameworks and methodologies.
  • Excellent analytical, problem solving and interpersonal skills.
  • Hybrid work schedule with in-office presence (Columbus OH, Chicago IL, Hartford CT or Charlotte NC) 3 days/week.
  • Authorized to work in the US without company sponsorship (no visa sponsorship or STEM OPT support).

The Hartford Financial Services Group, Inc. Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about The Hartford Financial Services Group, Inc. and has not been reviewed or approved by The Hartford Financial Services Group, Inc..

  • Retirement Support A 401(k) with matching plus an additional company contribution, alongside an employee stock purchase plan and no‑cost financial planning, signals robust long‑term savings support. HSAs/FSAs and related financial tools further strengthen overall financial well‑being.
  • Leave & Time Off Breadth At least 25 days of PTO to start, options to buy or roll over time, and paid parental leave indicate broad time‑off support. Paid leave for organ and bone marrow donation and generous disability coverage extend protection for significant life events.
  • Healthcare Strength Multiple medical, dental, and vision options with the company covering most medical and dental premiums reflect strong core health coverage. Wellness programs, fitness reimbursements, well‑being credits, and accessible behavioral health services expand depth and accessibility.

The Hartford Financial Services Group, Inc. Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Hartford, Connecticut
20,002 Employees
Year Founded: 1810

What We Do

Human achievement is at the heart of what we do. We put our belief into action by not only ensuring individuals and businesses are well protected, but by going even further – making an impact in ways that go beyond an insurance policy

Similar Jobs

DraftKings Logo DraftKings

Reliability Engineer

Digital Media • Gaming • Information Technology • Software • Sports • Esports • Big Data Analytics
Remote or Hybrid
United States
6400 Employees
168K-210K Annually
Remote or Hybrid
2 Locations
289097 Employees
Remote or Hybrid
3 Locations
289097 Employees

Vantage Data Centers Logo Vantage Data Centers

Reliability Engineer

Information Technology • Consulting
In-Office
5 Locations
1421 Employees
135K-145K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account