Site Reliability Engineer

Posted 9 Days Ago
2 Locations
In-Office
43K-43K Annually
Mid level
Food • Logistics • Retail
The Role
Ensure reliability, scalability and performance of critical services by applying software engineering to operations: design monitoring/alerting, automate tasks, manage incidents, analyse metrics, and support capacity planning and resilience improvements.
Summary Generated by Built In
Circa £43,000 (Depending on Skills & Experience)Permanent / Full-time / 37 hours per week (Flexible working opportunities available)Huntingdon or Lincoln - HybridMake every drop of your potential count!

As a Site Reliability Engineer (SRE) in the Digital Data & Technology team, you will ensure the reliability, scalability and performance of critical systems and services. You will apply software engineering principles to operations, automate processes, and improve system resilience. Working closely with development and operations teams, you will help build monitoring, alerting and incident response capabilities that minimise downtime and enhance service levels.

What you'll be doing

• Design and implement monitoring and alerting systems for critical services

• Automate operational tasks to improve efficiency and reduce manual effort

• Collaborate with development teams to enhance system reliability and performance

• Manage incident response and post-incident reviews

• Analyse system metrics to identify trends and areas for improvement

• Contribute to capacity planning and scalability strategies

What we’re looking for

• Experience in site reliability engineering or DevOps roles

• Strong scripting and automation skills (e.g., Python, Powershell, Azure Automation)

• Strong experience working in public cloud environments such as Microsoft Azure (preferred), AWS, GCP etc. 

• Knowledge of monitoring tools and observability practices

• Understanding of cloud infrastructure and containerisation

• Excellent problem-solving and analytical abilities

• Commitment to continuous improvement and operational excellence

Benefits

As a valued employee, you’ll be entitled to:
• 26 days annual leave plus bank holidays (increasing with length of service)
• Flexible working options
• Private health care
• Competitive pension scheme – Anglian Water double-matches your contributions up to 6%
• Life assurance at eight times your salary
• Annual bonus
• Personal medical assessments
• Virtual GP service
• Cancer screening
• Financial wellbeing support and salary finance benefits
• Lifestyle Savings including discounts on retail, travel, and utilities
• Employee Assistance Programme
• Volunteer days
• Environmental and wellbeing initiatives
 

Inclusion at Anglian Water

We’re committed to creating a workplace where everyone feels they belong. We’re proud signatories of the Social Mobility Pledge, Race at Work Charter, and Armed Forces Covenant, and we’re a Disability Confident employer.

Closing Date: 9th August 2026

#loveeverydrop!

#LI-LJ1

Skills Required

  • Experience in site reliability engineering or DevOps roles.
  • Strong scripting and automation skills (Python, Powershell, Azure Automation).
  • Strong experience working in public cloud environments (AWS, GCP, Microsoft Azure).
  • Microsoft Azure experience (preferred).
  • Knowledge of monitoring tools and observability practices.
  • Understanding of cloud infrastructure and containerisation.
  • Excellent problem-solving and analytical abilities.
  • Commitment to continuous improvement and operational excellence.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Kansas City, KS

What We Do

Associated Wholesale Grocers (AWG) is the nation's largest cooperative food wholesaler to independently owned supermarkets, providing distribution, marketing, and development services.

Similar Jobs

Andromeda (andromeda.ai) Logo Andromeda (andromeda.ai)

Site Reliability Engineer

Artificial Intelligence • Cloud • Information Technology • Software
In-Office or Remote
3 Locations
17 Employees
Remote or Hybrid
United States
1750 Employees

Akamai Technologies Logo Akamai Technologies

Site Reliability Engineer

Cloud • Security • Software • Cybersecurity
In-Office or Remote
2 Locations
10285 Employees
146K-264K Annually

Fabric Health Logo Fabric Health

Site Reliability Engineer

Artificial Intelligence • Healthtech • Software • Telehealth
In-Office or Remote
2 Locations
304 Employees
135K-160K Annually

Similar Companies Hiring

Scotch Thumbnail
Artificial Intelligence • eCommerce • Fintech • Payments • Retail • Software • Analytics
US
35 Employees
Amalgamated Sugar Thumbnail
Food • Greentech • Agriculture • Industrial • Manufacturing
Boise, Idaho
768 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account