Platform Operations Engineer

Posted Yesterday
Be an Early Applicant
Hiring Remotely in United Kingdom
Remote
Junior
Cloud • Software • Cybersecurity
The Role
Provides first-line operational support for a SaaS platform by monitoring production systems, responding to alerts, performing deployments and builds, troubleshooting Kubernetes and platform issues, handling operational requests, and escalating complex incidents. The role follows runbooks, supports follow-the-sun coverage and on-call rotations, coordinates incident handoffs, and improves operational documentation. It offers progression into broader DevOps, infrastructure, automation, observability, and site reliability engineering responsibilities.
Summary Generated by Built In
Platform Operations EngineerRole Summary

We are looking for a Platform Operations Engineer to join our DevOps/SRE organization and provide first-line operational support for our SaaS platform.

This role is focused on the day-to-day operation of our production and delivery environments: monitoring systems, responding to alerts, performing deployments and builds, handling incoming operational requests, and providing initial troubleshooting and escalation during incidents.

Candidates must have hands-on experience working with Linux and Kubernetes and should have practical familiarity with most of the following.

The goal of this role is to provide reliable operational coverage while allowing senior DevOps/SRE engineers to focus on infrastructure, reliability, automation, and longer-term engineering projects.

We are building this function around a follow-the-sun operational model, initially with coverage in India and Europe. Engineers will work scheduled shifts designed to provide coverage across regions. Some flexibility in working hours may be required based on operational needs, incidents, and on-call responsibilities.

This is also intended to be a growth role. Over time, successful engineers will take on deeper infrastructure, automation, reliability, and SRE responsibilities.

Responsibilities
  • Monitor production and SaaS environments and proactively identify potential issues.

  • Respond to monitoring alerts and perform initial investigation and remediation.

  • Perform routine application and platform deployments using established processes and tooling.

  • Create, trigger, and monitor software builds and release workflows.

  • Handle incoming operational requests from engineering and other internal teams.

  • Acknowledge and take ownership of requests until they are resolved or successfully handed off to the appropriate team.

  • Follow documented procedures and runbooks for common operational tasks and incidents.

  • Perform first-line troubleshooting of production issues using logs, metrics, Kubernetes tooling, and other available diagnostic information.

  • Escalate complex incidents to senior DevOps/SRE or Development engineers when appropriate.

  • Initiate incident or war-room coordination when necessary and ensure the appropriate technical teams are engaged.

  • Perform known and approved production remediation steps, such as restarting or scaling workloads, when appropriate.

  • Participate in follow-the-sun operational coverage and provide clear handoffs for active incidents, deployments, and unresolved requests.

  • Participate in an on-call rotation.

  • Help create, maintain, and improve operational documentation and runbooks as procedures and recurring issues are identified.

Technical Background

Candidates should have working familiarity with most of the following. Deep expertise is not required, but candidates should understand the fundamentals and be comfortable working with these technologies:

  • Linux

  • Kubernetes

  • Helm

  • Git

  • CI/CD pipelines and deployment workflows

  • Terraform and infrastructure-as-code concepts

  • Public cloud platforms such as AWS, Azure, or GCP

  • Monitoring, logging, and alerting systems such as Datadog or similar tools

  • Basic networking and troubleshooting concepts

  • Basic Bash scripting

Experience with one major cloud platform is sufficient. We value transferable cloud and infrastructure fundamentals more than experience with a specific provider.

Familiarity with databases, storage, IAM, DNS, and cloud networking is helpful but not required.

Qualifications
  • Approximately 1–3 years of experience in DevOps, SRE, cloud operations, infrastructure operations, production support, or a related technical role.

  • Strong entry-level candidates with relevant hands-on experience and solid technical fundamentals may also be considered.

  • Comfortable working with production systems and following controlled operational procedures.

  • Able to investigate technical issues, gather useful diagnostic information, and recognize when escalation is appropriate.

  • Clear written and verbal communication skills, particularly when documenting incidents, handing off work, or escalating issues.

  • Ability to work independently during assigned shifts while collaborating with a globally distributed engineering team.

  • Willingness to participate in on-call responsibilities.

  • Willingness to learn new systems and grow into deeper DevOps/SRE responsibilities.

A specific degree or professional certification is not required.

What Success Looks Like

Within the first several months, a successful Platform Operations Engineer should be able to:

  • Independently handle routine deployments and builds.

  • Monitor systems and respond appropriately to common alerts.

  • Handle and route day-to-day operational requests without requiring senior engineers to manage the intake process.

  • Follow runbooks and established procedures for common issues.

  • Perform useful first-pass troubleshooting before escalating an incident.

  • Provide senior engineers with clear context, logs, symptoms, and actions already taken when escalation is required.

  • Reliably participate in follow-the-sun and on-call operational coverage.

  • Provide clear handoffs between regional shifts.

  • Identify opportunities to improve runbooks and recurring operational procedures.

Career Growth

This role is intended to provide a path into broader DevOps and Site Reliability Engineering responsibilities.

As experience grows, Platform Operations Engineers may take on additional ownership of infrastructure, Kubernetes, CI/CD systems, automation, observability, reliability engineering, and production architecture.

Location: United Kingdom — 100% Remote

Work Authorization: Candidates must already be authorised to work in the United Kingdom. RapidFort is unable to provide visa sponsorship for this position.
Schedule: This position participates in scheduled European operational coverage and an on-call rotation. Occasional flexibility outside normal working hours may be required, with on-call arrangements compensated separately or supported through time off in lieu.

Skills Required

  • Approximately 1–3 years of experience in DevOps, SRE, cloud operations, infrastructure operations, production support, or a related technical role
  • Hands-on experience working with Linux and Kubernetes
  • Working familiarity with Helm, Git, CI/CD pipelines, deployment workflows, Terraform, infrastructure-as-code concepts, and at least one public cloud platform
  • Familiarity with monitoring, logging, alerting systems, basic networking, troubleshooting, and Bash scripting
  • Comfort working with production systems and controlled operational procedures
  • Ability to investigate technical issues, gather diagnostic information, and recognize when escalation is appropriate
  • Clear written and verbal communication skills for incident documentation, handoffs, and escalations
  • Ability to work independently during assigned shifts with a globally distributed engineering team
  • Willingness to participate in on-call responsibilities and scheduled operational coverage
  • Authorization to work in the United Kingdom without visa sponsorship
  • Familiarity with databases, storage, IAM, DNS, and cloud networking
  • Specific degree or professional certification
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
119 Employees
Year Founded: 2020

What We Do

RapidFort is a software supply chain security company that helps enterprises and government agencies secure containerized applications. Its platform automatically transforms vulnerable container images into hardened, production-ready artifacts, combining container hardening, vulnerability remediation, runtime intelligence, and policy automation. By reducing attack surface and operational complexity, RapidFort enables engineering teams to deliver secure software faster across cloud-native environments and modern development workflows.

Similar Jobs

Kraken Digital Asset Exchange Logo Kraken Digital Asset Exchange

Platform Engineer

Blockchain • Financial Services • Cryptocurrency • Web3
Remote
22 Locations
2900 Employees
Remote
2 Locations
125 Employees
220K-240K Annually

GitLab Logo GitLab

Engagement Manager

Cloud • Security • Software • Cybersecurity • Automation
Easy Apply
Remote
United Kingdom
2500 Employees

Coursera + Udemy  Logo Coursera + Udemy

Program Manager

Artificial Intelligence • Consumer Web • Edtech • Enterprise Web • HR Tech • Social Impact • Generative AI
Remote or Hybrid
United Kingdom
1500 Employees

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account