Find your next role
Strengthen your profile
RAPIDFORT
Cybersecurity
See how your profile stacks up against this role.
We compared the job requirements to your profile to show where you're strong and where you fall short.
Platform Operations Engineer
Role Summary
We are looking for a Platform Operations Engineer to join our DevOps/SRE organization and provide first-line operational support for our SaaS platform.
This role is focused on the day-to-day operation of our production and delivery environments: monitoring systems, responding to alerts, performing deployments and builds, handling incoming operational requests, and providing initial troubleshooting and escalation during incidents.
Candidates must have hands-on experience working with Linux and Kubernetes and should have practical familiarity with most of the following.
The goal of this role is to provide reliable operational coverage while allowing senior DevOps/SRE engineers to focus on infrastructure, reliability, automation, and longer-term engineering projects.
We are building this function around a follow-the-sun operational model, initially with coverage in India and Europe. Engineers will work scheduled shifts designed to provide coverage across regions. Some flexibility in working hours may be required based on operational needs, incidents, and on-call responsibilities.
This is also intended to be a growth role. Over time, successful engineers will take on deeper infrastructure, automation, reliability, and SRE responsibilities.
Responsibilities
Monitor production and SaaS environments and proactively identify potential issues.
Respond to monitoring alerts and perform initial investigation and remediation.
Perform routine application and platform deployments using established processes and tooling.
Create, trigger, and monitor software builds and release workflows.
Handle incoming operational requests from engineering and other internal teams.
Acknowledge and take ownership of requests until they are resolved or successfully handed off to the appropriate team.
Follow documented procedures and runbooks for common operational tasks and incidents.
Perform first-line troubleshooting of production issues using logs, metrics, Kubernetes tooling, and other available diagnostic information.
Escalate complex incidents to senior DevOps/SRE or Development engineers when appropriate.
Initiate incident or war-room coordination when necessary and ensure the appropriate technical teams are engaged.
Perform known and approved production remediation steps, such as restarting or scaling workloads, when appropriate.
Participate in follow-the-sun operational coverage and provide clear handoffs for active incidents, deployments, and unresolved requests.
Participate in an on-call rotation.
Help create, maintain, and improve operational documentation and runbooks as procedures and recurring issues are identified.
Technical Background
Candidates should have working familiarity with most of the following. Deep expertise is not required, but candidates should understand the fundamentals and be comfortable working with these technologies:
Linux
Kubernetes
Helm
Git
CI/CD pipelines and deployment workflows
Terraform and infrastructure-as-code concepts
Public cloud platforms such as AWS, Azure, or GCP
Monitoring, logging, and alerting systems such as Datadog or similar tools
Basic networking and troubleshooting concepts
Basic Bash scripting
Experience with one major cloud platform is sufficient. We value transferable cloud and infrastructure fundamentals more than experience with a specific provider.
Familiarity with databases, storage, IAM, DNS, and cloud networking is helpful but not required.
Qualifications
Approximately 1–3 years of experience in DevOps, SRE, cloud operations, infrastructure operations, production support, or a related technical role.
Strong entry-level candidates with relevant hands-on experience and solid technical fundamentals may also be considered.
Comfortable working with production systems and following controlled operational procedures.
Able to investigate technical issues, gather useful diagnostic information, and recognize when escalation is appropriate.
Clear written and verbal communication skills, particularly when documenting incidents, handing off work, or escalating issues.
Ability to work independently during assigned shifts while collaborating with a globally distributed engineering team.
Willingness to participate in on-call responsibilities.
Willingness to learn new systems and grow into deeper DevOps/SRE responsibilities.
A specific degree or professional certification is not required.
What Success Looks Like
Within the first several months, a successful Platform Operations Engineer should be able to:
Independently handle routine deployments and builds.
Monitor systems and respond appropriately to common alerts.
Handle and route day-to-day operational requests without requiring senior engineers to manage the intake process.
Follow runbooks and established procedures for common issues.
Perform useful first-pass troubleshooting before escalating an incident.
Provide senior engineers with clear context, logs, symptoms, and actions already taken when escalation is required.
Reliably participate in follow-the-sun and on-call operational coverage.
Provide clear handoffs between regional shifts.
Identify opportunities to improve runbooks and recurring operational procedures.
Career Growth
This role is intended to provide a path into broader DevOps and Site Reliability Engineering responsibilities.
As experience grows, Platform Operations Engineers may take on additional ownership of infrastructure, Kubernetes, CI/CD systems, automation, observability, reliability engineering, and production architecture.
Location: United Kingdom — 100% Remote
Work Authorization: Candidates must already be authorised to work in the United Kingdom. RapidFort is unable to provide visa sponsorship for this position.
Schedule: This position participates in scheduled European operational coverage and an on-call rotation. Occasional flexibility outside normal working hours may be required, with on-call arrangements compensated separately or supported through time off in lieu.
After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.
Marcus Rivera
Chief Revenue Officer

iFIT

Vultr

Amplity Health

The Stratagem Group

Recorded Future

RAPIDFORT

RAPIDFORT

RAPIDFORT