Match working

Site Reliability Engineer

72% Flex
Full Remote
  • Remote from:United Kingdom
Request priority access (3/3)

Site Reliability Engineer

72% Flex
Remote: Full Remote
Work from: United Kingdom...

Offer summary


Experience in performance monitoring and analysis, Infrastructure as Code experience, especially Terraform, Familiarity with SRE processes and DevOps principles.

Key responsabilities:

  • Proactively monitor platform performance
  • Assist with implementing SLOs
  • Ensure service is highly available and resilient
  • Collaborate with teams for scalability
  • Conduct capacity assessments and planning
Arbor Education logo
Match working

Arbor Education


51 - 200 Employees

Job description

Logo Jobgether

Your missions

Location: Remote

Salary: £55,000 - £65,000

About us

At Arbor, we’re on a mission to transform the way schools work for the better.

You’ve probably seen the headlines. Heavy workloads, constant change, admin pressure on teachers and staff at every level… sometimes it feels like this is just part and parcel of school life today. But it doesn’t have to be this way.

We passionately believe that there’s a better way to work. And it starts by giving everyone the right tools and technology for the job.

We’re building a platform and products we believe in - as well as a strong, diverse team of experienced specialists, ex-teachers and Edtech engineers passionate about making a difference to the sector.

Ultimately, we’re here to help make our schools and trusts stress a little less, and focus on what matters most - improving the lives of teachers and outcomes of students everywhere.

About the role

We are looking for an enthusiastic and proactive Site Reliability Engineer to join our SRE team and help us ensure we provide world-class resilience and performance across the platform. The remit and focus of the role is to advise on all aspects of site reliability including availability, scalability, observability and capacity planning. It’s a broad and exciting role, so we’re looking for someone up for a challenge - if you’re an energetic and a collaborative Site Reliability Engineer, this is the role for you.

Core responsibilities

  • Proactively monitor and analyse platform performance.

  • Collaborate with engineering teams to address performance bottlenecks and ensure scalability.

  • Assist engineering teams with implementing and reviewing SLOs

  • Continually improve observability through monitoring and alerting, and dashboards, using tools such as DataDog or Prometheus for example.

  • Work with other teams to ensure it is effective and provides full coverage.

  • Ensure the service is highly available and resilient

  • Champion best practices in design for high availability

  • Devise runbooks and run game sessions to test our DR plan, H/A and backups

  • Conduct assessments of capacity and plan for scaling to meet current and future business needs.

  • Work closely with the Head of Platform Engineering and Head of SRE to strategize and implement scalable solutions.

  • Work closely with the Platform team, feature teams and, 2nd line support and other stakeholders to ensure a good level of service is provided for our customers and embed SRE practices.

  • Key player in the response and troubleshooting of incidents, ensuring rapid resolution and minimising downtime.

  • Participate in blameless postmortems to identify root cause and corrective actions

  • Develop and maintain playbooks and documentation

About you

  • Experience in performance monitoring and analysis

  • Capacity planning experience

  • Scripting and automation skills, with experience in relevant technologies.

  • Experience with Infrastructure as Code, in particular, Terraform

  • Understanding of relational database technologies and their cloud versions (e.g. AWS Aurora)

  • Experience with messaging and distributed asynchronous workloads

  • Experience with nginx or similar technologies

  • Familiarity with SRE processes.

  • Aware of DevOps principles like the 3 ways and 5 ideals.

Bonus Skills

  • Experience with other database technologies and cloud platforms.

  • Past experience with enterprise solutions running at scale

  • Familiarity with kanban and agile development processes

  • Experience with containerisation, for example Docker

  • Familiarity with software best practices such as Refactoring, Clean Code, Domain-Driven Design and Test-Driven Development.

What we offer

The chance to work alongside a team of hard-working, passionate people in a role where you’ll see the impact of your work everyday. We also offer:

  • A dedicated wellbeing team who champion initiatives such as mindfulness, lunch n learns, manager training, mental health first aid training and much more!

  • 32 days holiday (plus Bank Holidays). This is made up of 25 days annual leave plus 7 extra company wide days given over Easter, Summer & Christmas

  • Private Bupa Dental Insurance

  • Enhanced maternity and adoption leave (20 weeks full pay) and paternity (6 weeks full pay) pay

  • 5 free return to work maternity coaching sessions, helping you adapt to this new exciting time of life!

  • Access to services such as Calm, Bippit (financial wellbeing coaching) and Health Assured (Employee assistance programme)

  • All of our roles champion flexible working and we are happy to discuss what this means to you!

  • Social committees that plan team, office and company wide events to bring people together and celebrate success

  • Dedicated professional development training budget (CPD courses, upskilling resources, professional memberships etc)

  • Volunteer with a charity of your choice for a day each year

  • Dog friendly offices!

Interview process

  1. Phone screen

  2. 1st stage

  3. 2nd stage

We are committed to a fair and comfortable recruitment process, so if you require any reasonable adjustments during your application or interview process, please reach out to a member of the team at

Our commitment is also backed by our partnership with Neurodiversity Consultancy, Lexxic who provide us with training, support and advice.

Arbor Education is an equal opportunities organisation

Our goal is for Arbor to be a workplace which represents, celebrates and supports people from all backgrounds, and which gives them the tools they need to thrive - whatever their ambitions may be so we support and promote diversity and equality, and actively encourage applications from people of all backgrounds.

Refer a friend: Know someone else who would be good for this role? You can refer a friend, family member or colleague, if they are offered a role with Arbor, we will say thank you with a voucher valued up to £200! Simply email:

Please note: We are unable to provide visa sponsorship at this time.

See more

Required profile

Match working


Industry :
Spoken language(s)
Check out the description to know which languages are mandatory.

Soft Skills

  • Analytical mindset
  • Effective team player

Go Premium: Access the World's Largest Selection of Remote Jobs!

  • Largest Inventory: Dive into the world's largest remote job inventory. More than half of these opportunities can't be found on standard platforms.
  • Personalized Matches: Our AI-driven algorithms ensure you find job listings perfectly matched to your skills and preferences.
  • Application fast-lane: Discover positions where you rank in the TOP 5% of applicants, and get personally introduced to recruiters with Jobgether.
  • Try out our Premium Benefits with a 7-Day FREE TRIAL.
    No obligations. Cancel anytime.

Find other similar jobs

🚀 Go Premium Today!
Unlock Unlimited Access to the Largest Remote Job Platform!


Go Premium Today!
Unlock Unlimited Access to the Largest Remote Job Platform!

  • Discover all Matching Remote Jobs available Worldwide
  • Boost your hiring chances: Apply faster and gain Priority Access to Recruiters
Start Your Free TrialDon’t ask again