Now hiring

Site Reliability Engineer @ ReVybe IT Recruitment Limited

City Road 124, LondonOnsiteFull-time
Apply with ResuMinder

Opens on the employer's site

About this role

Salary: £70,000 - 85,000 per year

Requirements: Proven commercial experience working as an SRE, DevOps Engineer, Platform Engineer, or similarStrong hands-on experience with AWSStrong experience working with KubernetesExcellent experience with Terraform and Infrastructure as CodeStrong experience building and managing GitHub Actions CI/CD pipelinesSolid experience with monitoring and observability toolingStrong understanding of metrics, logging, tracing, alerting, and system healthExperience troubleshooting complex production environmentsUnderstanding of SLIs, SLOs, SLAs, and error budgetsExperience with incident management and root cause analysisGood understanding of cloud networking, security, and infrastructure fundamentalsStrong scripting/automation skillsA strong understanding of reliability, scalability, performance, and availabilityExcellent communication skills and the ability to work closely with software engineering teamsA proactive mindset and genuine passion for automation and continuous improvement Responsibilities: Design, build, and maintain highly available and scalable AWS infrastructureManage and optimise Kubernetes environments and containerised workloadsBuild and maintain infrastructure using Terraform and Infrastructure as Code principlesDevelop and optimise CI/CD pipelines using GitHub ActionsBuild and improve comprehensive monitoring and observability across the platformImplement and maintain effective logging, metrics, tracing, alerting, and dashboardsDefine and improve SLIs, SLOs, and reliability metricsProactively identify and resolve performance, availability, and reliability issuesLead and contribute to incident response, troubleshooting, and root cause analysisAutomate operational processes and eliminate repetitive manual tasksWork closely with software engineers to improve deployment processes, system reliability, and developer experienceHelp improve platform resilience, scalability, and disaster recovery capabilitiesContribute to capacity planning and performance optimisation as the platform scalesEstablish and champion SRE best practices across the wider engineering function Technologies: AWSCI/CDCloudDevOpsGitHubIncident ManagementKubernetesSecurityTerraform More:

We are partnering with a fast-growing SaaS company that is going through an exciting period of growth and investing heavily in its engineering and platform capabilities. We are offering a Site Reliability Engineer role in Central London on a hybrid basis, with 2/3 days a week in the office, and a salary of up to £85,000 plus bonus and benefits. This is a hands-on opportunity to work with a modern AWS and Kubernetes environment, take ownership of reliability, automation, and platform performance, and have genuine influence over engineering and platform decisions as the business continues to scale. We work closely with talented software engineering teams and offer clear opportunities to progress while helping shape and mature our SRE practices.

last updated 35 week of 2026

Ready to apply?

Install the ResuMinder extension and we'll auto-fill the application in seconds — no rewriting.

See how your CV scores