About this role
Salary: £40,000 - 70,000 per year
Requirements: Previous experience as a Site Reliability Engineer or similar (cloud, infrastructure, DevOps or platform engineering) within a customer-facing environment, e.g. SaaS or MSPStrong hands-on experience with both Azure and on-premise virtual machinesExperience with Infrastructure as Code / TerraformExperience with container orchestration such as Kubernetes or AKSExperience with monitoring and observability tooling such as Prometheus, Grafana, Datadog, or Azure MonitorAbility to implement new processes and tools, ensuring wider development and support teams adopt new ways of workingAbility to communicate across internal teams and external customersSkilled in networking across both cloud (Azure) and on-premise environmentsWindows or Linux systems administration skills, with Linux preferredPrevious use of AI tools to enhance efficiency Responsibilities: Utilise technologies such as Terraform and Kubernetes to manage, provision, and configure servers and networks, and automate application lifecyclesRegularly use Datadog and other observability tools for application performance monitoringImplement new ways of working, helping to shape how we respond to and recover from incidentsTake ownership of incident resolutionsActively drive down key reliability metrics such as MTTR, incident frequency, and on-call toil by evaluating key incidentsWork on an on-call rota, ensuring we are available to respond to incidents during this timeIdentify areas for automation and help implement changes that raise the bar for reliability Technologies: AIAzureCloudDatadogDevOpsGrafanaSupportKubernetesLinuxPrometheusTerraformWindows More:
We are a well-established B2B SaaS company going through a significant platform transformation, and we are hiring a Senior Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep. This role will also play a key part in implementing and adopting new tooling and processes across the wider transformation. It is a genuine opportunity to own and operate how our cloud function works, with the potential to progress into a team lead position as our team grows. The role is based in Milton Keynes, with two days on site each week, and offers a salary of up to 75,000 plus bonus and on-call allowance.
last updated 40 week of 2026