About this role
Salary: £60,000 - 65,000 per year
Requirements: Experience in a Site Reliability Engineering, Production Engineering, Cloud Operations, or NOC environmentExposure to Linux systems administrationExposure to AWS cloud infrastructureExposure to Kubernetes and DockerExposure to production support and incident managementExposure to Python, Bash, or Go scriptingExposure to monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk, or CloudWatchExposure to networking fundamentals including DNS, TCP/IP, and load balancingA passion for automation, continuous improvement, and operational excellenceExperience with Infrastructure as Code such as Terraform would be beneficialExperience with SRE principles such as SLIs and SLOs would be beneficialExperience in regulated environments would be beneficial Responsibilities: Monitor and maintain highly available production platforms running in AWSRespond to and manage production incidents across a 24/7 serviceInvestigate complex technical issues and restore services quickly and effectivelyDevelop automation to reduce manual operational tasks and improve platform resilienceBuild and improve monitoring, alerting, and observability across cloud environmentsWork alongside Software, Platform, Cloud, and Security Engineers to improve reliability and operational excellenceContribute to post-incident reviews and drive continuous service improvementsSupport containerised workloads using Kubernetes and Docker Technologies: AIAWSBashCloudCloudWatchDatadogDockerGrafanaSupportKubernetesLinuxLoad BalancingPrometheusPythonSecuritySplunkTCP/IPTerraformDevOps More:
We are a global leader in AI-powered customer experience and cloud technology, and we are expanding our engineering teams following the award of a major government programme. We are building and supporting highly secure, cloud-native platforms that deliver sensitive communication services. This is a fully remote role in the UK on a 24/7 shift pattern with a 28-day rota including days and nights. We offer a competitive salary, bonus, and excellent benefits, and you will join an engineering-led organisation where reliability, automation, and continuous improvement are central to our platform.
last updated 29 week of 2026