Now hiring

Site Reliability Engineer (AWS) @ Spectrum IT Recruitment

Enterprise Road, SouthamptonOnsiteFull-time
Apply with ResuMinder

Opens on the employer's site

About this role

Salary: £60,000 - 60,000 per year

Requirements: We ideally want experience in a Site Reliability Engineering, Production Engineering, Cloud Operations or NOC environment.We are looking for exposure to Linux systems administration.We want experience with AWS cloud infrastructure.We are looking for experience with Kubernetes and Docker.We want experience in production support and incident management.We are looking for scripting experience in Python, Bash or Go.We want familiarity with monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch.We are looking for networking fundamentals including DNS, TCP/IP and load balancing.We value a passion for automation, continuous improvement and operational excellence.Experience with Infrastructure as Code such as Terraform, SRE principles such as SLIs and SLOs, or regulated environments would be beneficial but is not essential. Responsibilities: We monitor and maintain highly available production platforms running in AWS.We respond to and manage production incidents across a 24/7 service.We investigate complex technical issues and restore services quickly and effectively.We develop automation to reduce manual operational tasks and improve platform resilience.We build and improve monitoring, alerting and observability across cloud environments.We work alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence.We contribute to post-incident reviews and drive continuous service improvements.We support containerised workloads using Kubernetes and Docker. Technologies: AIAWSBashCloudCloudWatchDatadogDockerGrafanaIncident ManagementSupportKubernetesLinuxLoad BalancingPrometheusPythonSecuritySplunkTCP/IPTerraformDevOps More:

We are a global leader in AI-powered customer experience and cloud technology, expanding our engineering teams following the award of a major government programme. We are building and supporting highly secure, cloud-native platforms that deliver sensitive communication services. This is a fully remote role in the UK on a 24/7 shift pattern across a 28-day rota including days and nights, with a competitive salary, bonus and excellent benefits. You will join an engineering-led organisation where reliability, automation and continuous improvement are central to the platform, and you will work as part of a collaborative SRE team focused on resilient cloud services.

last updated 31 week of 2026

Ready to apply?

Install the ResuMinder extension and we'll auto-fill the application in seconds — no rewriting.

See how your CV scores