Now hiring

Site Reliability Engineer @ Biometric Talent

Chorley Old Road 108, ManchesterOnsiteFull-time
Apply with ResuMinder

Opens on the employer's site

About this role

Salary: £40,000 - 65,000 per year

Requirements: Strong software engineering experience, particularly with Python or GolangExperience with monitoring, alerting and observabilityKnowledge of OpenTelemetry and modern observability practicesExperience establishing proactive monitoring and alerting for complex platformsStrong understanding of SRE principles, including SLIs and SLOsExperience with modern software development practices and lifecyclesProficiency in shell scriptingExperience with Infrastructure as Code, automation and orchestration, ideally using Terraform and AnsibleExperience with tools such as Grafana, Splunk, New Relic and PagerDutyExperience working within large-scale, 24/7 enterprise environments where availability and stability are criticalStrong incident management, troubleshooting and root cause analysis experienceHands-on experience using LLM platforms and coding assistants to improve productivity and qualityExperience or interest in using AI for telemetry, predictive insights and root-cause analysis Responsibilities: Write and contribute to code that improves service reliability and observabilityDevelop tools, operational APIs and automation to improve system managementEstablish proactive monitoring and alerting across complex platformsImplement service instrumentation using OpenTelemetryBuild sophisticated dashboards using Grafana, Splunk and New RelicAutomate manual processes and reduce operational toilWork with Infrastructure as Code and orchestration technologiesSupport live incident resolution and contribute to post-mortem analysisCarry out root cause analysis and implement effective remediationDrive initiatives to improve system reliability, performance and observabilityMaintain and administer existing monitoring and analytics platformsWork with IT Operations to provide critical tooling and capabilitiesShare knowledge and mentor colleagues on new technologies and practicesUse AI tools, LLM platforms and coding assistants in day-to-day work to improve productivity, reduce toil and explore new approaches to autonomous operations, telemetry and system health Technologies: AIAnsibleGolangGrafanaIncident ManagementSupportLLMOpenTelemetryPagerDutyPythonSplunkTerraformJavaScript More:

We are supporting our client with the appointment of an experienced Site Reliability Engineer (SRE) to join their Platform Engineering team in Manchester, with 2 days per week onsite. This is a software-focused SRE role where software engineering skills are used to solve operational problems and improve reliability, observability and performance in a large-scale production environment. The role offers a salary of £40,000 to £65,000 per annum depending on experience and skillset, along with a performance-based bonus, pension scheme, hybrid working, flexible working hours, 25 days holiday plus your birthday off and bank holidays, the option to buy or sell up to 5 additional days, and free gym membership. We work closely with Development, Platform Delivery and IT Operations teams to build the tooling, automation, monitoring and practices that keep critical systems reliable and resilient, and we place a strong focus on AI-enabled engineering using AI tools, LLM platforms and coding assistants.

last updated 39 week of 2026

Ready to apply?

Install the ResuMinder extension and we'll auto-fill the application in seconds — no rewriting.

See how your CV scores