Now hiring

Lead Site Reliability Engineer @ Hackajob Ltd

Appold Street 9, GlasgowOnsiteFull-time
Apply with ResuMinder

Opens on the employer's site

About this role

Salary: £100,000 - 100,000 per year

Requirements: Formal training or certification on site reliability engineering concepts and advanced applied experienceExperience designing and coding complex problems in public cloud environments such as AWSDeep proficiency in reliability, scalability, performance, security, enterprise system architecture, toil reduction, and other site reliability best practicesFluency in Python and deep knowledge of software applications and technical processesProficiency in observability, including white-box and black-box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, and SplunkProficiency and experience with continuous integration and continuous delivery tools such as Jenkins, GitLab, and TerraformExperience with containers and container orchestration such as ECS, Kubernetes, and DockerExperience troubleshooting common networking technologies and issuesAbility to identify and solve problems related to complex data structures and algorithmsDrive to self-educate, evaluate new technology, teach new programming languages to team members, and collaborate across different levels and stakeholder groupsDemonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflowsAbility to evaluate AI-assisted operational recommendations for correctness and risk, define appropriate guardrails, and ensure outcomes align to resiliency and security expectations Responsibilities: Demonstrate and champion site reliability culture and practices and exert technical influence throughout our teamLead initiatives to improve the reliability and stability of our teams applications and platforms using data-driven analytics to improve service levelsCollaborate with team members to identify comprehensive service level indicators and establish reasonable service level objectives and error budgets with customersDemonstrate a high level of technical expertise within one or more technical domains and proactively identify and solve technology-related bottlenecksAct as the main point of contact during major incidents for our application and identify and solve issues quickly to avoid financial lossesDocument and share knowledge within our organization via internal forums and communities of practiceUse enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirementsLead reuse-first adoption of AI-assisted reliability workflows across SDLC and toolchain practices, ensuring traceability, auditability, resiliency, and security controlsTake lead on resiliency design reviewsBreak up complex problems into digestible work for other engineersAct as a technical lead for medium to large-sized productsProvide advice and mentoring to other engineers Technologies: AIAWSCloudDatadogDockerDynatraceGitLabGrafanaSupportJenkinsKubernetesMarketingPrometheusPythonSecuritySplunkTerraformCI/CD More:

hackajob is partnering directly with JPMorganChase to hire for this role. We are seeking a Lead Site Reliability Engineer within our Infrastructure Platforms team to help define the future of a globally recognized firm. This role places you in a leadership position where you will advise on technical and business issues, support major incident response, and help drive reliability across our applications and platforms. We are a global leader in financial services, providing strategic advice and products to prominent corporations, governments, wealthy individuals, and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do, and we build trusted, long-term partnerships to help clients achieve their business objectives. We value the diverse talents our people bring to our global workforce and are committed to diversity, inclusion, and reasonable accommodations. Our professionals in Corporate Functions support finance, risk, human resources, marketing, and other essential areas that help set our businesses, clients, customers, and employees up for success.

last updated 35 week of 2026

Ready to apply?

Install the ResuMinder extension and we'll auto-fill the application in seconds — no rewriting.

See how your CV scores