Now hiring

Lead SRE - Azure & GCP @ Hackajob Ltd

Appold Street 9, GlasgowOnsiteFull-time
Apply with ResuMinder

Opens on the employer's site

About this role

Salary: £61,000 - 101,000 per year

Requirements: We have Google and Azure cloud expertise in a mission-critical production environment.We have a strong understanding of container technologies such as Docker, Kubernetes, GKE, and Helm.We have programming experience in Python, shell scripting, or Go, along with a good understanding of REST APIs.We have hands-on experience with cloud-based technologies and tools for deployment, monitoring, and operations, such as Google Observability, Azure Monitor, Datadog, Prometheus, Splunk, Elasticsearch, and Grafana.We have demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows, with strong validation habits and awareness of data sensitivity.We can evaluate AI-assisted operational recommendations for correctness and risk, define guardrails for team usage, and ensure outcomes align with resiliency and security expectations.We have a strong understanding of Google Cloud governance, compliance, and cost management.We have working knowledge of modern development technologies and tools such as Agile, CI/CD, Git, Infrastructure as Code, Terraform, and Jenkins.We have Google Cloud certification or equivalent technical experience in the public cloud.We have a good understanding of Agentic AI SDKs and GitHub Copilot skills.We have a good understanding of operating systems such as Windows and Linux (Red Hat/Ubuntu).We have a good understanding of LLM and other AI/ML frameworks that can be used in AIOps. Responsibilities: We lead and implement SRE frameworks to support global Google Cloud environments and ensure the highest level of SLOs through operational excellence.We apply mastery of application, data, infrastructure, and Agentic AI disciplines.We use enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage, troubleshooting, and post-incident analysis, while validating outputs and handling operational data according to sensitivity and security requirements.We provide support to develop and improve the quality of technical engineering documentation.We provide technical supervision, oversight, and problem resolution for engineering activities.We champion a DevOps model so services are automated and elastic across all platforms. Technologies: Agentic AIAIAzureCI/CDCloudCopilotDatadogDevOpsDockerElasticSearchGitGitHubGrafanaHelmSupportJenkinsKubernetesLLMLinuxPrometheusPythonRESTSecuritySplunkTerraformUbuntuWindowsNetwork More:

We are J.P. Morgan, a global leader in financial services, and we are partnering with hackajob to hire a Lead Site Reliability Engineer for our Google Cloud Site Reliability Engineering team within the Infrastructure Platform - Cloud Foundational Services SRE organization. You will join our global follow-the-sun support model and contribute to our Corporate Technology team, supporting corporate functions across Global Finance, Corporate Treasury, Risk Management, Human Resources, Compliance, Legal, and the Corporate Administrative Office. We value diversity and inclusion, and we offer a collaborative environment focused on trusted long-term partnerships, technology innovation, and our technology controls agenda.

last updated 37 week of 2026

Ready to apply?

Install the ResuMinder extension and we'll auto-fill the application in seconds — no rewriting.

See how your CV scores