About this role
Salary: £61,000 - 101,000 per year
Requirements: Formal training or certification in site reliability engineering concepts, with proficient applied experienceProficiency in site reliability culture and principles, and the ability to implement site reliability practices within an application or platformProficiency in at least one programming language, such as Python, Java/Spring Boot, or .NETExperience with observability practices, such as white- and black-box monitoring, service level objective alerting, and telemetry collectionProficient knowledge of software applications and technical processes within a technical discipline, such as cloud, AI, or mobile platformsWorking knowledge of using enterprise-authorized AI capabilities to support site reliability engineering workflows, with strong validation habits and awareness of data sensitivityAbility to review and validate AI-assisted operational recommendations before applying changes, escalate when uncertain, and follow security and data-handling requirementsPreferred: Experience with continuous integration and continuous delivery toolingPreferred: Familiarity with container technologies and container orchestration platformsPreferred: Experience troubleshooting common networking technologies and issues Responsibilities: Guide and support team members in building appropriate-level designs, gaining peer consensus, and adopting site reliability engineering best practicesCollaborate with software engineers and cross-functional teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelinesImplement infrastructure, configuration, and network as code for applications and platforms within our scopePartner with technical experts, key stakeholders, and team members to resolve complex problems and proactively address issues using service level indicators and objectives before they affect customersIdentify and address roadblocks, propose improvements to solve business problems, and explore new technologies where appropriateApply availability, reliability, and scalability principles to iteratively improve outcomes in collaboration with partnersUse enterprise-authorized AI capabilities to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirementsUse enterprise-authorized AI capabilities to identify patterns in operational signals that indicate reliability risks or recurring toil, prioritizing reuse-first improvements tied to service level objective outcomesConfigure, maintain, monitor, and optimize applications and their associated infrastructureShare knowledge of end-to-end operations, availability, reliability, and scalability with the team Technologies: AICloudSupportJavaMarketingMobileNetworkPythonSecuritySpringSpring BootASP.NET More:
At JPMorganChase, were part of a rapidly growing technology field, applying our skills to drive innovation and modernize complex, mission-critical systems. As a Site Reliability Engineer II in Corporate Risk Technology, youll help solve broad business problems with straightforward solutions and contribute to the reliability and resilience of important platforms. We value curiosity, collaboration, and continuous improvement. J.P. Morgan is a global financial services company providing strategic advice and products to corporations, governments, wealthy individuals, and institutional investors. Our Corporate Functions professionals work across areas including finance, risk, human resources, and marketing, supporting our businesses, clients, customers, and employees. We value diversity and inclusion, are an equal opportunity employer, and provide reasonable accommodations for applicants and employees.
last updated 39 week of 2026