About this role
J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.
We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.
As a Software Engineer, Site Reliability Engineer Support within Private Banking Technology at JPMorganChase, you will collaborate with stakeholders to drive the adoption of Site Reliability Engineering (SRE) tools, practices, and culture. You will partner with Product teams, other Line of Business (LOB) SMEs, and leadership to define and deliver SRE objectives. You will lead programs and initiatives that enable Product teams to define non-functional requirements (NFRs) and availability targets for their services. You will ensure these NFRs are incorporated into product design and testing, and that firmwide SRE practices are integrated into Product teams' SDLC lifecycles. Job responsibilities
• Demonstrate and promote site reliability principles and practices within the application-aligned domain. • Manage and troubleshoot production incidents as an active member of an application-aligned SRE team; participate in blameless post-mortems. • Provide flexibility with occasional early or late coverage for APAC businesses. • May occasionally work outside regular office hours (e.g., weekends) on an as-needed basis. • Establish and follow processes for visualizations, including SLO reviews, error budget tracking, and observability operational models. • Evolve and debug critical components of applications and platforms, developing expertise in your remit and understanding interdependencies and limitations. • Drive automation initiatives and deliver on non-functional requirements, leveraging common programming languages and CI/CD tooling to automate operational processes and implement end-to-end (E2E) pipelines (including system start-of-day checks where applicable). • Plan and conduct chaos engineering testing for private and public cloud applications, with a strong understanding of resiliency patterns. • Champion SRE culture throughout the organization via programs, initiatives, and innovation. • Collaborate with all IPB product SRE teams to build consistent visualization and application performance monitoring. • Stay current and be comfortable adopting AI-enabled technologies and approaches (including MCP) to improve reliability, operations, and engineering workflows. Required qualifications, capabilities, and skills
• Bachelor's degree in Computer Science, Mathematics or a related field • Formal Training and certification on software engineering and 3+ years of applied experience • Proven ability to engineer toil reduction and automation solutions using common programming languages (e.g., Python, Java). • Excellent communication and user-facing skills, with a positive, collaborative attitude. • Hands-on exposure to private and public cloud computing environments (e.g., AWS, Private Cloud). • Experience troubleshooting issues with common technologies (e.g., MSSQL, MongoDB, Oracle, MQ, Kafka). • Strong knowledge of site reliability culture and principles, with demonstrated ability to implement SRE practices within an application or platform. • Firsthand experience in observability, including white and black box monitoring, service level objectives, alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc. • Hands-on experience with CI/CD tools (e.g., Jules, Jenkins, GitHub, Terraform). • Practical experience with chaos engineering, resiliency testing, and a self-motivated approach to learning and evaluating modern technologies.
Preferred qualifications, capabilities, and skills
• Experience supporting trading platforms and front-to-middle office environments; exposure to FIX is preferred but not mandatory. • Exposure to Temenos Transact is an added advantage.