Site Reliability Engineer II - AI & Corporate Risk Tech
JPMorgan Chase & Co. Glasgow, Scotland, United Kingdom
Financial Services · 10,001+ employees
About the role
The Site Reliability Engineer will configure, monitor, and optimize mission-critical systems using cloud infrastructure and automated pipelines. They will also collaborate with cross-functional teams to implement reliability best practices and resolve complex operational issues.
What they look for
Requirements
Candidates must have formal training or certification in site reliability engineering and proficiency in at least one programming language like Python, Java, or .NET. Experience with observability practices, CI/CD tooling, and container technologies is required to support platform reliability.
Full description
There's nothing more exciting than being at the center of a rapidly growing field in technology and applying your skills to drive innovation and modernize some of the world's most complex and mission-critical systems. At JPMorganChase, you'll be part of a team that values curiosity, collaboration, and continuous improvement — where your contributions directly shape the reliability and resilience of platforms that matter.
As a Site Reliability Engineer II at JPMorganChase within Corporate Risk Technology, you will solve complex and broad business problems with simple, straightforward solutions. Through code and cloud infrastructure, you will configure, maintain, monitor, and optimize applications and their associated infrastructure — independently decomposing and iteratively improving on existing solutions. You are a meaningful contributor to your team, sharing your knowledge of end-to-end operations, availability, reliability, and scalability of your application or platform.
Job responsibilities
- Guide and support team members in building appropriate-level designs, gaining peer consensus, and driving adoption of site reliability engineering best practices across the team
- Collaborate with software engineers and cross-functional teams to design, develop, test, and implement deployment and reliability approaches using automated continuous integration and continuous delivery pipelines
- Implement infrastructure, configuration, and network as code for the applications and platforms within your scope
- Partner with technical experts, key stakeholders, and team members to resolve complex problems and proactively address issues using service level indicators and objectives before they impact customers
- Identify and address roadblocks, propose improvements to solve business problems, and explore new technologies where appropriate
- Apply familiarity with availability, reliability, and scalability principles to iteratively improve outcomes in collaboration with partners
- Uses enterprise-authorized AI capabilities within the work environment to accelerate incident triage, troubleshooting, and post-incident analysis, validating outputs and handling operational data according to sensitivity and security requirements
- Applies enterprise-authorized AI capabilities within the work environment to identify patterns in operational signals that indicate reliability risk or recurring toil, prioritizing reuse-first improvements tied to service level objective outcomes
Required qualifications, capabilities, and skills
- Formal training or certification on site reliability engineering concepts and proficient applied experience
- Proficiency in site reliability culture and principles, with the ability to implement site reliability practices within an application or platform
- Proficiency in at least one programming language such as Python, Java/Spring Boot, or .NET
- Experience in observability practices such as white and black box monitoring, service level objective alerting, and telemetry collection
- Proficient knowledge of software applications and technical processes within a given technical discipline (e.g., cloud, AI, mobile platforms)
- Working knowledge of using enterprise-authorized AI capabilities within the work environment to support site reliability engineering workflows, with strong validation habits and awareness of data sensitivity
- Ability to review and validate AI-assisted operational recommendations before applying changes, escalating when uncertain and following security and data handling requirements
Preferred qualifications, capabilities, and skills
- Experience with continuous integration and continuous delivery tooling
- Familiarity with container technologies and container orchestration platforms
- Experience troubleshooting common networking technologies and issues
- There’s nothing more exciting than being at the center of a rapidly growing field in technology and applying your skillsets to drive innovation and modernize the world's most complex and mission-critical systems.
J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.
We recognize that our people are our strength and the diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin, gender, sexual orientation, gender identity, gender expression, age, marital or veteran status, pregnancy or disability, or any other basis protected under applicable law. We also make reasonable accommodations for applicants’ and employees’ religious practices and beliefs, as well as mental health or physical disability needs. Visit our FAQs for more information about requesting an accommodation.
Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we’re setting our businesses, clients, customers and employees up for success.
Similar roles
-
Senior Site Reliability Engineer
Zocdoc United States · $180K–$220K/yr
-
Sr. Site Reliability Engineer
CentralReach Holmdel Township, New Jersey, United States · $160K–$180K/yr
-
Senior AWS Site Reliability Engineer (REMOTE)
VIATEQ Corporation Fairfax County, Virginia, United States · $145K–$185K/yr
-
Senior Site Reliability and DevOps Engineer
Capco Toronto, Ontario, Canada · CA$118K–CA$152K/yr
-
Site Reliability and DevOps Engineer
Capco Toronto, Ontario, Canada · CA$92K–CA$118K/yr
-
(1112149) Site Reliability Engineer (SRE) - Application Support
Diversified Services Network, Inc. Peoria, Illinois, United States · $95K–$100K/yr