American Express

Manager Technology Ops Engineering

American Express · Mid Sussex, England, United Kingdom

Financial Services · 10,001+ employees

4 h ago Closes in 7d
Senior (5-10 yrs) Full-time Visa sponsorship United Kingdom
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The manager leads and mentors Site Reliability Engineering teams to foster a culture of continuous improvement and system resilience. They are responsible for overseeing 24x7 support, incident management, and collaborating with engineering teams to ensure high-quality performance of enterprise platforms.

What they look for

Site Reliability Engineering Splunk Prometheus Grafana Kubernetes Docker AWS Azure Google Cloud Java Go Git Agile Incident Management Networking Microservices

Requirements

Candidates must have a bachelor's degree in computer science or a related field and significant experience in SRE functions and application support. Proficiency in observability tools, cloud platforms, and containerization technologies is required, along with domain knowledge of card payment systems.

Full description

The Enterprise Technology Services organization partners with every part of the American Express business to power the company’s growth and innovation with trust and efficiency, and drive competitive differentiation with speed. We support the delivery and operations of technology, digital, and data capabilities, platforms, and services globally. Specifically, our team is responsible for the company’s technology engineering, architecture, and infrastructure, providing 24x7 support to ensure an uninterrupted, high-quality experience for customers and colleagues. We also provide product management for core enterprise platforms, and lead technology risk and information security, enterprise data governance and platforms, digital product and design, and enterprise AI platforms on behalf of the company.

Manager, Site Reliability Engineering leads and mentors Site Reliability Engineering (SRE) teams, fostering a culture of continuous improvement and inclusivity, while collaborating across the organization to enhance system resilience, scalability, and alignment with business objectives.