Apple

SRE, London

Apple London, England, United Kingdom

Computers and Electronics Manufacturing · 10,001+ employees

14 h ago
sre Senior (5-10 yrs) Full-time United Kingdom
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

You will be responsible for managing and scaling large-scale distributed infrastructure to support Apple's global services. This includes automating operations, debugging system performance, and collaborating with development teams to ensure high availability.

What they look for

Site Reliability Engineering Linux Distributed Systems Go Java Python Automation Networking Protocols Cloud Infrastructure System Configuration Management Software Deployment Monitoring Kubernetes Troubleshooting Capacity Planning Disaster Recovery

Requirements

Candidates must have significant experience as an SRE or infrastructure engineer with a strong understanding of Linux and distributed systems. Proficiency in programming languages like Go, Java, or Python and experience with cloud environments are required.

Full description

People at Apple don’t just build products — they craft the kind of experience that have revolutionized entire industries. The diverse collection of our people and their ideas inspire innovation in everything we do. Imagine what you could do here! Join Apple, and help us leave the world better than we found it.

The Apple Services Engineering (ASE) team builds and provides systems and infrastructure that fuel Apple’s services (such as iCloud, iTunes, Siri, and Maps). We are the foundation on which Apple’s software developers build the products that our customers love. We are looking for passionate and talented Site Reliability Engineers to continue our focus in providing our customers the highest quality Apple Services experience. Our services have to scale globally, stay highly available, and "just work.” If you love designing, engineering and running systems and infrastructure that will help millions of customers, then this is the place for you!

Description

FoundationDB infrastructure is BIG. Operating at our scale, across multiple geographically dispersed data centers and servicing hundreds of millions of users presents unique challenges. As an SRE at Apple, you'll need to solve these problems using data, teamwork, and your own expertise. SREs at Apple own the full infrastructure stack; from device driver performance debugging to content delivery network traffic management — our responsibilities are both broad and deep.

FoundationDB runs its systems on Linux. We run a mix of open source, vendor licensed, and internally developed tools to perform functions such as system configuration management, provisioning, software deployment, logging, and monitoring. You'll learn these tools and have opportunities to improve them. Our team is collaborative; we work closely with the development teams we support to deliver the best results for Apple. We think critically and strive to balance the best solution with the need to get things done for each engineering challenge we face. Good ideas are heard and results are rewarded.

FoundationDB SRE is a small team with huge scale. We serve as the database for much of CloudKit's use cases, including Mail, Contacts, and Keychain. We serve hundreds of millions of customers every day and are a fundamental piece of the Apple device experience.

Minimum Qualifications

* Strong sense of ownership and integrity demonstrated through clear communication and collaboration * Experience in managing and scaling distributed systems in a public, private, or hybrid cloud environment The ability to design, author, and release code in languages like (but not limited to) Go, Java or Python Acute drive to automate manual operations and to improve them through repeated iteration Understanding of the Linux Operating System, standard networking protocols, and components Experience with deploying, supporting and monitoring new and existing services, platforms, and application stacks Experience with scale testing, disaster recovery, and capacity planning In depth experience as a SRE, PE, Infra software engineer or equivalent role

Preferred Qualifications

Hands-on experience managing large numbers of diverse systems with configuration management or software delivery platforms (such as Puppet, Chef, Ansible, and Spinnaker) Excellent troubleshooting and problem solving skills Experience with scale testing, disaster recovery, and capacity planning Familiarity with microservices architecture and container orchestration with Kubernetes

Similar roles