Senior Site Reliability Engineer
Apple Cupertino, California, United States
Computers and Electronics Manufacturing · 10,001+ employees
About the role
Design, build, and operate reliable, high-performance services and infrastructure for media ingestion, processing, and management across Apple services. Develop automation and tooling, support observability and monitoring, and help transition services from bare-metal systems to cloud infrastructure and Kubernetes-based microservices.
What they look for
Requirements
Requires a bachelor's degree or equivalent with at least six years of experience, or a master's degree or equivalent with at least four years of experience, plus at least six years in reliability engineering, DevOps, or infrastructure-focused work. Candidates need advanced programming experience in Golang, Python, Java, or C++, strong systems and infrastructure knowledge, and excellent troubleshooting and problem-solving skills; Kubernetes, microservices, automation, and service operations experience are preferred.
Full description
The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple Music, Apple Podcasts, and Apple Books. And they do it on a massive scale, meeting Apple’s high expectations with high performance to deliver a huge variety of entertainment in over 35 languages to more than 150 countries.
These engineers build secure, end-to-end solutions. They develop the custom software used to process all the creative work, the tools that providers use to deliver that media, all the server-side systems, and the APIs for many Apple services.
Thanks to Apple’s unique integration of hardware, software, and services, engineers here partner to get behind a single unified vision. That vision always includes a deep commitment to strengthening Apple’s privacy policy, one of Apple’s core values. Although services are a bigger part of Apple’s business than ever before, these teams remain small, forward-thinking, and cross-functional, offering greater exposure to the array of opportunities here.
Description
Media Platforms SRE is responsible for designing, building and running a diverse set of services that ingest, transform and manage all media content across products like the App Store, Apple Music, Apple Fitness+ and, Apple TV+. Whether it’s the latest episode of the Morning Show, a new album from your favorite artist or a brand new iOS app destined for the App Store. You will make sure our systems are highly available, fast and efficient. You’ll interact with a diverse range of system types -- public facing web & API developer services like App Store Connect & TestFlight, Media Processing pipelines that enables audio & video transcoding at scale as well as internal business tools/systems and machine learning models. Culturally we believe in a close partnership with our development teams and aim to design & build new services together. We're passionate about software and automation in SRE and develop a variety of tooling and infrastructure. Our services run on mixed platforms - many run on our bare metal fleet which you will help transition to modern Cloud infrastructure and Kubernetes micro-services and work with state of the art observability tools and monitoring platforms.
Minimum Qualifications
BS degree in computer science or equivalent field with 6+ years experience or MS degree in computer science or equivalent field with 4+ years experience At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role Advanced experience with programming languages (Golang, Python, Java, C++) Deep systems and infrastructure knowledge Excellent troubleshooting and problem solving skills
Preferred Qualifications
Passion for designing and building reliable systems Familiarity with microservices architecture and container orchestration with Kubernetes Demonstrated ability to deliver results on time with high quality Automation advocate - you truly believe in removing operation load with software Experience with deploying, supporting and monitoring new and existing services, platforms, and application stacks
Similar roles
-
Principal Site Reliability Engineer, Platform Engineering: Dedicated
GitLab Canada · $223K–$380K/yr
-
Senior Site Reliability Engineer (SRE)
UJET United States · $140K–$180K/yr
-
Site Reliability Engineer (High Performance Computing)
SpaceX Hawthorne, California, United States · $125K–$195K/yr
-
Site Reliability Engineer
N26 Barcelona, Catalonia, Spain
-
SRE
Hitachi Solutions pune, Maharashtra, India
-
Principal SRE
Azira Bengaluru, Karnataka, India