Senior Site Reliability Engineer - iCloud / Apple Services Engineering
Apple Seattle, Washington, United States
Computers and Electronics Manufacturing · 10,001+ employees
About the role
Responsible for the reliability and performance of the server software stack powering Apple services like iCloud and Mail. You will leverage data and automation to manage the end-to-end SDLC while collaborating with product development teams.
What they look for
Requirements
Requires 5+ years of experience in Infrastructure Ops, SRE, or DevOps and a BS degree in computer science or equivalent. Candidates must demonstrate fluency in Java, Python, or Go and possess strong knowledge of Linux, networking, and distributed systems.
Full description
People at Apple don’t just build products — they craft experiences our customers love and depend on. Apple Services Engineering (ASE) builds and supports the systems that make many of these daily experiences possible. If you’ve used Apple products, you’ve likely interacted with us. Apple Services Site Reliability Engineering (SRE) teams are responsible for the systems and services that directly support those customers and their experiences. We are looking for an SRE with experience in building and supporting highly available customer-facing services.
Description
Apple Services’ scale is BIG. Operating at our scale, across multiple geographies and servicing hundreds of millions of users presents unique challenges. As a Software Developer in SRE at Apple, you'll need to solve these problems using data, teamwork, and your own expertise. ASE Products Site Reliability teams are responsible for the reliability and performance of the server software stack that powers products like iCloud Photos, Mail, Drive, Backup and many more. We do that by focusing on reliability best practices from service inception to production, collaborating deeply with product development teams to deliver a superlative product and shared vision while leveraging data and automation as first principles. We run a mix of open source, vendor licensed, and internally developed tools to manage the end to end SDLC of our products. You'll learn these tools and have opportunities to improve them.
Minimum Qualifications
5+ years in a Infrastructure Ops, Site Reliability Engineering, or DevOps focused role. BS degree in computer science or equivalent field with 5+ years of experience. Knowledge of Linux operating system principles, networking fundamentals, and systems management. Demonstrable fluency in at least one of the following languages: Java, Python, or Go. Experience in managing and scaling distributed systems in a public, private, or hybrid cloud environment. Familiarity with micro-services architecture and container orchestration with Kubernetes. Awareness of key security principles including encryption, keys (types and exchange protocols). Understanding of SRE principals including monitoring, alerting, error budgets, fault analysis, and automation. Strong sense of ownership, with a desire to communicate and collaborate with other engineers and teams. Ability to identify and communicate technical and architectural problems, while working with partners and their team to iteratively find solutions.
Preferred Qualifications
Experience implementing automation Scripting experience in Python Enjoy building partnerships
Similar roles
-
Technical Operations Engineer (SRE) for an Real Estate Company (US-Based/Remote)
Paired Argentina
-
Staff Site Reliability Engineer
IonQ Santa Clara, California, United States · $152K–$228K/yr
-
Senior Site Reliability Engineer (SRE) – CloudVision as a Service (CVaaS)
Arista Networks Vancouver, British Columbia, Canada · $95K–$145K/yr
-
FedRAMP Site Reliability Engineer (FedSRE) - CloudVision
Arista Networks United States · $101K–$161K/yr
-
Cloud Site Reliability Engineer (SRE)
ECS Tech Inc Arlington, Virginia, United States · $130K–$180K/yr
-
Staff Software Engineer - SRE & AIOps
ServiceNow Atlanta, Georgia, United States