Apple

Site Reliability Engineer, Enterprise Technology Services

Apple Sunnyvale, California, United States

Computers and Electronics Manufacturing · 10,001+ employees

3 h ago
sre Senior (5-10 yrs) Full-time United States
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The Site Reliability Engineer will manage global-scale platforms, including identity, device security, and anti-abuse systems. They are responsible for ensuring the reliability and security of infrastructure that supports Apple's manufacturing, supply chain, and product launch operations.

What they look for

Site Reliability Engineering DevOps Java Python Bash LUA Oracle MongoDB Linux Networking Kubernetes Cloud Computing CI/CD Observability Incident Management Cryptography

Requirements

Candidates must have at least 5 years of experience in SRE, DevOps, or Software Engineering with proficiency in Java, Python, or Bash. A bachelor's or master's degree in Computer Science or a related field is required.

Full description

Imagine what we could do together. At Apple, new ideas have a way of becoming excellent products, services, and customer experiences very quickly. Bring passion and dedication to your job and there’s no telling what you could accomplish! The people here at Apple don’t just build products - they craft the kind of wonder that’s revolutionized entire industries. It’s the diversity of those people and their ideas that encourages the innovation that runs through everything we do, from amazing technology to industry-leading environmental efforts.

Enterprise Technology Services (ETS) is part of IS&T and delivers global-scale platforms and services that keep Apple's operations secure and running. The team manages identity, device security, and anti-abuse platforms — covering everything from manufacturing and repairs to software updates and activations. ETS also oversees supply chain, manufacturing, and partner integration platforms, protecting data on more than 2.5 billion devices worldwide. And when Apple prepares for a global product launch, ETS owns the systems that ramp factory production — managing serial numbers, network credentials, and verified software.

At ETS, we take pride in developing groundbreaking, world-changing platforms and services. Our ETS applications play a crucial role in supporting the Apple ecosystem by offering identity management, factory and device support, infrastructure support, platform support, and collaboration tools. Whether you're logging into Apple, making a purchase, or enabling Apple devices, our applications are there every step of the way, ensuring a flawless and secure experience.

Description

THIS ROLE IS DESIGNED FOR DRIVEN INDIVIDUALS WHO: Love learning new technologies and thrive in solving sophisticated challenges. Are independent, motivated, and excited to take on ambitious projects. Excel at collaborating with engineering teams and can stay calm under pressure. Have a passion for delivering quality, reliable solutions in a dynamic, high-energy workplace

Minimum Qualifications

5+ years of experience in Site Reliability Engineering, DevOps, Software Engineering, or a related field Strong foundation in programming language (Java) or scripting (Python / Bash / LUA) Hands on experience in one or more databases (Relational / NoSQL like Oracle, MongoDB) Education: Bachelor’s or Master’s degree in Computer Science or a related field (equivalent practical experience)

Preferred Qualifications

Hands on experience with monitoring and logging tools (e.g., Prometheus, Splunk, Grafana, CloudWatch) Proficient in Linux, Networking concepts (TLS/SSL, DNS, Load Balancers, etc..) and troubleshooting skills in large scale environments Source control management such as Git / Understanding of CI/CD, Release Engineering and DevOps Understanding of security standards, policies, and cryptography Experience with Incident / Problem management and RCA Strong Network, Load Balancing (Nginx, Envoy, NetScaler) experience is a huge plus Good solid understanding using Kubernetes concepts such as networking, Storage, Secrets, Deployments, Containers. AWS or GCP are preferred. Knowledge or experience in Governance and Compliance. Understanding of SRE principles, including observability, error budgeting, service reliability measurements through SLA & SLO & SLI, corresponding telemetry standards and practices, and product feedback. Strong analytical skills

Similar roles