Site Reliability Engineer, Apple Ads
Apple Cupertino, California, United States
Computers and Electronics Manufacturing · 10,001+ employees
About the role
The Site Reliability Engineer will maintain the uptime, scalability, and stability of mission-critical ad-tech systems. They will collaborate with developers and architects to implement infrastructure improvements and ensure system reliability.
What they look for
Requirements
Candidates must have at least 3 years of experience with internet-facing production systems and distributed cloud infrastructure. Proficiency in Python, Go, or Java, along with hands-on experience in Linux, AWS, and Infrastructure as Code (Terraform), is required.
Full description
At Apple, we focus deeply on our customers’ experience. Apple Ads brings this same approach to advertising, helping people find exactly what they’re looking for and helping advertisers grow their businesses.
Our technology powers ads and sponsorships across Apple Services, including the App Store, Maps, Apple News, and MLS Season Pass. Everything we do is designed for trust, connection, and impact: We respect user privacy, integrate advertising thoughtfully into the experience, and deliver value for advertisers of all sizes—from small app developers to big, global brands. Because when advertising is done right, it benefits everyone.
Description
As a Site Reliability Engineer, you will be responsible for providing the platform for mission-critical ad-tech systems to maintain constant uptime, scale seamlessly, and allow for new applications and services to flourish.
The successful candidate will be highly self-motivated and passionate about excellence, quality, and detail. The SRE will not only support operations but also work closely with the developers and architects within the team to aid in the design and assist with the implementation to improve stability, security, and scalability.
Minimum Qualifications
3+ years of experience supporting internet-facing production systems and distributed cloud infrastructure. Strong programming skills in at least one of: Python, Go, or Java. Proven expertise with AWS-managed infrastructure Hands-on experience with Linux systems and deep knowledge of its internals. Demonstrated experience with Infrastructure as Code, especially Terraform. Strong foundation in SRE concepts: Monitoring, alerting, and observability, incident response and root cause analysis, error budgets, SLAs/SLOs, and system reliability
Preferred Qualifications
Demonstrated experience designing, building, or integrating AI/LLM-powered tooling and automations Passion for customer privacy Built tools or services that automate platform operations, reduce toil, or improve cost efficiency. Experience managing Kubernetes clusters at scale in production environments. Hands-on experience troubleshooting distributed systems under real-world load. Experience building and operating infrastructure at scale. Experience building solutions that reduces friction in software delivery. Clear communication skills and comfort collaborating across engineering, infrastructure, and product teams. AWS certifications or broad experience across multiple AWS services is a plus.
Similar roles
-
Manager, Site Reliability Engineering
LayerZero Labs Vancouver, British Columbia, Canada
-
Lead Site Reliability Engineer - Network
JPMorgan Chase & Co. Columbus, Ohio, United States
-
Site Reliability Engineer
Klaviyo GMI Boston, Massachusetts, United States · $138K–$150K/yr
-
SRE | Site Reliability Engineering
C6 Bank São Paulo, São Paulo, Brazil
-
Lead Site Reliability Engineer
JPMorgan Chase & Co. Seattle, Washington, United States · $157K–$215K/yr
-
Engineering Manager, SRE
Remote São José da Laje, Alagoas, Brazil · $75K–$170K/yr