Apple

Engineering Manager, Data Services, IS&T Ai & Data Platforms

Apple · Shanghai, Shanghai, China

Computers and Electronics Manufacturing · 10,001+ employees

8 h ago
Principal (10+ yrs) Full-time China
Log in to apply, save this posting, or score it against your profile with AI.

About the role

Lead the Data Services SRE organization to manage and scale critical database platforms across global data centers. Define technical strategy for engine-agnostic platform capabilities, including provisioning, observability, and data migration tooling.

What they look for

Site Reliability Engineering People Management PostgreSQL CockroachDB Couchbase OpenSearch MongoDB Oracle Apache Doris Python Go Infrastructure Operations Automation Cloud Architecture Data Migration Observability

Requirements

Requires 12+ years of experience in SRE or infrastructure roles with at least 4 years of people management. Candidates must have deep hands-on experience with major database engines and proficiency in Python or Go.

Full description

Apple's AI & Data Platform (AiDP) Data Services organization is seeking an experienced, versatile engineering leader to head our Data Services SRE organization. This leader will manage a group of engineering managers and technical leads responsible for cross-cutting software and platform tooling that manages a fleet of relational, distributed-SQL, document, search, and analytics engines powering some of Apple's most critical internet services, deployed at massive scale across data centers worldwide. In AiDP, your leadership will directly shape the reliability and scalability of platforms benefiting hundreds of millions of users, and will be critical to the success of some of the most visible current and future Apple features.

Description

The AiDP Data Services SRE organization builds engine-agnostic platform capabilities — provisioning, backup/restore, observability, and self-service — spanning PostgreSQL, CockroachDB, Couchbase, OpenSearch, MongoDB, Oracle, and Apache Doris. As the leader of this organization, you will set technical direction and organizational strategy for engine-agnostic data migration tooling across hybrid-cloud environments, ensuring minimal downtime, guaranteed data integrity, and adherence to local data residency. This role requires exceptional cross-functional leadership, executive-level communication, deep partnership with Core Storage and Analytics leadership, and the ability to build, mentor, and scale a distributed organization of managers and senior engineers.

Minimum Qualifications

BS or MS in Computer Science / related field or equivalent experience, with 12+ years in Site Reliability Engineering / Infrastructure roles, including 4+ years of people management experience. Proven track record building, scaling, and retaining high-performing SRE/infrastructure engineering organizations. Deep hands-on background (prior to or alongside management) with at least two of PostgreSQL, Oracle, MongoDB, CockroachDB, Couchbase, OpenSearch, or Apache Doris, supporting internet-facing production services via On Call and Incident Management. Demonstrated experience overseeing large-scale infrastructure operations with heavy reliance on automation tooling, across Datacenter and Cloud architectures (including Alibaba Cloud/Ali Cloud and/or AWS). Strong executive communication and technical writing skills; ability to translate deep technical context into leadership-level narratives and drive decisions across senior stakeholders. Working knowledge of one or more of the following programming languages: Python, Go — sufficient to engage credibly in technical design discussions and code/architecture reviews.

Preferred Qualifications

Experience sponsoring or overseeing the build of database-as-a-service control planes, provisioning APIs, or self-service tooling at an organizational level. Proven track record leading engine-agnostic, cross-cloud / hybrid-cloud data migration programs at scale. Experience defining organization-wide engine-agnostic observability (SLIs/SLOs), backup, and DR standards across a heterogeneous fleet. Experience leading teams operating real-time OLAP/analytics engines (e.g., Apache Doris, ClickHouse, StarRocks) at scale. Proficiency in Mandarin (spoken and written), to support coordination with local Ali Cloud teams and regional vendors/partners. Track record of sponsoring contributions to open-source database projects or internal data-platform tooling.