Apple

Software Engineer, Data Services, IS&T Ai & Data Platforms

Apple Shanghai, Shanghai, China

Computers and Electronics Manufacturing · 10,001+ employees

Aug 08
Mid (2-5 yrs) Full-time China
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The engineer will manage and maintain a large-scale transactional database fleet, including provisioning, backup, and observability tasks. They will also support complex data migration efforts across hybrid-cloud environments while ensuring data integrity and regulatory compliance.

What they look for

Site Reliability Engineering Oracle PostgreSQL MongoDB CockroachDB Python Go Kubernetes Data Migration Infrastructure Automation Observability Incident Management Cloud Architecture Distributed Systems Database Administration Troubleshooting

Requirements

Candidates must have 3-5 years of experience in Site Reliability Engineering or infrastructure roles with hands-on production experience in at least one major database engine. Proficiency in Python or Go and experience with automation tooling in cloud or datacenter environments are required.

Full description

AI & Data Platforms (AiDP) is IS&T's engine for AI-powered innovation. The team brings together data, application development, and machine learning — including generative AI — along with data services and customer success functions, to help IS&T build solutions more efficiently and streamline the adoption and embedding of generative AI across Apple.

Apple's AI & Data Platform (AiDP) Data Services organization is seeking a motivated database systems engineer to join our Data Services SRE team, focused on our transactional database fleet. Engineers on this team develop and contribute to platform tooling that manages relational and distributed-SQL/document engines powering some of Apple's most critical internet services at massive scale. In AiDP, your work will benefit hundreds of millions of users.

Description

The AiDP Data Services SRE team builds engine-agnostic platform capabilities — provisioning, backup/restore, observability, and self-service — spanning Oracle, PostgreSQL, MongoDB, and CockroachDB. This role involves supporting data migration efforts across hybrid-cloud environments with minimal downtime and guaranteed data integrity, following established runbooks and adhering to local data residency and regulatory requirements. This role requires good communication and collaboration with Core Storage teams and colleagues across a distributed team.

Minimum Qualifications

BS or MS in Computer Science / related fields or equivalent work experience, with 3–5 years in a Site Reliability Engineering / Infrastructure focused role. Hands-on production experience with at least one of Oracle, PostgreSQL, MongoDB, or CockroachDB, including exposure to On Call and Incident Management. Exposure to running infrastructure with reliance on automation tooling, across Datacenter and Cloud architectures (including Alibaba Cloud/Ali Cloud and/or AWS). Solid troubleshooting skills, a resourceful first-principles approach to problem solving, and good technical writing habits. Good understanding in one or more of the following programming languages: Python, Go.

Preferred Qualifications

Some operational experience with stateful services on Kubernetes (operators, StatefulSets, CSI storage). Exposure to database-as-a-service control planes, provisioning APIs, or self-service tooling. Experience assisting with cross-cloud / hybrid-cloud data migrations for transactional systems. Familiarity with observability concepts (SLIs/SLOs), backup, and DR practices for OLTP/distributed-SQL engines. Proficiency in Mandarin (spoken and written), to support coordination with local Ali Cloud teams and regional vendors/partners. Contributions to open-source database projects or internal data-platform tooling.