Apple

Senior Data Engineer, iCloud

Apple Shanghai, Shanghai, China

Computers and Electronics Manufacturing · 10,001+ employees

4 h ago
data-engineer Senior (5-10 yrs) Full-time China
Log in to apply, save this posting, or score it against your profile with AI.

About the role

You will design and maintain scalable, high-performance data infrastructure to support iCloud services and cross-functional data needs. The role involves solving complex data engineering problems and leading technical strategy to improve product data utilization.

What they look for

Spark Hadoop Presto Flink Druid Java Python Scala Data Architecture Data Modeling SQL Distributed Systems AWS Data Pipelines NoSQL Search Systems

Requirements

Candidates must have 8+ years of experience in distributed data processing and proficiency in Java, Python, or Scala. A strong background in data modeling, SQL, and distributed technologies like Spark is required, along with a degree in a technical field.

Full description

Would you like to drive the future of Apple’s products, applications, and platform while having the unique opportunity to impact some of the most far-reaching software applications in the world?

iCloud Data organization enables Apple to ship better products for iCloud users to access all their content across apps (Photos, Mail, Messages, FaceTime, Calendar, etc) from all of their devices all the time by providing consistent, scalable, timely, accurate, complete and fully integrated data infrastructure and capabilities to accurately surface relevant information. If this excites you and you are energized by solving hard, high-leverage problems at scale, we'd love to hear from you!

Description

We’re looking for exceptional data engineers who have a strong background in distributed data processing, have great and demonstrable data intuition, and share our passion for continuously improving the ways we use data to make Apple’s products, applications, and platform better.

Minimum Qualifications

8+ years of experience working with Spark and other distributed data technologies (e.g. Hadoop, Presto, Flink, Druid) for building efficient & large scale data pipelines Highly proficient in at least one of Java, Python or Scala Deep expertise in Data Principles, Data Architecture & Data Modeling, Strong SQL skills Strong problem solver with meticulous attention to detail, capable of taking on loosely defined problems Experience working in a complex, matrixed organization involving cross-functional, and/or cross-business projects Strong communication and collaboration skills & ability to lead high-level discussions on technology strategy and approach Conceptually familiar with AWS cloud resources (S3, EC2, RDS etc) MS or BS in Computer Science, Engineering, Mathematics, Statistics or a related field OR equivalent practical experience in Software or Data Engineering

Preferred Qualifications

Experience with Cloud Computing platforms like Amazon AWS, Google Cloud Experience with building stream-processing applications using Apache Flink, Spark-Streaming, Apache Storm, Kafka Streams or others Experience with Search systems (such as ElasticSearch, Solr), NoSQL datastores (such as HBase, Cassandra, MongoDB) Experience building distributed, high-volume data services is a plus

Similar roles