Senior Data Engineer, iCloud
Summary
Build and optimize large-scale data pipelines for Apple’s iCloud to ensure seamless, real-time data access across devices and apps using Spark, Hadoop, and distributed systems.
Would you like to drive the future of Apple’s products, applications, and platform while having the unique opportunity to impact some of the most far-reaching software applications in the world?
iCloud Data organization enables Apple to ship better products for iCloud users to access all their content across apps (Photos, Mail, Messages, FaceTime, Calendar, etc) from all of their devices all the time by providing consistent, scalable, timely, accurate, complete and fully integrated data infrastructure and capabilities to accurately surface relevant information. If this excites you and you are energized by solving hard, high-leverage problems at scale, we'd love to hear from you!
We’re looking for exceptional data engineers who have a strong background in distributed data processing, have great and demonstrable data intuition, and share our passion for continuously improving the ways we use data to make Apple’s products, applications, and platform better.
Minimum Qualifications
- 8+ years of experience working with Spark and other distributed data technologies (e.g. Hadoop, Presto, Flink, Druid) for building efficient & large scale data pipelines
- Highly proficient in at least one of Java, Python or Scala
- Deep expertise in Data Principles, Data Architecture & Data Modeling, Strong SQL skills
- Strong problem solver with meticulous attention to detail, capable of taking on loosely defined problems
- Experience working in a complex, matrixed organization involving cross-functional, and/or cross-business projects
- Strong communication and collaboration skills & ability to lead high-level discussions on technology strategy and approach
- Conceptually familiar with AWS cloud resources (S3, EC2, RDS etc)
- MS or BS in Computer Science, Engineering, Mathematics, Statistics or a related field OR equivalent practical experience in Software or Data Engineering
Preferred Qualifications
- Experience with Cloud Computing platforms like Amazon AWS, Google Cloud
- Experience with building stream-processing applications using Apache Flink, Spark-Streaming, Apache Storm, Kafka Streams or others
- Experience with Search systems (such as ElasticSearch, Solr), NoSQL datastores (such as HBase, Cassandra, MongoDB)
- Experience building distributed, high-volume data services is a plus