Skip to content
Skip to content
DevOps Jobs
Apple

Site Reliability Engineer, Apple Data Platform

Apple

Location
Onsite (Austin, Texas)
Employment
Full-time
Level
Senior Level
Posted 4 days ago

About the Role

Apple is seeking a Site Reliability Engineer to manage planetary-scale infrastructure and applications for its global data platform. The role ensures high availability, performance, and scalability for services supporting Apple Music, iCloud, Siri, and other core products.

Skills

Site Reliability Engineering Go Python Java Kubernetes Linux Cloud Computing Distributed Systems Capacity Planning Disaster Recovery Performance Tuning Data Analytics Infrastructure Automation Troubleshooting Networking Protocols Virtualization

Full job details

People at Apple don’t just build products — they craft the kind of experience that have revolutionized entire industries. The diverse collection of our people and their ideas inspire innovation in everything we do. Imagine what you could do here! Join Apple, and help us leave the world better than we found it. Apple Services Engineering (ASE) is responsible for designing and maintaining the systems, platforms, and infrastructure that support Apple's global services, such as Apple Music, iCloud, Siri, Maps, and many more. Our work forms the foundation upon which our world-class software developers build the products our customers love. We are seeking innovative and dedicated Site Reliability Engineers to help us sustain our mission of providing the highest quality experience for our customers. ASE services must scale globally, remain highly available and consistently performant. If you are passionate about designing, engineering, and running systems and infrastructure that will help millions of customers, then this is the place for you!

Description


Apple Services infrastructure is planetary scale. Our Data Platform Site Reliability Engineering team manages the infrastructure and applications on bare-metal and cloud computing platforms to deliver data processing, governance, and storage for many of Apple’s global products and organizations. Our platform teams work with exabytes of data, terabytes of memory, and hundreds of thousands of jobs running millions of executors to support predicable and performant data analytics. Our platform enables key features in Apple Music, TV, Maps, News, and other world class products. Ensuring all of these technologies in geographically distributed data centers work together in harmony presents unique challenges.

Minimum Qualifications


BS/MS in Computer Science or Equivalent 5+ years of software development or production operations experience in a large-scale environment Proficiency in authoring and releasing code in Go, Python, or Java using common configuration management and software delivery platforms Experience operating production applications at scale, including well designed performance testing, HA and disaster recovery concepts, capacity planning, and managing distributed systems on internal and public cloud infrastructure, principally Kubernetes Understanding of the Linux Operating System, containers and virtualization, standard networking protocols, and components Strong sense of ownership and integrity demonstrated through clear communication and collaboration Demonstrates excellent troubleshooting and problem solving skills using the scientific method

Preferred Qualifications


Proficiency with the architecture, deployment, performance tuning, and troubleshooting of open source data analytics or governance technologies such as Flink, Hive, Hadoop/HDFS, Trino, and/or Druid. Proficiency in managing applications and infra on AWS, GCP and Ali Cloud. The successful candidate is frustrated with toil and has an acute drive to both automate manual operations and evolve them into automatic processes.