Site Reliability Engineer - Data, Apple Ads
Apple
- Location
- Onsite (Austin, Texas)
- Employment
- Full-time
- Level
- Mid Level
Posted 2 weeks ago
About the Role
Apple Ads is seeking a Site Reliability Engineer to design, build, and manage cloud-based infrastructure for storing, processing, and analyzing large datasets. This role ensures data systems remain accessible, reliable, and secure within AWS and other cloud environments, supporting advertising services across Apple's ecosystem.
Skills
Infrastructure Engineering
Cloud Infrastructure
AWS
Apache Spark
Flink
Kafka
Iceberg
Kubernetes
Machine Learning
GenAI
Java
Scala
Kotlin
Python
Helm
Infrastructure as Code
Full job details
At Apple, we focus deeply on our customers’ experience. Apple Ads brings this same approach to advertising, helping people find exactly what they’re looking for and helping advertisers grow their businesses.
Our technology powers ads and sponsorships across Apple Services, including the App Store, Apple News, MLS Season Pass and Apple Maps. Everything we do is designed for trust, connection, and impact: We respect user privacy, integrate advertising thoughtfully into the experience, and deliver value for advertisers of all sizes—from small app developers to big, global brands. Because when advertising is done right, it benefits everyone.
As a Data Site Reliability Engineer in Apple Ads, you will focus on designing, building, and managing cloud-based systems and infrastructure that stores, process, transforms and analyzes large datasets. You will ensure data and corresponding infrastructure is accessible, reliable, and secure within the cloud environment, working with various Apple Internal and 3rd party cloud providers like AWS.
2+ years of experience in Infrastructure Engineering with a strong focus on building, scaling and operating cloud based distributed data systems. Proven expertise running cloud infrastructure built on managed services offered by cloud providers like AWS Strong technical grasp and experience working on Open Source technologies designed for large scale data processing, like Apache Spark, Flink, Kafka, Iceberg or other similar technologies Strong expertise and experience with container orchestration systems like Kubernetes Experience leveraging ML and GenAI capabilities to improve infrastructure and Operational efficiency
Big Data processing and storage systems (streaming and batch) Track record of driving automation, cost optimization and performance tuning at scale for data systems. Experience designing, analyzing and troubleshooting large-scale distributed systems. Experience driving adoption of new technologies and influencing designing of systems. Strong programming skills on object-oriented, functional or procedural programming languages, preferably Java / Scala / Kotlin / Python Experience designing and managing Infrastructure as Code with Helm and CRD, ensuring repeatable, secure, and scalable deployments.
Description
As a Data Site Reliability Engineer in Apple Ads, you will focus on designing, building, and managing cloud-based systems and infrastructure that stores, process, transforms and analyzes large datasets. You will ensure data and corresponding infrastructure is accessible, reliable, and secure within the cloud environment, working with various Apple Internal and 3rd party cloud providers like AWS.
Minimum Qualifications
2+ years of experience in Infrastructure Engineering with a strong focus on building, scaling and operating cloud based distributed data systems. Proven expertise running cloud infrastructure built on managed services offered by cloud providers like AWS Strong technical grasp and experience working on Open Source technologies designed for large scale data processing, like Apache Spark, Flink, Kafka, Iceberg or other similar technologies Strong expertise and experience with container orchestration systems like Kubernetes Experience leveraging ML and GenAI capabilities to improve infrastructure and Operational efficiency
Preferred Qualifications
Big Data processing and storage systems (streaming and batch) Track record of driving automation, cost optimization and performance tuning at scale for data systems. Experience designing, analyzing and troubleshooting large-scale distributed systems. Experience driving adoption of new technologies and influencing designing of systems. Strong programming skills on object-oriented, functional or procedural programming languages, preferably Java / Scala / Kotlin / Python Experience designing and managing Infrastructure as Code with Helm and CRD, ensuring repeatable, secure, and scalable deployments.