Senior Site Reliability Engineer, Apple Data Platform SRE / Apple Services Engineering
Core
Principal SRE leading the Apple Data Platform Compute team to ensure petabyte-scale analytics infrastructure reliability and efficiency.
Role type
Principal/Lead Site Reliability Engineer
Builds
Bare-metal and cloud-based infrastructure for Apple services (iCloud, iTunes, Siri, Maps)
Domain
Cloud infrastructure & Big Data
Deliverable
production ML models | infrastructure
Required skills
Hadoop, Kubernetes, Python, Golang, Linux, Networking, Containers, Infrastructure-as-Code, Capacity Planning, Fleet Management
Preferred skills
Scale testing, Disaster Recovery, Technical Roadmap Definition, Cross-functional Alignment
Technologies
Hadoop, Kubernetes, Python, Golang, Generative AI tooling
Responsibilities
Mentor engineers and partner teams on service architecture and tooling; Manage infrastructure fleet capacity and hardware lifecycle; Provide technical leadership for Hadoop/Kubernetes infrastructure; Develop mission-critical automation and tools; Handle production on-call and incident management.
Seniority
Principal, hands-on IC with leadership
