SRE Software Engineer
Core
Deploy and manage large-scale, multi-tenant Kubernetes environments powering Apple's services, ensuring global scale, high availability, and reliability.
Role type
Senior IC Site Reliability Engineer (Kubernetes/Infrastructure)
Builds
Production-scale Kubernetes platform, controllers, namespace management infrastructure, and AI-assisted operational tooling
Domain
Cloud infrastructure, container orchestration, and system reliability
Required skills
Linux systems administration, containerization, Python or Go scripting, networking fundamentals, version control, configuration management
Preferred skills
Kubernetes orchestration, cloud platforms (AWS/GCP/Azure), cloud-native observability, CI/CD pipelines, OS security hardening
Technologies
Kubernetes, Docker, Prometheus, Thanos, Splunk, Git, Puppet, Ansible, RHEL, Oracle Linux, CentOS
Responsibilities
Deploy and maintain large-scale Kubernetes environments; write operational tooling to improve reliability; implement reliability standards (SLOs, error budgets, alerting); contribute to CI/CD pipelines; take on-call and troubleshoot production issues; enforce security best practices
Seniority
Mid-level to Senior, hands-on IC
