Sr. Site Reliability Engineer, Infrastructure Services
Core
Ensure reliability, scalability, and performance of Apple's Backbone Network including hardware, software, and toolset used to configure/monitor the environment.
Role type
Senior Site Reliability Engineer (Infrastructure)
Builds
Apple Cloud services (iCloud, iTunes, Siri, Maps) and internal developer platforms
Domain
Enterprise IT infrastructure, Cloud Networking, DevOps
Required skills
Python, Bash, Go, Java, CI/CD tools, monitoring and logging tools, Linux, Networking concepts, troubleshooting in large-scale environments
Preferred skills
Cloud platforms (AWS, GCP), container technologies (Docker, Kubernetes), databases (Relational/NoSQL), event-driven architectures, security and authentication
Technologies
Jenkins, GitHub Actions, GitLab, Splunk, AppDynamics, Datadog, Prometheus, Grafana, AWS, GCP, Docker, Kubernetes, Oracle, MongoDB, Kafka, RabbitMQ, OAuth, SAML, SSO
Responsibilities
Ensure reliability, scalability, and performance of application and networking platforms, APIs, and integrations; Collaborate with development teams to design resilient hosting architectures and streamline CI/CD pipelines; Implement observability/monitoring best practices; Troubleshoot incidents; Promote DevOps culture to improve release velocity
Seniority
Senior, hands-on IC