Site Reliability Engineer
Core
Design, build, and deliver highly scalable, reliable, secure cloud infrastructure powering Apple's customer support applications and services.
Role type
Senior Site Reliability Engineer
Builds
Cloud infrastructure, CI/CD pipelines, automation tools, and monitoring systems for customer-facing services
Domain
Enterprise technology, cloud infrastructure, customer support systems
Required skills
Kubernetes, Spinnaker, Helm, Splunk, ELK stack, Grafana, Prometheus, Alertmanager, Jenkins, Linux Shell Scripting, Python, Terraform, AWS S3, Cassandra, MongoDB, Couchbase, Java application support, networking protocols (DNS, TCP, HTTP/HTTPS)
Preferred skills
Infrastructure as Code, production troubleshooting, automation tool development, cross-functional collaboration
Technologies
Kubernetes, Spinnaker, Helm, Splunk, ELK, Grafana, Prometheus, Alertmanager, Jenkins, Terraform, AWS S3, Cassandra, MongoDB, Couchbase
Responsibilities
Design and build resilient large-scale cloud and on-prem infrastructure; manage Kubernetes clusters; set up and manage CI/CD pipelines; monitor systems using specified tools; troubleshoot production issues; collaborate with distributed teams to define and implement technical requirements
Seniority
Senior, hands-on IC