Site Reliability Engineer (SRE) - Elastic Disk
Core
Building and running next-generation distributed storage systems for Apple's critical cloud services.
Role type
Senior IC Site Reliability Engineer (Storage)
Builds
Distributed storage systems, block/object/file storage layers, and content delivery infrastructure for Apple Cloud.
Domain
Cloud Infrastructure / Distributed Storage Systems
Required skills
Linux system internals, distributed systems architecture, block/object/file storage solutions, Go/Java/Python/Rust, capacity planning, disaster recovery, data migration
Preferred skills
Kubernetes, microservices architecture, configuration management (Puppet/Chef/Ansible), database management (Cassandra/Postgres/RocksDB), automation of manual operations
Technologies
Linux, Go, Java, Python, Rust, Kubernetes, Puppet, Chef, Ansible, Spinnaker, Ceph, Gluster, NFS, S3, LVM, XFS, ext4
Responsibilities
Design and release code for storage systems, take ownership of projects to shape platform direction, provide technical feedback to colleagues, drive technical standards across the team, handle on-call alerts and critical issues to maintain service availability
Seniority
Senior, hands-on IC