Sr Mgr, Site Reliability Engineer (SRE)
Core
Provide strategic leadership for multiple SRE teams to ensure alignment with organizational priorities and drive operational excellence across commerce platforms.
Role type
Senior Manager, Site Reliability Engineering
Builds
Highly scalable, fault-tolerant systems across cloud and on-prem environments powering guest experiences
Domain
Entertainment technology, Commerce platforms, Cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
Strategic leadership, Observability implementation, Cloud platform expertise, Container orchestration, Infrastructure-as-code, CI/CD automation, Stakeholder influence, Team mentorship
Preferred skills
FAANG-level engineering standards experience, Serverless architecture knowledge, Organizational transformation management
Technologies
AWS, GCP, Azure, Kubernetes, Terraform, CloudFormation, Ansible, Harness, GitLab
Responsibilities
Lead and inspire multiple SRE teams to foster a culture of reliability and continuous improvement; Oversee design and delivery of scalable systems; Implement advanced telemetry and monitoring practices leveraging AI/ML; Guide teams in automating infrastructure and CI/CD pipelines; Develop and execute departmental plans aligned with business objectives; Mentor and develop leaders while setting clear OKRs
Seniority
Senior, hands-on IC with management responsibilities