Site Reliability Engineer
Core
Senior SRE leading technology initiatives in cloud infrastructure, software delivery, and observability to ensure platform reliability and operational excellence.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Cloud infrastructure, automated tooling, CI/CD pipelines, and self-service developer platforms.
Domain
Cloud infrastructure and Platform Engineering
Required skills
Kubernetes orchestration, Terraform, Python, CI/CD automation, observability (alerting/tracing), root cause analysis, system optimization, Bash/Golang/Java scripting, GCP/AWS/Azure cloud platforms
Preferred skills
Helm, Ansible, Splunk, GCP Monitoring, GitHub, Artifactory, Agile methodology
Technologies
Kubernetes, Terraform, Python, Helm, Docker, GCP, CI/CD tools, Splunk, GCP Monitoring, Bash, Golang, Java, AWS, Azure
Responsibilities
Build and develop tooling, policies, and processes to advance scale and performance; implement observability, alerting, tracing, automation, and self-healing capabilities; coordinate end-to-end across platforms for issue response and escalation; develop maintenance and operations automation through CI/CD; participate in on-call rotation for production troubleshooting and root cause analysis.
Seniority
Senior, hands-on IC