Staff Site Reliability Engineer (f/m/d)
Core
Design, implement, migrate, and operate highly available distributed management services on Linux and Kubernetes.
Role type
Staff Site Reliability Engineer (Infrastructure)
Builds
Distributed, highly available services and in-house K8s operators
Domain
Cloud Infrastructure / DevOps / Linux / Kubernetes
Required skills
Linux administration, Kubernetes (K8s), Infrastructure as Code, Go, Python, Observability (Loki, Prometheus, Mimir, Grafana), System-level development, Automation
Preferred skills
Open Source contribution, Security vetting readiness
Technologies
Linux, Kubernetes, Go, Python, Loki, Prometheus, Mimir, Grafana
Responsibilities
Provision, operate, and migrate distributed services; Develop and maintain K8s operators; Support devs in automating operational tasks; Troubleshoot complex infrastructure; Participate in on-call rotation; Manage lifecycle and security of infrastructure
Seniority
Staff, hands-on IC
