Senior Site Reliability Engineer (f/m/x)
Core
Ensure high availability and reliability of B2C and B2B sports booking applications on AWS.
Role type
Senior Site Reliability Engineer
Builds
Production reliability, observability, and automation for sports booking platform
Domain
Sports technology / Cloud infrastructure
Required skills
AWS operations, SRE frameworks (SLIs/SLOs), incident management, Infrastructure as Code (AWS CDK), relational databases (MySQL/Postgres), observability (OpenTelemetry/Grafana), CI/CD, scripting (TypeScript/Python/Bash)
Preferred skills
Coaching teams on reliability practices, German language skills
Technologies
AWS, ECS Fargate, Lambda, SQS, SNS, EventBridge, RDS MySQL, GitHub Actions, OpenTelemetry, Grafana, Prometheus, TypeScript, Node.js, Bash, Python
Responsibilities
Own production monitoring, alerting, and observability; lead incident response and root-cause analysis; enable team ownership of application health; automate operational toil; strengthen backup and disaster recovery practices; partner with Platform Engineering on CI/CD initiatives
Seniority
Senior, hands-on IC