Site Reliability Engineer - Kafka
Core
Building and running large-scale, fault-tolerant data streaming infrastructure and platform services using Kafka.
Role type
Senior Site Reliability Engineer (Kafka/Data Streaming)
Builds
Next-generation Kafka infrastructure, deployment automation, monitoring tooling, and control plane enhancements.
Domain
Cloud infrastructure, distributed systems, data streaming
Deliverable
infrastructure
Required skills
Kafka infrastructure management, Kubernetes operations, performance engineering, incident management, automation tooling, multi-datacenter system design, IaC (Terraform), cloud architecture (AWS/GCP), Java/Go/Python
Preferred skills
Messaging services management, deep troubleshooting of distributed systems and database storage engines
Sourced via apple · Listed on CareerPlan, which tracks 853,000+ jobs from 20+ sources.