Site Reliability Engineer
Core
Ensure cloud platform security, reliability, scalability, and operational excellence for IoT smart home products and networking services.
Role type
Site Reliability Engineer (IC)
Builds
Cloud-based microservices on Multi-Cloud Platform (AWS, OCI, Azure, GCP)
Domain
Networking, IoT, Cloud Infrastructure
Required skills
Kubernetes, Cloud Operations, DevOps, Python, Go, Bash, Chaos Engineering, Load Testing, Incident Response, Technical Documentation, Security Compliance (ISO27001, SOC2, GDPR), KPI Definition (SLA/SLO/SLI)
Preferred skills
AWS/Azure/GCP Solutions Architect Certifications, Container Orchestration
Technologies
Kubernetes, AWS, OCI, Azure, GCP, Python, Go, Bash, Java, PowerShell
Responsibilities
Implement and operate microservices on Kubernetes; Conduct load and chaos tests; Build observability; Write automation scripts; Define and maintain KPIs; Participate in incident response and post-incident analysis; Create technical documentation; Ensure security and compliance adherence; Mentor less senior members; Participate in on-call rotation.
Seniority
Mid-level, hands-on IC