SRE高级软件工程师(云原生方向)-基础架构
Core
Design and maintain production clusters and services to ensure high availability, scalability, and stability for large-scale distributed systems.
Role type
Senior Site Reliability Engineer (Cloud Native)
Builds
Automated platforms for rapid iteration of large-scale online clusters and core systems like big data and distributed storage.
Domain
Cloud Native, Big Data, Distributed Systems
Required skills
Go, Python, Java, C++, Linux, Network IO, Distributed Systems, Big Data, System Design, Problem Solving
Preferred skills
Large-scale distributed system design, Product thinking, Data Structures, English fluency
Responsibilities
Participate in architecture planning, review, design, deployment, and launch of production clusters; Ensure high availability and performance of core systems; Build automated engineering solutions to prevent issues; Optimize service governance practices and troubleshoot performance bottlenecks.
Seniority
Senior, hands-on IC