大数据高级研发工程师
Core
Design and develop high-performance data systems for international recommendation scenarios, ensuring system stability and high availability.
Role type
Senior IC big data engineering specialist (real-time streaming & distributed systems)
Builds
Standardized high-performance feature computation, stable production systems, and industry-leading distributed streaming frameworks
Domain
Internet / Recommendation Systems / Big Data Infrastructure
Deliverable
production ML models
Required skills
Flink (DataStream, SQL, Checkpoint, State), real-time stream processing, distributed systems design, Java, C++, Scala, Python, Kafka, RocketMQ, data lake technologies (Hudi, Iceberg, DeltaLake), PB-scale data processing
Preferred skills
Source code reading experience, data lake development, storage systems (HBase, Cassandra, RocksDB), YARN, K8S, Spark, Kudu
Responsibilities
Design and implement data systems for large-scale recommendation systems; troubleshoot production systems and build mechanisms/tools for stability; develop distributed streaming frameworks for massive data and large-scale business systems