Research Engineer, AI/LLM/HPC Network
Core
Designing, building, and operating global intelligent network infrastructure to support high-performance AI/LLM applications at hyperscale.
Role type
Research Engineer, AI/LLM/HPC Network
Builds
Next-generation AI network infrastructure, high-performance communication frameworks, and network protocol stacks.
Domain
Data center networking, AI/LLM systems, high-speed networking
Deliverable
production ML models
Required skills
High-speed network systems, network programming, C/C++, Python, Go, RDMA, congestion control, AI network optimization, host-network-application codesign
Preferred skills
High-performance communication frameworks (NCCL, MPI, RPC), AI network diagnosis and performance optimization
Responsibilities
Design, implementation, and deployment of high-speed network technologies for AI/LLM; Develop platforms for monitoring, analysis, and diagnosis of large-scale AI/LLM networks; Research and develop high-performance AI communication frameworks and protocol stacks; Build next-generation AI network infrastructure supporting heterogeneous hardware.
