Senior Manager - AI Network Engineering
Core
Lead the full NPI lifecycle for high-performance NICs supporting OCI GPU and AI infrastructure, including architecture, qualification, and production readiness.
Role type
Senior Manager, AI Network Engineering
Builds
High-performance NICs, OCI Collective Communication Library, and large-scale AI cluster networking solutions.
Domain
Cloud infrastructure, AI/ML, High-Performance Computing (HPC)
Required skills
High-performance networking (Ethernet, RDMA, RoCE, PCIe, DMA), GPU systems and distributed communication, hardware/software co-design, cross-functional program leadership, performance analysis, NPI lifecycle management
Preferred skills
Experience guiding new hardware technologies from early engineering to production, silicon vendor partnerships, large-scale GPU cluster networking (thousands to tens of thousands of accelerators), NCCL, MPI, UCX, CUDA, SmartNICs/DPUs
Technologies
CUDA, Ethernet, RDMA, RoCE, PCIe, DMA, NCCL, MPI, UCX, SmartNICs, DPUs
Responsibilities
Lead the full NPI lifecycle for current and next-generation high-performance NICs; build and lead the engineering organization for the OCI Collective Communication Library; improve collective communication and topology-aware algorithms across clusters; own GPU cluster networking performance and develop benchmarking capabilities; drive hardware and software co-design; establish automated qualification and testing processes; lead complex cross-functional programs and present technical strategy to senior leadership
Seniority
Senior Manager, hands-on IC with team leadership
