Amazon Web Services in Cupertino, CA leads the development of the network stack for EC2 distributed AI/ML systems, enabling NVIDIA GPUs to work with AWS machines. You will direct a team of engineers to deliver features for the largest AI models and workloads.
The role requires 5+ years in system design and the software development lifecycle, with strong C/C++ background and prior tech-lead/mentorship experience.
#J-18808-LjbffrLead ML Network Stack Engineer (RDMA, CUDA, NCCL) in cupertino at Unknown Company
This position is listed as full time and onsite.