Annapurna Labs (U.S.) Inc. at Amazon Web Services (AWS) seeks a software engineer to design and optimize high-performance compute kernels for ML operations on the Neuron architecture.
You will leverage the Neuron architecture and programming models to push the limits of ML inference and training on Inferentia and Trainium. Based in Cupertino, you will analyze kernel-level performance, implement compiler optimizations such as fusion, sharding, tiling, and scheduling, and collaborate with
#J-18808-LjbffrSenior ML Kernel Performance Engineer — AI Acceleration in cupertino at Unknown Company
This position is listed as full time and onsite.