Unknown Company

Senior ML Kernel Performance Engineer — AI Acceleration

cupertino, ca • Posted 2 days ago
Onsite Full Time IT & Technology

Annapurna Labs (U.S.) Inc. at Amazon Web Services (AWS) seeks a software engineer to design and optimize high-performance compute kernels for ML operations on the Neuron architecture.

You will leverage the Neuron architecture and programming models to push the limits of ML inference and training on Inferentia and Trainium. Based in Cupertino, you will analyze kernel-level performance, implement compiler optimizations such as fusion, sharding, tiling, and scheduling, and collaborate with

#J-18808-Ljbffr

Senior ML Kernel Performance Engineer — AI Acceleration in cupertino at Unknown Company

This position is listed as full time and onsite.

Back to Job Search