Amazon is seeking a software engineer for the AWS Neuron Compiler team in Cupertino, CA. You will build the next‑generation Neuron compiler to translate ML models from PyTorch, TensorFlow, and JAX for deployment on AWS Inferentia/Trainium in the cloud.
You will work with OpenXLA, StableHLO and MLIR, design optimization passes, and collaborate with chip architects, runtime engineers, scientists, and ML apps teams to deliver high performance and a great developer experience.
#J-18808-LjbffrML Compiler Engineer for AI Inference Accelerators in cupertino at Unknown Company
This position is listed as full time and onsite.