Unknown Company

ML Compiler Engineer for AI Inference Accelerators

cupertino, ca • Posted Today
Onsite Full Time Electrical & Energy Engineering

Amazon is seeking a software engineer for the AWS Neuron Compiler team in Cupertino, CA. You will build the next‑generation Neuron compiler to translate ML models from PyTorch, TensorFlow, and JAX for deployment on AWS Inferentia/Trainium in the cloud.

You will work with OpenXLA, StableHLO and MLIR, design optimization passes, and collaborate with chip architects, runtime engineers, scientists, and ML apps teams to deliver high performance and a great developer experience.

#J-18808-Ljbffr

ML Compiler Engineer for AI Inference Accelerators in cupertino at Unknown Company

This position is listed as full time and onsite.

Back to Job Search