Amazon Inc. in Seattle, WA is seeking a Sr. Software Development Engineer for the Inference Team on AWS Neuron to deliver high-performance model inference for customer workloads on Inferentia- and Trainium-powered instances.
You will lead customization of open-source inference frameworks (vLLM, SGLang) and drive kernel-level and framework optimizations, collaborating across model development, compiler, and runtime teams to ensure scalable, production-ready performance.
#J-18808-LjbffrSenior ML Inference Engineer – High-Performance Serving in seattle at Unknown Company
This position is listed as full time and onsite.