Hewlett Packard Enterprise Development LP is seeking a Senior Software Engineer to advance the model runtime for our AI inference platform. You will design and implement core components, optimize batching and KV cache strategies, and collaborate with cross-functional teams to push performance on enterprise hardware.
You will work in a hybrid setup with occasional on-site requirements in US locations, contributing to scalable, low-latency inference in air-gapped and regulated environments.
#J-18808-LjbffrSenior LLM Inference Engineer – Hybrid (Remote/On-site) in fort collins at Unknown Company
This position is listed as full time and hybrid.