NVIDIA is seeking a highly technical Product Manager to own AI inference optimization on NVIDIA hardware, from single-GPU workstations to large data centers. You will translate deep optimization techniques into capabilities that customers can adopt across the inference stack.
The role spans framework integration with TensorRT-LLM, vLLM, SGLang, and NVIDIA Dynamo, plus benchmarking, release readiness, and cross-functional collaboration to deliver measurable performance gains and compelling
#J-18808-LjbffrSenior AI Inference Platform PM — Performance & Scale in santa clara at Unknown Company
This position is listed as full time and onsite.