NVIDIA is seeking a Senior Deep Learning Algorithms Engineer to advance Dynamo, our open-source distributed inference platform for large-scale AI services. You’ll lead architecture and performance work across Dynamo and open source frameworks, collaborating across research, software, systems, and hardware teams to make AI inference faster and easier to deploy.
The role focuses on integrating with vLLM, SGLang, and TRTLLM, reducing latency, and improving throughput.
#J-18808-Ljbffr