NVIDIA is seeking a Senior Deep Learning Algorithms Engineer to advance Dynamo, our open-source distributed inference platform for large-scale, low-latency AI services. You’ll lead architecture and performance work across Dynamo and open source frameworks, collaborating with research, software, systems, and hardware teams to improve AI inference speed and efficiency.
You will work with vLLM, SGLang, and TensorRT-LLM, contributing to cutting-edge inference techniques, scheduling optimizations,
#J-18808-Ljbffr