NVIDIA is seeking a Senior Deep Learning Algorithms Engineer to advance Dynamo, our open-source distributed inference platform for large-scale, low-latency AI services. You’ll lead architecture and performance work across Dynamo and open source frameworks, collaborating with research, software, systems, and hardware teams to accelerate AI inference.
You will design, implement, and optimize integrations with vLLM, SGLang, and TRTLLM, while partnering with open source communities to improve
#J-18808-LjbffrSenior DL Inference Architect (Hybrid) in santa clara at Unknown Company
This position is listed as full time and onsite.