NVIDIA is seeking a Software Engineer to bring up, triage, benchmark, and optimize distributed training and inference workloads across GPU platforms at scale.
You will work on multi-GPU and multi-node LLM workloads, build benchmarking tooling, and collaborate with framework, systems, and platform teams to deliver data-driven recommendations.
A strong background in Python and C/C++, CUDA-enabled distributed execution, and debugging at scale is required. Equity and benefits are included.
#J-18808-LjbffrDistributed AI Systems Engineer - LLM Benchmarking in santa clara at Unknown Company
This position is listed as full time and onsite.