NVIDIA is seeking a senior technical leader to architect and drive the global strategy for scaled-out AI inference. You will design high-throughput, low-latency distributed pipelines and model serving strategies for massive scale and production reliability across enterprise and cloud environments.
You will define the technical roadmap for deployment, versioning, and automated scaling while coordinating with leadership to enable NVIDIA's AI models to run efficiently on our accelerated computing
#J-18808-Ljbffr