NVIDIA is seeking a leader to drive the global strategy for scaled-out AI inferencing. You will architect high-throughput, low-latency distributed pipelines and model serving strategies for massive scale and reliability across enterprise and cloud environments.
You will define the technical roadmap for full lifecycle deployment, versioning, and automated scaling, collaborating with leadership to ensure peak efficiency on NVIDIA hardware.
#J-18808-LjbffrSenior Architect, Scaled AI Inference & Orchestration in california at Unknown Company
This position is listed as full time and onsite.