NVIDIA is seeking an experienced Solutions Architect to redefine AI inference infrastructure. You will lead high-stakes engagements, design scalable, disaggregated inference architectures, and mentor teams across customers and partners.
Expect deep collaboration with partners to push performance and reliability across GPU clusters and cloud environments. You will work with Dynamo, Triton Inference Server, and TensorRT-LLM to optimize models and pipelines, while guiding adoption of best-practices
#J-18808-LjbffrSenior AI Inference Architect — Equity & Scalable GPU Solutions in santa clara at Unknown Company
This position is listed as full time and onsite.