NVIDIA is seeking an engineer to design, develop, and operate distributed systems for large-scale AI workloads, spanning GPUs, storage, and multi-region clusters. You will build software and automation to orchestrate workloads across thousands of GPUs and petabytes of storage.
Collaborate with AI/ML teams to translate requirements into scalable, high-performance solutions and drive improvements in reliability, performance, and observability to meet exascale standards.
#J-18808-Ljbffr