NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models and AI workloads.
You will collaborate across teams, work on CUDA C/C++, Triton, and cutting-edge MLIR-based tooling, and participate in open source projects.
#J-18808-Ljbffr