Unknown Company

Senior AI Inference Systems Engineer | GPU Kernels & Runtime

santa clara, ca • Posted 5 days ago
Onsite Full Time Software Architecture & Engineering

NVIDIA in Santa Clara, CA is seeking outstanding AI systems engineers to advance the inference software stack. You will design and optimize kernels, build new abstractions for LLM serving engines, and contribute to accelerators and runtimes that power large language models and AI workloads.

You will collaborate across teams, work on CUDA C/C++, Triton, and cutting-edge MLIR-based tooling, and participate in open source projects.

#J-18808-Ljbffr
Back to Job Search