NVIDIA is seeking outstanding AI systems engineers to advance inference software, building libraries, code generators, and GPU kernel technologies for our hardware architecture. You will design abstractions for LLM serving engines and just-in-time compilers to accelerate large language models and agents.
You will work with NVIDIA teams on deep learning frameworks and kernels, contributing to open source ecosystems like FlashInfer, vLLM, and SGLang while shaping high-performance AI workloads
#J-18808-LjbffrSenior AI Inference Engineer: GPU Kernels & LLM Runtimes in redmond at Unknown Company
This position is listed as full time and onsite.