NVIDIA is seeking a skilled engineer in Santa Clara to optimize GenAI inference on advanced accelerators. Responsibilities include enhancing industry benchmarks and defining next-gen workloads through collaboration across various teams.
Ideal candidates will possess experience in Python and C++, a strong understanding of LLM mechanics, and a proven ability to deliver performance improvements in high-stakes settings. NVIDIA values diversity and committed to inclusive hiring practices.
#J-18808-Ljbffr