Unknown Company

AI Inference Performance Architect

santa clara, ca • Posted 6 days ago
Onsite Full Time IT & Technology

NVIDIA is seeking a skilled engineer in Santa Clara to optimize GenAI inference on advanced accelerators. Responsibilities include enhancing industry benchmarks and defining next-gen workloads through collaboration across various teams.

Ideal candidates will possess experience in Python and C++, a strong understanding of LLM mechanics, and a proven ability to deliver performance improvements in high-stakes settings. NVIDIA values diversity and committed to inclusive hiring practices.

#J-18808-Ljbffr
Back to Job Search