Unknown Company

Staff GenAI Inference Engineer: Optimize LLM Serving Latency

san francisco, ca • Posted 1 weeks ago
Onsite Full Time Software Development
A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong software engineering background and a proven ability to collaborate with researchers and drive architectural decisions. Competitive compensation is offered, with a salary range of $190,900 to $232,800 USD.
#J-18808-Ljbffr
Back to Job Search