Unknown Company

Staff GenAI Inference Engineer: Optimize LLM Serving Latency

san francisco, ca • Posted 1 months ago
Onsite Full Time Software Development
A leading data and AI company is seeking a Staff Software Engineer for GenAI inference to lead the architecture and optimization of the inference engine. The role requires expertise in CUDA, GPU programming, and distributed systems design. Ideal candidates will have a strong software engineering background and a proven ability to collaborate with researchers and drive architectural decisions. Competitive compensation is offered, with a salary range of $190,900 to $232,800 USD.
#J-18808-Ljbffr

Staff GenAI Inference Engineer: Optimize LLM Serving Latency in san francisco at Unknown Company

This position is listed as full time and onsite.

Back to Job Search