Unknown Company

LLM Inference Engineer — Scalable AI Serving

san francisco, ca • Posted 5 days ago
Onsite Full Time Networks & Systems
A tech startup in AI model serving located in San Francisco is seeking a qualified candidate to architect scalable inference systems. The role focuses on optimizing model serving performance and integrating advanced techniques for AI deployment. Candidates should have strong expertise in Python and PyTorch, along with low-level systems knowledge. This in-person position offers a fast-paced work environment that is ideal for those eager to tackle complex challenges and shape the future of AI technology.
#J-18808-Ljbffr
Back to Job Search