A tech startup in AI model serving located in San Francisco is seeking a qualified candidate to architect scalable inference systems. The role focuses on optimizing model serving performance and integrating advanced techniques for AI deployment. Candidates should have strong expertise in Python and PyTorch, along with low-level systems knowledge. This in-person position offers a fast-paced work environment that is ideal for those eager to tackle complex challenges and shape the future of AI technology.
#J-18808-Ljbffr
#J-18808-Ljbffr