vCluster Labs seeks a Senior Inference Engineer to own and build the platform’s inference layer from the ground up. You will work with the CTO to deploy LLMs on GPU infrastructure and shape the roadmap for scalable inference.
In this role you will stand up serving frameworks, optimize for latency and cost, and push production-grade software in Python or Golang. Remote-first, with global team collaboration.
#J-18808-LjbffrSenior Inference Engineer — Production LLMs at Scale in Location not specified at Unknown Company
This position is listed as full time and able to be worked remotely.