Talanto in the United States is seeking an ML Infrastructure Engineer, Model Inference to build and optimize our core inference infrastructure powering AI models. You will work with Infrastructure and Research to deploy, optimize, and orchestrate AI models across scalable systems.
The role requires 5+ years in production ML, strong Kubernetes skills, and experience with model serving frameworks and GPU optimization. Hybrid options may apply with a focus on scalable backend infrastructure.
#J-18808-LjbffrML Inference Infrastructure Engineer — Scale & GPU in northern at Unknown Company
This position is listed as full time and hybrid.