Thinking Machines is hiring a Software Engineer, Inference to own the reliability, scale, and efficiency of the systems that serve our models to real users. This production-facing role sits at the center of a fast-growing platform, bridging cutting-edge inference techniques with real traffic, including rollouts, capacity planning, incidents, and resiliency.
You will operate and scale live traffic systems, own model rollout processes, and collaborate with research teams to productionize new
#J-18808-LjbffrProduction ML Inference Engineer — Scale & Reliability in san francisco at Unknown Company
This position is listed as full time and onsite.