Thinking Machines Lab Inc. is hiring a Software Engineer, Inference to own the reliability, scale, and efficiency of serving our AI models to real users in a production environment.
You will bridge cutting-edge inference techniques with real traffic, rolling out new models safely and improving observability. The role focuses on multi-tenant serving, capacity planning, and incident response to keep the platform fast and resilient as usage grows.
#J-18808-LjbffrProduction ML Inference Engineer — Unlimited PTO in san francisco at Unknown Company
This position is listed as full time and onsite.