Thinking Machines in San Francisco is seeking a Site Reliability Engineer to enhance the reliability of their Tinker platform. You'll define the end-to-end reliability processes and collaborate with both engineering and research teams.
Ideal candidates will have a Bachelor's degree, experience in cloud infrastructure, and proficiency in software reliability solutions. The role offers a competitive salary between $350,000 – $475,000 USD along with comprehensive benefits.
#J-18808-LjbffrSRE for Distributed AI Training Infra | Unlimited PTO in san francisco at Unknown Company
This position is listed as full time and onsite.