Together AI is building the best inference infrastructure for voice applications. We seek a Staff ML Engineer to own the model serving stack and optimize latency and throughput for real-time voice workloads.
You'll work with state-of-the-art accelerators and collaborate with model partners to bring models to production on Together's platform. This is a foundational role on a small, high-impact team.
#J-18808-LjbffrStaff ML Engineer — Real-Time Voice Inference Architect in san francisco at Unknown Company
This position is listed as full time and onsite.