Cohere is seeking a Lead Member of Technical Staff to drive the Model Serving platform, delivering low-latency NLP model deployments via scalable, reliable APIs. You will mentor engineers, define architecture direction, and coordinate across teams to meet customers’ needs.
The role focuses on leading design and deployment of high-performance, distributed ML infrastructure with strong emphasis on Kubernetes, cloud platforms, and accelerator technologies.
#J-18808-LjbffrLead ML Infra Architect for High-Performance NLP in new york at Unknown Company
This position is listed as full time and onsite.