Cohere is seeking a Lead Member of Technical Staff to drive the architecture and deployment of scalable NLP model-serving platforms. You will own the end-to-end design of low-latency, high-throughput API endpoints and guide cross-functional teams to deliver production-grade infrastructure.
You will mentor engineers, set technical direction across multiple teams, and optimize compute/storage networks for cost efficiency while supporting GPU-accelerated workloads in hybrid multi-cloud environments.
#J-18808-LjbffrLead ML Infrastructure Engineer (Kubernetes + GPUs) in california at Unknown Company
This position is listed as full time and hybrid.