Thomson Reuters is seeking a Senior Inference Engineer, AI to collaborate with platform teams to productionize and scale AI and LLM workloads across AWS, Azure, GCP and internal Kubernetes clusters.
You will optimize models for low latency, develop containerized inference pipelines, implement routing and observability, and work with data science and product teams to onboard new research models into production.
#J-18808-LjbffrSenior AI Inference Engineer - Hybrid Scale for LLMs in eagan at Unknown Company
This position is listed as full time and onsite.