NVIDIA is seeking a talented Machine Learning Engineer to drive end-to-end lifecycle management of AI-powered systems across distributed infrastructure. You will deploy and scale models, manage GPU orchestration, and build automated testing and CI/CD pipelines using GitLab.
The role demands strong Python engineering skills, experience with Kubernetes, Ray, Slurm, and ML frameworks, and a track record of delivering production-grade AI workflows at scale in a fast-paced environment.
#J-18808-LjbffrML Systems Engineer - Distributed AI & GPU in santa clara at Unknown Company
This position is listed as full time and onsite.