Unknown Company

Senior ML Engineer: Quantized Inference & Pipelines

redmond, wa • Posted 1 weeks ago
Onsite Full Time Software Architecture & Engineering

NVIDIA AI is seeking an engineer to implement quantized and sparse recipes in inference engines and to manage model export pipelines for correct serialization. You will build benchmarking harnesses and data analysis tools to improve developer productivity through infrastructure and CI improvements.

The role requires strong Python and C++ skills, experience with ML accelerators, and familiarity with PyTorch internals, with 4+ years in software engineering. MS/PhD in CS is preferred.

#J-18808-Ljbffr

Senior ML Engineer: Quantized Inference & Pipelines in redmond at Unknown Company

This position is listed as full time and onsite.

Back to Job Search