Unknown Company

Inference AI/ML Engineer for GPU Model Serving

sunnyvale, ca • Posted Today
Onsite Full Time IT & Technology

CoreWeave is seeking an IC1 engineer to join the Inference team and ship production features for model serving on our GPU platform. You will implement well-scoped changes, learn our practices, and grow quickly with mentorship from experienced engineers.

You will work on Python/Go/C++ services such as Triton, vLLM, and TensorRT-LLM, write tests and docs, and contribute to metrics, dashboards, and runbooks. This is an opportunity to advance in a fast-paced AI infrastructure company in Sunnyvale,

#J-18808-Ljbffr

Inference AI/ML Engineer for GPU Model Serving in sunnyvale at Unknown Company

This position is listed as full time and onsite.

Back to Job Search