CoreWeave is seeking an IC1 engineer to join the Inference team and ship production features for model serving on our GPU platform. You will implement well-scoped changes, learn our practices, and grow quickly with mentorship from experienced engineers.
You will work on Python/Go/C++ services such as Triton, vLLM, and TensorRT-LLM, write tests and docs, and contribute to metrics, dashboards, and runbooks. This is an opportunity to advance in a fast-paced AI infrastructure company in Sunnyvale,
#J-18808-LjbffrInference AI/ML Engineer for GPU Model Serving in sunnyvale at Unknown Company
This position is listed as full time and onsite.