Unknown Company

Founding ML Inference Performance Engineer

san francisco, ca • Posted 5 days ago
Onsite Full Time IT & Technology

uRun, located in San Francisco, is seeking a founding ML Performance Engineer to drive AI infrastructure performance. In this role, you will write custom CUDA kernels and optimize model inference for real-time applications, significantly impacting performance across the stack.

The ideal candidate will possess deep knowledge of CUDA, experience with AI workloads, and a strong capacity for optimization. The position offers a competitive salary, equity, and top-tier tools for an exceptional contributor.

#J-18808-Ljbffr
Back to Job Search