Acceler8 Talent in San Francisco seeks a Member of Technical Staff focusing on Kernels & GPU Performance to push the limits of production AI inference across diverse accelerators.
You will implement low-level kernels, analyze memory hierarchies, and collaborate with compiler, ML systems, and runtime teams to optimize latency and throughput.
This on-site role offers a full-time path at a fast-growing AI infrastructure company.
#J-18808-LjbffrKernel Engineer: GPU Performance & Inference in san francisco at Unknown Company
This position is listed as full time and onsite.