Unknown Company

Performance Engineer: GPU Kernel & Inference Optimize

san francisco, ca • Posted 3 days ago
Onsite Full Time Engineering

World Labs in San Francisco is hiring a Performance Engineer to accelerate training and serving of large world models. You will identify bottlenecks across kernels, serving paths, and GPUs, then implement concrete improvements that raise throughput and lower latency.

You’ll work end-to-end, from CUDA kernels to fleet-wide serving, partner with researchers, and own numerical correctness across precision and hardware changes.

#J-18808-Ljbffr
Back to Job Search