World Labs in San Francisco is hiring a Performance Engineer to accelerate training and serving of large world models. You will identify bottlenecks across kernels, serving paths, and GPUs, then implement concrete improvements that raise throughput and lower latency.
You’ll work end-to-end, from CUDA kernels to fleet-wide serving, partner with researchers, and own numerical correctness across precision and hardware changes.
#J-18808-Ljbffr