SpaceX in Palo Alto is seeking a Software Engineer, Inference (AI Data Engineering) to design and optimize large-scale model serving systems. You will own distributed infrastructure, drive high-throughput, low-latency inference for SpaceX's critical workloads.
You will work on GPU kernels, quantization, and acceleration techniques, collaborating with diverse teams to deliver reliable, scalable AI infrastructure that powers our ambitious missions. On-site work in Palo Alto is required.
#J-18808-LjbffrAI Inference Systems Engineer (High-Throughput, Low-Latency) in palo alto at Unknown Company
This position is listed as full time and onsite.