Lead the design and implementation of scalable GPU infrastructure and high-performance AI model serving APIs. Architect robust distributed systems to optimize AI workloads and mentor engineers to set the technical vision for AI infrastructure.
Requirements
Requires over 5 years of experience in systems infrastructure with deep expertise in Kubernetes, cloud platforms, and AI serving frameworks. Proficiency in Python, Go, or C++ and a proven track record of managing large-scale GPU fleets are essential.
Key Skills
GPU Infrastructure, AI Model Serving, Distributed Systems, Kubernetes, Cloud-native Infrastructure, API Optimization, Python, Go, C++, GPU Orchestration, TensorFlow Serving, Triton, TorchServe, Model Deployment, MLOps, System Architecture
Benefits
Equity, Comprehensive health benefits, Monthly stipends, Company retreats
#J-18808-LjbffrSoftware Engineer, AI Infra in palo alto at Unknown Company
This position is listed as full time and onsite.