Parallax Systems is building the next generation of inference infrastructure for enterprise AI workloads. We need an AI Infrastructure Lead to own the design and operation of our GPU cluster management layer, model serving pipeline, and low-latency routing system.
You will work directly with the CTO and a team of four senior engineers. This is a high-ownership role in a fast-moving environment — you will make architectural decisions that affect thousands of enterprise clients.
Core Requirements
- Proven experience designing and operating large-scale GPU infrastructure and model serving systems.
- Deep knowledge of distributed training, inference optimisation, and containerised workloads.
- Hands‑on expertise with AWS, GCP, or Azure AI/ML services and Kubernetes.
- Strong background in monitoring, alerting, and incident response for critical AI systems.
Rolling contract with 6‑month minimum commitment.
Work load 40 HRS/WK
Full‑time availability required. On‑call rotation included.
#J-18808-Ljbffr