ScitiX seeks a Kubernetes-focused systems engineer in San Francisco to own deployment, daily operation, and maintenance of training clusters. You will diagnose faults, optimize performance, and ensure reliable AI workflows across multi-environment infrastructures.
You will design monitoring and automation for the cluster management platform, troubleshoot containers, OS, networking, and storage issues, and contribute to capacity planning and resource governance.
#J-18808-LjbffrAI Training Kubernetes SRE — Cloud-Native Platform Ops in san francisco at Unknown Company
This position is listed as full time and onsite.