To support the expansion of AI capabilities, the full-time Lead AI Infrastructure Operations Engineer will enable, operate, and continuously improve infrastructure for production AI solutions while working remotely or in a hybrid environment. Key responsibilities Partner with various teams to translate AI solution designs into infrastructure requirements and coordinate provisioning with managed services Implement and maintain observability and monitoring for production AI applications, while diagnosing and resolving production issues Support AI governance by operationalizing controls, monitoring performance, and analyzing trends to identify potential risks and issues Required qualifications Bachelor's degree in computer science, engineering, information technology, or a related field, or equivalent practical experience 8+ years of information technology experience, including 5+ years in infrastructure operations or related fields Experience with cloud-hosted, data-intensive, or AI-enabled production applications Familiarity with AWS services and capabilities related to monitoring, logging, and infrastructure operations Experience supporting observability, logging, and operational reporting for production applications
AI Infrastructure Operations Engineer in workfromhome at Unknown Company
This position is listed as full time and able to be worked remotely.