Lead the day-to-day operation, support, and continuous improvement of enterprise infrastructure services — spanning cloud-hosted server environments, endpoint management, identity services, and enterprise application infrastructure
Manage and maintain cloud-hosted servers running enterprise directory services, database platforms, and business applications — owning the operating system layer while cloud engineering owns resource deployment and connectivity
Lead patch management, software installation, configuration management, permissions management, and services troubleshooting across all server and endpoint environments
Ensure reliable operation and patch compliance of server operating systems through enterprise patch management tooling and oversee endpoint management platforms to maintain a consistent and secure end-user computing environment
Lead vulnerability management prioritization and standard remediation workflows for infrastructure assets, operating under segregation of duty from application teams while ensuring proper cross-functional coordination and system availability
Oversee identity and directory services operations including service account management, group lifecycle, certificate-based authentication, and related operational processes in partnership with Cybersecurity
Monitor infrastructure health, capacity, and performance across all environments — proactively identifying risks, capacity constraints, and improvement opportunities before they become incidents
Coordinate with cloud engineering to ensure resource deployments, connectivity configurations, and infrastructure changes are aligned to operational requirements and do not disrupt production services
Own routing and network-level operations across enterprise firewall, remote access, SD-WAN, switching, wireless, and network access control platforms — coordinating with managed services partners for operational support and incident response
Partner with Cybersecurity, for access control policies, firewall rule management, policy changes, and platform upgrades — providing technical expertise and operational insights to support seamless cross-team execution
Serve as the primary internal technical counterpart to managed services partners for network operations — driving monitoring, escalation management, and continuous service improvement
Serve as the highest-level technical escalation point for complex infrastructure incidents, outages, and critical changes — coordinating resolution across internal teams, managed services partners, vendors, and offshore resources
Establish and enforce domain engineering standards: configuration baselines, change management guardrails, operational documentation, and infrastructure-as-code patterns across systems and network disciplines
Coordinate incident response, problem management, and root cause analysis activities — driving recurring issue reduction and operational maturity across the team
Lead disaster recovery and business continuity planning for infrastructure services — ensuring DR runbooks are current, tested, and aligned to recovery objectives
Partner with enterprise governance bodies and Cybersecurity to ensure infrastructure changes are governance-approved, compliant, and aligned to enterprise architecture
Manage service provider and vendor relationships — tracking licensing, support contracts, renewals, and professional services engagements
Lead and develop onshore and offshore engineering resources through coaching, mentoring, performance management, and technical guidance — building independence, accountability, and technical maturity across the team
Requirements
Bachelor's degree in Information Technology, Computer Science, Engineering, or equivalent practical experience
10+ years of progressive experience in enterprise infrastructure engineering spanning systems administration, cloud infrastructure, and network operations
5+ years of experience leading technical teams, projects, or operational functions with direct people management responsibility, including experience managing offshore or distributed resources
Strong experience with enterprise server operating system administration, directory services, group policy management, and enterprise patch management platforms
Demonstrated experience managing server workloads in cloud-hosted IaaS environments, including OS-level management, patching, software installation, and configuration
Experience with enterprise identity and directory services, including service account management, group lifecycle, and certificate-based authentication
Experience with enterprise monitoring and observability tools for infrastructure health, capacity planning, and performance management
Strong hands-on experience with enterprise-grade switching, wireless, firewall, remote access, and SD-WAN platforms
Strong knowledge of routing protocols, VLANs, DHCP, DNS, and network access control solutions
Experience managing managed services provider relationships for network monitoring, incident escalation, and operational support coordination
Experience serving as a senior escalation resource for complex, multi-site infrastructure environments
Proven ability to independently research, diagnose, and resolve complex technical issues — including in unfamiliar technologies and without existing documentation
Strong knowledge of change management, operational governance, and risk-aware engineering practices.