Service Engineer – High Performance Computing (HPC)
Location: Redmond, WA (Hybrid) Job Type: Full Time
We are seeking an experienced Service Engineer (7–10 years) with strong expertise in HPC infrastructure management and system administration. The role involves supporting distributed computing platforms, integrating third-party applications, and ensuring smooth HPC operations.
Key Responsibilities
- Manage and optimize HPC and PLM infrastructure.
- Configure HPC schedulers (Windows HPC Pack, OpenPBS, Slurm).
- Support distributed workloads (MPI-based apps) and networking (TCP/IP, RDMA).
- Implement Azure CycleCloud and Azure Batch for HPC orchestration.
- Maintain Windows/Linux environments, VMs, and VMSS.
- Collaborate with teams to ensure performance and compliance.
Required Skills
- Development: 3+ years in PowerShell, Azure Bash, Go, or Python.
- Administration: 7+ years in system administration (PLM/HPC).
- Strong communication and backlog management skills.
Preferred Attributes
- Passion for HPC technologies.
- Quick learner with strong problem-solving skills.