HPC Storage Engineer
Location: Dallas, TX (Hybrid)
Type: Direct Hire
Benefits
- Competitive base salary + performance bonus
- 100% company-paid benefits
Overview
We are seeking an HPC Storage Engineer to support the design, deployment, and operation of large-scale, high-performance storage infrastructure supporting mission-critical compute environments.
This role sits within a high-performance computing (HPC) storage engineering team managing over 200PB of scalable storage. The environment supports research-driven workloads, making storage performance, reliability, and optimization a core component of overall infrastructure success.
The ideal candidate brings deep expertise in large-scale storage systems, clustered filesystems, and modern data centre storage architectures, along with a strong collaborative mindset to partner with researchers, software engineers, network teams, and external vendors.
Key Responsibilities
Storage Architecture & Engineering
- Design, implement, and manage multi-petabyte storage systems supporting high-performance compute environments
- Support the architecture and scaling of distributed and clustered storage platforms
- Act as a subject matter expert across storage technologies for infrastructure and engineering initiatives
Automation & Infrastructure as Code
- Develop and implement automated processes aligned with Infrastructure-as-Code principles
- Build consistent, repeatable deployment and operational workflows for storage platforms
- Support DevOps-driven approaches to infrastructure management and lifecycle operations
Performance Optimization & Troubleshooting
- Optimize storage performance to support demanding HPC and research workloads
- Troubleshoot complex storage, workflow, and performance issues across compute environments
- Partner with engineering and research teams to identify and resolve bottlenecks
Cross-Functional Collaboration
- Work closely with researchers and software engineers to streamline workflows and improve storage utilization
- Collaborate with network teams and infrastructure partners to ensure end-to-end system performance
- Interface with external vendors including Dell, VAST Data, HPE, and Rubrik
Required Experience
- Bachelor’s degree or equivalent practical experience
- Experience deploying and managing high-performance or clustered filesystems at scale
- Strong understanding of data centre storage technologies including file, block, and object storage
- Experience with scale-out storage architectures such as VAST Data and Isilon/PowerScale
- Hands‑on experience operating in large-scale, high-availability environments
Technical Skills
- Python and/or DevOps experience supporting infrastructure automation
- Experience working within ticketing systems such as JIRA
- Strong problem-solving skills with the ability to troubleshoot complex infrastructure issues
- Passion for storage technologies, hardware, and performance optimization
Preferred Experience
- Experience supporting HPC or research-driven environments
- Exposure to large-scale (multi-petabyte) storage platforms
- Experience working in highly collaborative, cross-functional engineering environments