Manager, HPC Storage Engineer
RunPod
United States
Workplace: RemoteFull timeUSD 150,000 - 240,000 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 3+ yearsSkills: ["People management","Operational discipline","Technical leadership","Communication","Incident response"]Lead the team responsible for distributed, GPU-centric storage infrastructure across all regions. Own the end-to-end storage stack—from NAND/NVMe media and controllers through SAN/NFS architectures, Lustre-style parallel filesystems, and cluster deployments—driving performance, reliability, and scalability for AI training, inference, checkpointing, and dataset access. Manage engineers, build observability and automation, and partner with networking, SRE, and product to deliver next-generation capabilities like NFS over RDMA and GPU Direct Storage.

