HPC Engineer – IFM
MBZUAI
Anywhere
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: []Join the IFM infrastructure team to support and maintain large-scale GPU computing clusters powering frontier AI research. You’ll assist researchers with job submission and troubleshooting, monitor cluster health and performance, and handle issues across Linux, hardware, storage, networking, and software. The role includes Slurm administration, cluster deployment and upgrades, building automation scripts, documenting operations, and collaborating with researchers and vendors.

