Senior HPC AI Cluster Engineer
Design and maintain large-scale HPC/AI clusters with a focus on GPU-accelerated computing and deep learning platforms. Develop automation tooling for deployment, monitoring, and self-service infrastructure, while collaborating with researchers and engineers to optimise performance at scale. Work across bare metal, OS, networking, and application layers to troubleshoot and improve system reliability and efficiency.