Senior HPC AI Cluster Engineer
Design and maintain large-scale HPC/AI clusters with a focus on system automation, monitoring, and performance tuning across bare metal, OS, and application layers. Develop CI/CD pipelines and self-service tooling for infrastructure management, supporting cutting-edge AI and GPU computing workflows. Collaborate with researchers and engineers to deploy and optimize accelerated computing platforms at scale.