Description
NVIDIA is hiring a Senior HPC Cluster Administrator to lead the design, deployment, provisioning, monitoring, automation, and reliability of large-scale GPU compute clusters supporting deep learning training, inference, and high-performance computing workloads. The role covers heterogeneous Linux environments, storage and networking architectures, Slurm scheduling, observability, distributed training optimization, and infrastructure automation, while mentoring junior engineers. Candidates need a BS/MS in a relevant field or equivalent hands-on experience, at least five years of large-scale HPC or ML training cluster experience, and expertise in Linux, Python or bash, Ansible, Terraform, container technologies, high-speed networking, and distributed filesystems. The base salary for Poland is 221,250 PLN to 383,500 PLN for Level 3 and 292,500 PLN to 507,000 PLN for Level 4.
