Description
Arm is hiring a Principal Engineer for its AI Compute Infra team to design, build, and operate large-scale infrastructure supporting AI training, fine-tuning, evaluation, and inference. The role focuses on Kubernetes clusters, CPU and GPU enablement, workload scheduling, high-performance networking, storage, capacity management, reliability, and automation, while partnering with AI researchers and engineers. Candidates need at least 8 years of production cloud, compute, HPC, or distributed infrastructure experience, systems programming experience, and knowledge of Kubernetes, containers, Linux, networking, and storage. The role offers a salary of $262,700-$355,400 per year and includes visa sponsorship support for candidates who require it.
