Description
The Workload Orchestration Engineer will lead SLURM architecture, scheduling, resource optimization, and observability across Roche’s heterogeneous HPC and AI environments. The role bridges traditional scientific computing with Kubernetes and cloud-native workflows, solves complex multi-tenant and GPU allocation issues, establishes technical standards and governance, and mentors engineering colleagues. It requires extensive systems engineering and SLURM expertise, experience with HPC, life sciences or pharmaceutical research environments, Kubernetes and container knowledge, GPU and interconnect familiarity, and automation skills.
