Description
NVIDIA is hiring a Senior AI/ML Performance and Efficiency Engineer, GPU Clusters to improve the efficiency, productivity, and cost-effectiveness of AI/ML research workloads. The role involves collaborating with researchers and engineering teams to identify infrastructure and application bottlenecks, build performance tools and frameworks, analyze GPU-cluster utilization, and deliver scalable solutions. Candidates need a computer science or related background, at least five years of experience designing and operating large-scale compute infrastructure, expertise in ML performance, distributed training, NVIDIA debugging tools, and programming languages such as Python, Go, and Bash. The posting lists base salaries of 152,000–241,500 USD for Level 3 and 184,000–287,500 USD for Level 4, along with equity and benefits, and states that applications were accepted until March 23, 2026.
