Description
NVIDIA is hiring an AI Server Performance Architect to drive server-level performance across CPU, GPU, memory, interconnect, networking, and storage subsystems. The role involves workload characterization, bottleneck analysis, profiling, trade-off studies, performance-gap closure, automation, architecture reviews, and analytical modeling for next-generation AI server platforms. Candidates need a BS, MS, or PhD in a relevant engineering or computer science field, at least 10 years of server or system performance architecture experience, and strong knowledge of modern server architectures, GPU-accelerated compute, and performance analysis tools.
