Description
Hewlett Packard Enterprise is hiring an onsite Performance Engineering Architect to characterize and optimize the performance of generative AI, machine learning, and other AI workloads across GPUs, servers, storage, networking, system software, and AI frameworks. The role involves benchmarking, workload analysis, automation, bottleneck identification, performance baselines, sizing evidence, and collaboration with engineering, product, field, and partner teams. Candidates need strong AI and GPU performance knowledge, Linux and Python skills, experience with inference frameworks and server hardware, and the ability to document reproducible performance methodologies.
