Description
Micron Technology is hiring a GPU Performance Engineer to architect and optimize large-scale GPU-based machine learning and AI workloads, including model training, fine-tuning, inference, and autonomous AI agents. The role focuses on distributed training, GPU resource management, workload profiling, high-performance kernel development, performance regression testing, and collaboration with hardware architects and data teams. Candidates need a technical degree, extensive GPU and distributed systems experience, strong C++ and CUDA or other GPGPU expertise, and at least five years of experience in performance optimization, parallel computing, or low-level systems programming.
