Description
AWS is hiring a Kernel Engineer to design and implement high-performance compute kernels for machine learning operations on the Neuron architecture, optimize kernel-level performance across Inferentia and Trainium hardware, analyze bottlenecks with profiling tools, and apply compiler optimizations such as fusion, sharding, tiling, and scheduling. The role also involves customer-facing model enablement and optimization, cross-team collaboration, research publication, and mentoring experienced engineers. Candidates need substantial software development, system architecture, full software development lifecycle, and engineering leadership experience, along with accelerator and GPU optimization expertise.
