Description
The role focuses on developing and optimizing CUDA programming techniques for deep learning, graphics, machine learning, and data analytics across current and next-generation GPU architectures. It involves analyzing application performance, collaborating with NVIDIA architecture, research, libraries, tools, and system software teams, and working with customers to provide GPU-based AI solutions. The position requires a relevant engineering or computer science degree or equivalent experience, at least three years of experience, C/C++ and/or Python proficiency, strong mathematical fundamentals, parallel programming experience, and the ability to travel for conferences and on-site visits.
