Description
NVIDIA is hiring a Deep Learning Engineer to integrate advanced communication technologies into AI frameworks and toolkits, including PyTorch, vLLM, SGLang, TRT-LLM, and veRL. The role involves analyzing multi-GPU workloads, developing communication and fused compute-communication kernels, researching GPU performance, and building fault-tolerant and elastic solutions for large-scale training and inference. Candidates should be pursuing a master’s or doctoral degree in computer engineering, computer science, or electrical engineering, with strong communication, kernel, and AI training or inference experience, and proficiency in Python, C++, CUDA, or related DSLs.
