Description
NVIDIA is hiring a Software Engineer for its AI networking acceleration team to develop a highly optimized, open-source inference framework. The role involves building low-level infrastructure that runs on large supercomputers and data centers, using hardware offloads, GPU kernels, and RDMA network cards. Candidates need a B.Sc. or equivalent in Computer Science or Software Engineering, 8 years of modern C++/C/Rust development experience, 3 years of Linux experience, and deep TCP/IP knowledge, with LLM inference, distributed storage, Linux internals, CUDA, parallel programming, and high-performance computing experience preferred.
