Description
NVIDIA is hiring a Local AI Software Engineer to build and maintain the software stack that enables large language models and generative AI applications to run efficiently on NVIDIA edge AI hardware. The role focuses on evaluating open-source LLM inference frameworks, analyzing model architectures and inference algorithms against GPU hardware, characterizing multi-node inference behavior, producing performance analysis reports, owning model validation workflows, developing developer-facing inference recipes, and supporting community and partner model bring-up activities. The position requires a computer science, computer engineering, electrical engineering, or equivalent background, 12+ years of software engineering experience, strong Python or C++ skills, GPU kernel development expertise, LLM inference knowledge, and container engineering expertise. The base salary range is 224,000 USD to 356,500 USD for Level 5 and 272,000 USD to 431,250 USD for Level 6, with equity and benefits available.
