Skip to main content

AI Computing Software Development Engineer, TensorRT-LLM at NVIDIA

Department: Architect

Language
Setup
On-site
Location
Taipei · Hsinchu
Type
Full-time
Level
mid
Posted

Description

NVIDIA is hiring a Software Development Engineer for LLM inference to develop and optimize scalable inference software for TensorRT-LLM. The role involves performance analysis, kernel and runtime implementation, architecture and hardware feedback, and collaboration with software, research, and product teams. Candidates need a master's degree or equivalent experience, at least three years of relevant software development experience, strong Python and C/C++ skills, and experience with deep learning frameworks, LLM inference, and GPU programming.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation