Skip to main content

Senior Inference Engineer, GPU Kernel Optimization at NVIDIA

Department: Engineer, Sys SW

Compensation

$184,000 – $287,500/yr

Setup
On-site
Location
California · Austin, Texas · +2
Type
Full-time
Level
senior
Posted

Description

NVIDIA is hiring a Senior Inference Engineer focused on GPU kernel optimization for LLM inference. The role develops silicon-measured kernel benchmarking infrastructure, model-level performance analysis, and agentic kernel-optimization systems, then collaborates with compiler, hardware, kernel, and framework teams to improve production inference performance. Candidates need a master's or PhD in a relevant field or equivalent experience, at least six years of industry experience, expertise in Python and C++, GPU profiling tools, LLM inference frameworks, and GPU kernel optimization techniques.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation