Skip to main content

Senior Deep Learning Research Engineer, LLM Inference at NVIDIA

Department: Architect

Language
Setup
On-site
Location
Tel Aviv, Tel-Aviv District
Type
Full-time
Level
senior
Posted

Description

NVIDIA is seeking a Deep Learning Research Engineer to develop and improve LLM inference algorithms, benchmarks, profiling workflows, and experimental frameworks. The role involves prototyping low-latency and high-throughput inference algorithms, profiling performance on NVIDIA hardware, identifying optimization opportunities, and translating research into production software. Candidates need an MSc or equivalent industrial research experience, at least five years of applied research, research engineering, or algorithm engineering experience, strong Python and PyTorch skills, and experience with large-scale GPU clusters.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation