Skip to main content

Senior ML Engineer (Token Factory) at Nebius

Department: ML

Language
Location
Amsterdam, North Holland
Level
not_specified
Posted

Description

Nebius is hiring a GPU Inference Software Engineer to develop and optimize low-level kernels and runtime components for AI inference, improve GPU platform performance, profile and debug system- and hardware-level issues, and integrate support for new GPU architectures. The role requires strong C++ or GPU programming expertise, systems-level software experience, profiling and debugging skills, and knowledge of CPU/GPU architecture and memory hierarchy. Preferred qualifications include CUDA, ROCm, CUTLASS, Triton, TensorRT, and other inference-engine technologies.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation