Skip to main content

Senior Site Reliability Engineer (SRE, Compute Node Team) at Nebius

Department: SRE

Language
Setup
Remote
Location
Amsterdam, North Holland
Level
senior
Posted

Description

Nebius is hiring a Senior Site Reliability Engineer (SRE) for its Compute Node team to build and operate the cluster scheduler and node-level services that manage virtual machines across cloud regions. The role focuses on Linux systems engineering, virtualization, containerization, production troubleshooting, observability, incident response, root-cause analysis, and reliability improvements. It requires strong Linux, QEMU/KVM, containerization, debugging, and SRE experience, with optional expertise in Kubernetes, low-level Linux tools, large-scale compute platforms, open-source infrastructure, and hardware or GPU debugging.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation