Skip to main content

GPU Cluster Architect at Nebius

Department: Product & Infrastructure

Language
Location
Singapore · India
Level
senior
Posted

Description

Nebius is seeking a GPU Cluster Architect to design and optimize large-scale AI infrastructure across compute, networking, storage, and control planes. The role involves architecting GPU cluster topologies, modeling AI/ML workload performance, validating low-latency interconnects, integrating storage, analyzing monitoring signals, and collaborating with reliability, networking, storage, and data center engineering teams. Candidates need at least five years of cluster-design experience, knowledge of modern GPU architectures, experience with InfiniBand and RoCE, systems architecture and hardware reliability expertise, and scripting experience with Python or Go.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation