说明
Nebius 正在寻找一位 GPU 集群架构师,负责在计算、网络、存储和控制平面等方面设计并优化大规模 AI 基础设施。该岗位涉及架构 GPU 集群拓扑、对 AI/ML 工作负载性能进行建模、验证低延迟互连、集成存储、分析监控信号,并与可靠性、网络、存储以及数据中心工程团队协作。候选人需要至少五年集群设计经验,具备现代 GPU 架构知识,具备 InfiniBand 和 RoCE 的经验,拥有系统架构与硬件可靠性方面的专业知识,并具备使用 Python 或 Go 进行脚本编写的经验。
AI cloudAmsterdam
Nebius is an AI cloud company delivering a unified platform across the AI journey, from data and model training and tuning to production runtime and deployment. Its purpose-built AI cloud is built for AI developers and serves AI builders and enterprises worldwide across industries including healthcare and life sciences, robotics and physical AI, financial services, media and entertainment, and retail. Nebius offers flexible consumption options, serverless inference, and dedicated capacity; the company is listed on Nasdaq (NBIS) and headquartered in Amsterdam.
客户与案例
查看客户与案例