Description
Nebius is hiring a Senior Machine Learning Engineer to own the infrastructure and experiments for large-scale AI model training and post-training, including SFT, continued pretraining, preference optimization, and RL methods. The role involves building reward functions, synthetic data pipelines, distributed training and RL infrastructure, parallelism strategies, GPU performance optimizations, evaluations, and experiment documentation, while collaborating with research and platform teams. Candidates need strong Python and PyTorch skills, hands-on experience across at least two relevant areas, and practical knowledge of LLM training, transformer bottlenecks, and distributed systems.
