Skip to main content

Distributed Training Engineer at Reflection

Department: Research

Setup
On-site
Location
San Francisco, California
Level
senior
Posted

Description

Reflection is hiring a distributed-training engineer to build and scale systems for frontier-model pre-training. The role involves designing and operating large-scale training runs, developing infrastructure across thousands of GPUs, optimizing throughput and GPU utilization, building training pipelines, and debugging distributed-training bottlenecks. Candidates should have experience with large-scale distributed training frameworks, model parallelism, GPU communication libraries, and foundation-model training workflows.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation