Description
Reflection is hiring a GPU Infrastructure Engineer to design, build, and operate large-scale GPU infrastructure for high-throughput model inference, mid-training, synthetic data generation, and reinforcement learning workloads. The role focuses on optimizing throughput, latency, and GPU utilization; developing distributed systems and inference platforms; improving kernel-level and runtime performance; and diagnosing bottlenecks across GPUs, networking, and distributed execution. Candidates should have experience with production GPU systems, modern LLM inference frameworks, distributed reinforcement learning infrastructure, and low-level performance optimization.
