Skip to main content

GPU Infrastructure Engineer at Reflection

Department: Research

Setup
On-site
Location
San Francisco, California
Level
senior
Posted

Description

Reflection is hiring a GPU Infrastructure Engineer to design, build, and operate large-scale GPU infrastructure for high-throughput model inference, mid-training, synthetic data generation, and reinforcement learning workloads. The role focuses on optimizing throughput, latency, and GPU utilization; developing distributed systems and inference platforms; improving kernel-level and runtime performance; and diagnosing bottlenecks across GPUs, networking, and distributed execution. Candidates should have experience with production GPU systems, modern LLM inference frameworks, distributed reinforcement learning infrastructure, and low-level performance optimization.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation