Description
SpaceX is hiring a Site Reliability Engineer focused on high-performance computing to operate and improve its shared compute platform used for vehicle and structures simulation, machine learning, AI inference, and other engineering workloads. The role owns Linux infrastructure, infrastructure as code, storage, user-facing applications, observability, automation, capacity planning, and sustainable incident response, while collaborating with HPC systems engineers and other disciplines. It is based in Hawthorne, California, primarily on-site, requires production infrastructure experience, and offers a base salary of $125,000 to $195,000 depending on level, along with medical, vision, dental, retirement, and other benefits.
