Skip to main content

Sr. Site Reliability Engineer, AI Infrastructure (Starshield) at SpaceX

Department: Starshield Software EngineeringEducation: education_required

Compensation · Listed in posting

$165,000 – $265,000/yr

Location
Hawthorne, California
Type
Full-time
Level
senior
Posted

Description

SpaceX is hiring a Senior Site Reliability Engineer focused on Starshield’s software and GPU infrastructure. The role manages GPU and CPU deployments in Top Secret data centers, supports GPU-as-a-service, designs and productizes AI clusters at 100k+ GPU scale, automates on-premise Kubernetes and AI clusters, manages databases and distributed storage, and collaborates with AI engineers. The position requires a relevant bachelor’s degree and 5+ years of Linux experience, or 7+ years of software, DevOps, or site reliability experience; 5+ years of Kubernetes experience; and experience with Terraform, Ansible, containerization, scripting, and Python, C++, or Go. The role requires a Top Secret security clearance, extended hours, domestic and global travel, and ITAR eligibility. Compensation is $165,000–$265,000 for Level 3, with medical, vision, dental, retirement, and other benefits.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation