Skip to main content

Site Reliability Engineer, AI Infrastructure (Starshield) at SpaceX

Department: Starshield Software EngineeringEducation: education_required

Compensation · Listed in posting

$125,000 – $160,000/yr

Location
Washington, District of Columbia
Type
Full-time
Level
mid
Posted

Description

SpaceX is hiring a Site Reliability Engineer focused on Starshield’s software and GPU infrastructure. The role manages GPU and CPU deployments in Top Secret datacenters, supports GPU-as-a-service, designs and productizes AI clusters at 100k+ GPU scale, automates on-premise Kubernetes and AI clusters, and maintains databases, monitoring, and distributed storage. The position requires a relevant bachelor’s degree or equivalent experience, Linux, infrastructure automation, containerization, scripting, and development experience, along with a Top Secret security clearance and willingness to work extended hours, weekends, and travel.

For job seekers

Ready to find a role that actually fits?

Upload your résumé, start a Job Search Thread, and let Metaintro rank real openings against your experience — then guide you from search to offer.

Match

Compare live roles against your current evidence.

Position

Turn proof projects into role-specific applications.

Improve

Use market feedback to keep the skill plan current.

Return to navigation