Description
The company is hiring Data Center Infrastructure Software Engineers at Senior, Staff, and Principal levels to design, develop, configure, and automate AI infrastructure clusters. Responsibilities include infrastructure-as-code and provisioning systems, Kubernetes and container deployment, GPU and distributed computing optimization, troubleshooting, and reliability and observability practices. Candidates need at least five years of experience with large-scale Linux infrastructure, production Kubernetes and distributed systems, and infrastructure automation; bare-metal, IPMI, GPU, and automated OS deployment experience is preferred. The role is hybrid in Bellevue, Washington, with approximately three days per week in the office, and requires U.S. work authorization.
