Description
Lambda is hiring a Kubernetes Site Reliability Engineer to operate and maintain bare-metal Kubernetes clusters, handle incidents and cluster lifecycle changes, build control-plane services and automation, and support customers with Kubernetes, storage, authentication, and workload integration. The role requires 6+ years of SRE or operations experience, strong Go and Python skills, production Kubernetes expertise, and familiarity with observability and CI/CD tools. It is based in San Francisco, San Jose, or Bellevue with four days per week in the office.
