Description
Plaud is hiring an SRE or Platform Engineer to ensure the reliability and performance of its AI products at scale. The role owns production reliability, incident response, on-call practices, observability, reliability automation, SLO/SLI and error-budget management, postmortems, and continuous reliability improvement while partnering with product and engineering teams. Candidates need at least five years of SRE, infrastructure, or platform engineering experience, cloud-platform expertise, Kubernetes and distributed-systems knowledge, on-call and incident-management experience, and proficiency in Go, Python, or Java.

