Description
The Site Reliability Engineer will improve the reliability, scalability, and efficiency of large-scale distributed systems supporting Google Cloud services. Responsibilities span service design, software platform development, capacity planning, launch reviews, monitoring, automation, incident response, and blameless postmortems. The role requires a bachelor’s degree or equivalent practical experience, 8 years of software development experience, 3 years leading projects, and 3 years designing, analyzing, and troubleshooting distributed systems.
