Description
Google is hiring a Site Reliability Engineering Manager to build and scale critical AI infrastructure and manage a North American SRE team. The role partners with development teams and the Sydney counterpart to run a follow-the-sun rotation, guides resilient system architecture, leads incident response and operational automation, and develops technical roadmaps. It requires a bachelor’s degree or equivalent practical experience, 8 years of software development or data structures/algorithms experience, 3 years of large-scale distributed systems experience, and 3 years of engineering-team management experience. The position offers USD 207,000–300,000 plus a 20% bonus target, equity, and benefits.
