Description
The Site Reliability Engineer will architect resilient, automated production environments for Google Cloud services and lead the team through scaling, capacity, compliance, and accelerated product-development challenges. Responsibilities span service design, deployment, monitoring, incident response, and postmortems, while the role requires a bachelor's degree or equivalent practical experience, 8 years of software development experience, 3 years leading projects, and 3 years designing distributed systems. Preferred qualifications include a master's degree, engineering-mentoring experience, deep AI infrastructure expertise, and cross-functional technical strategy influence.
